IP Library Granted Patent US 9,293,140
Granted Patent B2
US 9,293,140 · App. 13/965,661 · Granted Mar 22, 2016

Speaker-identification-assisted speech processing systems and methods

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,293,140
App. No.
13/965,661
Granted
Mar 22, 2016
Kind
B2
Abstract

Methods, systems, and apparatuses are described for performing speaker-identification-assisted speech processing. In accordance with certain embodiments, a communication device includes speaker identification (SID) logic that is configured to identify a user of the communication device and/or the identity of a far-end speaker participating in a voice call with a user of the communication device. Knowledge of the identity of the user and/or far-end speaker is then used to improve the performance of one or more speech processing algorithms implemented on the communication device.

Claims (43)

1. A communication device, comprising:

speaker identification logic configured to apply a speaker identification algorithm to a speech signal to generate speaker identification information, the speaker identification information including at least an identifier that identifies a target speaker associated with the speech signal; and

speech processing logic comprising a plurality of speech signal processing stages, wherein each of the plurality of speech signal processing stages is configured to process the speech signal in accordance with a respective speech processing algorithm based on the speaker identification information provided by the speaker identification logic,

wherein the speaker identification logic is further configured to apply the speaker identification algorithm to the speech signal to generate a first measure of confidence that is indicative of the likelihood that the speech signal is associated with a target speaker;

wherein a first speech signal processing stage of the plurality of speech signal processing stages is configured to process the speech signal in accordance with a first speech processing algorithm in a manner that takes into account the first measure of confidence to produce a processed speech signal; and

wherein the speaker identification logic is further configured to apply the speaker identification algorithm to the processed speech signal to generate a second measure of confidence that is indicative of the likelihood that the processed speech signal is associated with the target speaker.

2. The communication device of claim 1 , wherein a second speech signal processing stage of the plurality of speech signal processing stages is configured to process the processed speech signal in accordance with a second speech processing algorithm in a manner that takes into account the second measure of confidence.

3. The communication device of claim 1 , wherein the speech signal is an uplink speech signal.

4. The communication device of claim 2 , wherein the speech signal is a downlink speech signal.

5. The communication device of claim 3 , wherein the speaker identification logic is configured to apply the speaker identification algorithm to the speech signal to generate the first measure of confidence by;

obtaining a speaker model; and

generating the first measure of confidence by comparing one or more features of at least a portion of the speech signal to one or more features of the speaker model.

6. The communication device of claim 5 , wherein obtaining the speaker model comprises:

obtaining the speaker model from a storage component of the communication device.

7. The communication device of claim 6 , wherein obtaining the speaker model comprises:

obtaining the speaker model based at least on analyzing a portion of the speech signal.

8. A method performed by a communication device, comprising:

making a first speaker identification determination as to an identity of a target speaker based on a speech signal received by the communication device;

processing the speech signal by a first speech processing stage that is configured to process the speech signal in accordance with a first speech processing algorithm based on the first speaker identification determination to produce a first processed speech signal; and

making a second speaker identification determination as to the identity of the target speaker based on the first processed speech signal.

9. The method of claim 8 , further comprising:

processing the first processed speech signal by a second speech processing stage that is configured to process the first processed speech signal in accordance with a second speech processing algorithm based on the second speaker identification determination to produce a second processed speech signal.

10. The method of claim 9 , wherein the speech signal is an uplink speech signal.

11. The method of claim 10 , wherein the speech signal is a downlink speech signal.

12. The method of claim 11 , wherein said making the first speaker identification determination comprises:

obtaining a speaker model; and

comparing one or more features of at least a portion of the speech signal to one or more features of the speaker model to determine a likelihood that the portion of the speech signal is associated with the target speaker.

13. The method of claim 12 , wherein obtaining the speaker model comprises:

obtaining the speaker model from a storage component of the communication device.

14. The method of claim 13 , wherein obtaining the speaker model comprises:

obtaining the speaker model based at least on analyzing a portion of the speech signal.

15. A communication device, comprising:

speaker identification logic configured to make a first speaker identification determination as to an identity of a target speaker based on a speech signal received by the communication device; and

a first speech processing stage configured to process the speech signal in accordance with a first speech processing algorithm based on the first speaker identification determination to produce a first processed speech signal, wherein

the speaker identification logic is further configured to make a second speaker identification determination as to the identity of the target speaker based on the first processed speech signal.

16. The communication device of claim 15 , further comprising a second speech processing stage that is configured to process the first processed speech signal in accordance with a second speech processing algorithm based on the second speaker identification determination to produce a second processed speech signal.

17. The communication device of claim 16 , wherein the speech signal is an uplink speech signal.

18. The communication device of claim 17 , wherein the speech signal is a downlink speech signal.

19. The communication device of claim 18 , wherein the speaker identification logic is further configured to:

obtain a speaker model; and

compare one or more features of at least a portion of the speech signal to one or more features of the speaker model to determine a likelihood that the portion of the speech signal is associated with the target speaker.

20. The communication device of claim 19 , wherein the speaker identification logic is configured to obtain the speaker model by:

obtaining the speaker model from a storage component of the communication device.

Assignments (7)
CORRECTIVE ASSIGNMENT TO CORRECT THE PATENT NUMBER 9,385,856 TO 9,385,756 PREVIOUSLY RECORDED AT REEL: 47349 FRAME: 001. ASSIGNOR(S) HEREBY CONFIRMS THE MERGER. Recorded Mar 22, 2019
From: AVAGO TECHNOLOGIES GENERAL IP (SINGAPORE) PTE. LTD.
To: AVAGO TECHNOLOGIES INTERNATIONAL SALES PTE. LIMITED
Reel/Frame 051144/0648 →
CORRECTIVE ASSIGNMENT TO CORRECT THE EFFECTIVE DATE PREVIOUSLY RECORDED ON REEL 047229 FRAME 0408. ASSIGNOR(S) HEREBY CONFIRMS THE THE EFFECTIVE DATE IS 09/05/2018. Recorded Oct 29, 2018
From: AVAGO TECHNOLOGIES GENERAL IP (SINGAPORE) PTE. LTD.
To: AVAGO TECHNOLOGIES INTERNATIONAL SALES PTE. LIMITED
Reel/Frame 047349/0001 →
MERGER Recorded Oct 4, 2018
From: AVAGO TECHNOLOGIES GENERAL IP (SINGAPORE) PTE. LTD.
To: AVAGO TECHNOLOGIES INTERNATIONAL SALES PTE. LIMITED
Reel/Frame 047229/0408 →
TERMINATION AND RELEASE OF SECURITY INTEREST IN PATENTS Recorded Feb 3, 2017
From: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
To: BROADCOM CORPORATION
Reel/Frame 041712/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 1, 2017
From: BROADCOM CORPORATION
To: AVAGO TECHNOLOGIES GENERAL IP (SINGAPORE) PTE. LTD.
Reel/Frame 041706/0001 →
PATENT SECURITY AGREEMENT Recorded Feb 11, 2016
From: BROADCOM CORPORATION
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 037806/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 13, 2013
From: CHEN, JUIN-HWEY; ZOPF, ROBERT W.; BORGSTROM, BENGT J.; NEMER, ELIAS; PANDEY, ASHUTOSH; THYSSEN, JES
To: BROADCOM CORPORATION
Reel/Frame 031000/0792 →