IP Library Granted Patent US 7,630,895
Granted Patent B2
US 7,630,895 · App. 11/033,697 · Granted Dec 8, 2009

Speaker verification method

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,630,895
App. No.
11/033,697
Granted
Dec 8, 2009
Kind
B2
Abstract

A speaker verification method consist of the following steps: (1) generating a code book ( 42 ) covering a number of speakers having a number of training utterances for each of the speakers; (2) receiving a number of test utterances ( 44 ) from a speaker; (3) comparing ( 46 ) each of the test utterances to each of the training utterances for the speaker to form a number of decisions, one decision for each of the number of test utterances; (4) weighting each of the decisions ( 48 ) to form a number of weighted decisions; and (5) combining ( 50 ) the plurality of weighted decisions to form a verification decision ( 52 ).

Claims (52)

1. A speaker verification method comprising:

performing operations executed by a speaker verification machine, the operations comprising:

comparing a plurality of different test utterances to a plurality of training utterances for a speaker to form a plurality of preliminary verification decisions, one preliminary verification decision for each of the plurality of different test utterances;

weighting each of the plurality of preliminary verification decisions based on a historical error rate corresponding to a respective one of the test utterances; and

combining the weighted preliminary verification decisions to form a verification decision.

2. The method of claim 1 , wherein the one preliminary verification decision is either a true or a false decision.

3. The method of claim 1 , further comprising generating a code book covering a plurality of speakers having the plurality of training utterances for each of the plurality of speakers, wherein a first utterance of the plurality of training utterances is not the same as a second utterance of the plurality of training utterances.

4. The method of claim 1 , wherein weighting each of the plurality of preliminary verification decisions includes determining a historical probability of false alarm for each of the plurality of different test utterances.

5. The method of claim 1 , wherein weighting each of the plurality of preliminary verification decisions includes evaluating a quality of the preliminary verification decision for each of the plurality of preliminary verification decisions.

6. The method of claim 1 , further comprising:

receiving a plurality of input utterances;

segmenting the plurality of input utterances into voiced sounds and unvoiced sounds; and

extracting features from the voiced sounds to form the plurality of different test utterances.

7. The method of claim 1 , further comprising generating a code book associated with a plurality of speakers and including the plurality of training utterances.

8. The method of claim 7 , wherein generating the code book comprises:

receiving a sample test utterance;

segmenting the sample test utterance into voiced sounds and unvoiced sounds;

extracting features from the voiced sounds; and

storing the features as one of the training utterances.

9. The method of claim 1 , further comprising selecting a male variance vector to weight each of the plurality of preliminary verification decisions based on whether the speaker is a male.

10. The method of claim 9 , further comprising determining whether the speaker purports to be a male.

11. A computer readable medium having instructions stored thereon that, when executed, cause a machine to:

compare a plurality of different test utterances to a plurality of training utterances for the speaker to form a plurality of preliminary verification decisions, one preliminary verification decision for each of the plurality of different test utterances;

weight each of the plurality of preliminary verification decisions based on a historical error rate corresponding to a respective one of the test utterances; and

combine the weighted preliminary verification decisions to form a verification decision.

12. The computer readable medium of claim 11 having instructions stored thereon that, when executed, cause the machine to:

receive a plurality of input utterances;

segment the plurality of input utterances into voiced sounds and unvoiced sounds; and

extract features from the voiced sounds to form the plurality of different test utterances.

13. The computer readable medium of claim 11 having instructions stored thereon that, when executed, cause the machine to generate a code book covering a plurality of speakers having the plurality of training utterances for each of the plurality of speakers, wherein a first utterance of the plurality of training utterances is not the same as a second utterance of the plurality of training utterances.

14. The computer readable medium of claim 11 having instructions stored thereon that, when executed, cause the machine to generate a code book associated with a plurality of speakers and including the plurality of training utterances.

15. The computer readable medium of claim 14 having instructions stored thereon that, when executed, cause the machine to generate the code book by:

receiving a sample test utterance;

segmenting the sample test utterance into voiced sounds and unvoiced sounds;

extracting features from the voiced sounds; and

storing the features as one of the training utterances.

16. The computer readable medium of claim 11 having instructions stored thereon that, when executed, cause the machine to select a male variance vector to weight each of the plurality of preliminary verification decisions based on whether the speaker is a male.

17. The computer readable medium of claim 16 having instructions stored thereon that, when executed, cause the machine to determine whether the speaker purports to be a male.

18. A computer readable medium having instructions stored thereon that, when executed, cause a machine to:

determine whether a speaker is male;

when the speaker is male, use a male variance vector to determine a weighted Euclidean distance between each of a plurality of test utterances and a plurality of training utterances;

form a preliminary verification decision for each of the plurality of test utterances to form a plurality of preliminary verification decisions;

combine the plurality of preliminary verification decisions to form a verification decision; and

weight each of the plurality of preliminary verification decisions based on a historical error rate corresponding to a respective one of the test utterances.

19. The computer readable medium of claim 18 having instructions stored thereon that, when executed, cause the machine to:

separate a plurality of speakers into a male group and a female group; and

determine the male variance vector based on the male group.

20. The computer readable medium of claim 18 having instructions stored thereon that, when executed, cause the machine to retrieve the male variance vector from a code book including the plurality of training utterances.

21. The computer readable medium of claim 18 having instructions stored thereon that, when executed, cause the machine to:

receive a plurality of input utterances;

segment the plurality of input utterances into voiced sounds and unvoiced sounds; and

extract features from the voiced sounds to form the plurality of test utterances.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY I, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041504/0952 →