IP Library Granted Patent US 9,978,373
Granted Patent B2
US 9,978,373 · App. 15/185,298 · Granted May 22, 2018

Method of accessing a dial-up service

Inventor: Robert Wesley Bossemeyer, Jr. (St. Charles, IL)
Assignee: NUANCE COMMUNICATIONS, INC.
G10L17/04G10L15/08G10L15/10G10L17/005G10L17/24G10L25/12G10L25/24H04M3/382H04M3/385H04M3/42204H04M3/493G10L2015/0638H04M3/38H04M2201/40
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,978,373
App. No.
15/185,298
Granted
May 22, 2018
Kind
B2
Abstract

A method of accessing a dial-up service is disclosed. An example method of providing access to a service includes receiving a first speech signal from a user to form a first utterance; recognizing the first utterance using speaker independent speaker recognition; requesting the user to enter a personal identification number; and when the personal identification number is valid, receiving a second speech signal to form a second utterance and providing access to the service.

Claims (43)

1. A method comprising:

comparing, via a processor, a feature coefficient generated from a speech signal to a user-specific codebook associated with a user who provided the speech signal, to yield a similarity value, wherein the user-specific codebook comprises codes that are formed from cepstrum coefficients generated from utterances spoken by the user;

when the similarity value meets a threshold:

adding the speech signal to a database of reference speech signals; and

adding the feature coefficient to the user-specific codebook; and

performing a speaker verification process of the user based on the database of reference speech signals and the user-specific codebook.

2. The method of claim 1 , wherein the user-specific codebook utilizes utterances from both the user and a group of non-users.

3. The method of claim 1 , further comprising:

mixing the speech signal with a second speech signal, to yield a mixed speech signal; and

adding the mixed speech signal to the database of reference speech signals.

4. The method of claim 2 , wherein the speech signal and the second speech signal are received from the user.

5. The method of claim 1 , wherein the threshold is determined using a Chi-squared detector.

6. The method of claim 1 , further comprising:

when the similarity value does not meet the threshold, requesting the speech signal be repeated.

7. The method of claim 1 , further comprising verifying an identity of the user based on the similarity value.

8. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

comparing a feature coefficient generated from a speech signal to a user-specific codebook associated with a user who provided the speech signal, to yield a similarity value, wherein the user-specific codebook comprises codes that are formed from cepstrum coefficients generated from utterances spoken by the user;

when the similarity value meets a threshold:

adding the speech signal to a database of reference speech signals; and

adding the feature coefficient to the user-specific codebook; and

performing a speaker verification process of the user based on the database of reference speech signals and the user-specific codebook.

9. The system of claim 8 , wherein the user-specific codebook utilizes utterances from both the user and a group of non-users.

10. The system of claim 8 , the computer-readable storage medium having additional instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

mixing the speech signal with a second speech signal, to yield a mixed speech signal; and

adding the mixed speech signal to the database of reference speech signals.

11. The system of claim 9 , wherein the speech signal and the second speech signal are received from the user.

12. The system of claim 8 , wherein the threshold is determined using a Chi-squared detector.

13. The system of claim 8 , the computer-readable storage medium having additional instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

when the similarity value does not meet the threshold, requesting the speech signal be repeated.

14. The system of claim 8 , the computer-readable storage medium having additional instructions stored which result in operations comprising verifying an identity of the user based on the similarity value.

15. A non-transitory computer-readable storage device having instructions stored which, when executed by a computing device, cause the computing device to perform operations comprising:

comparing a feature coefficient generated from a speech signal to a user-specific codebook associated with a user who provided the speech signal, to yield a similarity value, wherein the user-specific codebook comprises codes that are formed from cepstrum coefficients generated from utterances spoken by the user;

when the similarity value meets a threshold:

adding the speech signal to a database of reference speech signals; and

adding the feature coefficient to the user-specific codebook; and

performing a speaker verification process of the user based on the database of reference speech signals and the user-specific codebook.

16. The non-transitory computer-readable storage device of claim 15 , wherein the user-specific codebook utilizes utterances from both the user and a group of non-users.

17. The non-transitory computer-readable storage device of claim 15 , having additional instructions stored which, when executed by the computing device, cause the computing device to perform operations comprising:

mixing the speech signal with a second speech signal, to yield a mixed speech signal; and

adding the mixed speech signal to the database of reference speech signals.

18. The non-transitory computer-readable storage device of claim 15 , wherein the threshold is determined using a Chi-squared detector.

Assignments (8)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY I, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041504/0952 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 23, 2016
From: AMERITECH CORPORATION
To: AMERITECH PROPERTIES, INC.
Reel/Frame 039510/0878 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 23, 2016
From: AMERITECH PROPERTIES, INC.
To: SBC HOLDINGS PROPERTIES, L.P.
Reel/Frame 039510/0969 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 23, 2016
From: SBC HOLDINGS PROPERTIES, L.P.
To: SBC PROPERTIES, L.P.
Reel/Frame 039511/0058 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 23, 2016
From: BOSSEMEYER, ROBERT WESLEY, JR.
To: AMERITECH CORPORATION
Reel/Frame 039509/0743 →
CHANGE OF NAME Recorded Aug 23, 2016
From: AT&T KNOWLEDGE VENTURES, L.P.
To: AT&T INTELLECTUAL PROPERTY I, L.P.
Reel/Frame 039785/0043 →
CHANGE OF NAME Recorded Aug 23, 2016
From: SBC KNOWLEDGE VENTURES, L.P.
To: AT&T KNOWLEDGE VENTURES, L.P.
Reel/Frame 039787/0186 →
CHANGE OF NAME Recorded Aug 23, 2016
From: SBC PROPERTIES, L.P.
To: SBC KNOWLEDGE VENTURES, L.P.
Reel/Frame 039785/0033 →
Continuity (8)
Continuation 14268078 · May 2, 2014
Continuation 13873638 · Apr 30, 2013
Continuation 13251634 · Oct 3, 2011
Continuation 12029952 · Feb 12, 2008
Continuation 11004287 · Dec 3, 2004
Continuation 08863462 · May 27, 1997
Related Publication 20160300574A1 · Oct 13, 2016
Related Publication 20170287488A9 · Oct 5, 2017