IP Library Granted Patent US 8,117,033
Granted Patent B2
US 8,117,033 · App. 13/204,932 · Granted Feb 14, 2012

System and method for automatic verification of the understandability of speech

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,117,033
App. No.
13/204,932
Granted
Feb 14, 2012
Kind
B2
Abstract

Disclosed herein are systems, methods, and computer-readable storage media for processing a message received from a user to determine whether an estimate of intelligibility is below an intelligibility threshold. The method includes recognizing a portion of a user's message that contains the one or more expected utterances from a critical information list, calculating an estimate of intelligibility for the recognized portion of the user's message that contains the one or more expected utterances, and prompting the user to repeat at least the recognized portion of the user's message if the calculated estimate of intelligibility for the recognized portion of the user's message is below an intelligibility threshold. In one aspect, the method further includes prompting the user to repeat at least a portion of the message if any of a measured speech level and a measured signal-to-noise ratio of the user's message are determined to be below their respective thresholds.

Claims (37)

1. A method comprising:

receiving speech in a first language;

recognizing, via a processor, at least one phoneme within the speech with reference to a list of phonemes in the first language, to yield a recognized portion;

determining, based on the recognized portion and a phoneme distribution for the first language, an estimate of intelligibility;

if the estimate of intelligibility is below a threshold, applying a phoneme distribution for a second language to determine an estimated language; and

communicating to a user at least one of the estimate of intelligibility and the estimated language.

2. The method of claim 1 , further comprising if the estimate of intelligibility is above the threshold determining other words within the speech using the phoneme distribution for the first language.

3. The method of claim 1 , further comprising if the estimate of intelligibility is below the threshold requesting the user repeat the speech.

4. The method of claim 1 , wherein the phoneme distribution for the first language is a domain-specific phoneme distribution.

5. The method of claim 4 , further comprising if the estimate of intelligibility is above the threshold assigning a domain to the speech, wherein the domain is associated with the domain-specific phoneme distribution.

6. The method of claim 4 , wherein the estimate of intelligibility is based on a number of words within the speech found in the domain-specific phoneme distribution.

7. The method of claim 1 , wherein the first language is English and the second language is one of Spanish, German, French, Japanese, Chinese, and Portuguese.

8. A system comprising:

a processor;

a storage device storing instructions for controlling the processor to perform steps comprising:

receiving speech in a first language;

recognizing at least one phoneme within the speech with reference to a list of phonemes in the first language, to yield a recognized portion;

determining, based on the recognized portion and a phoneme distribution for the first language, an estimate of intelligibility;

if the estimate of intelligibility is below a threshold, applying a phoneme distribution for a second language to determine an estimated language; and

communicating to a user at least one of the estimate of intelligibility and the estimated language.

9. The system of claim 8 , the steps further comprising if the estimate of intelligibility is above the threshold determining other words within the speech using the phoneme distribution for the first language.

10. The system of claim 8 , the steps further comprising if the estimate of intelligibility is below the threshold requesting the user repeat the speech.

11. The system of claim 8 , wherein the phoneme distribution for the first language is a domain-specific phoneme distribution.

12. The system of claim 11 , the steps, further comprising if the estimate of intelligibility is above the threshold assigning a domain to the speech, wherein the domain is associated with the domain-specific phoneme distribution.

13. The system of claim 11 , wherein the estimate of intelligibility is based on a number of words within the speech found in the domain-specific phoneme distribution.

14. The system of claim 8 , wherein the first language is English and the second language is one of Spanish, German, French, Japanese, Chinese, and Portuguese.

15. A non-transitory computer-readable storage medium storing instructions which, when executed by a computing device, cause the computing device to perform steps comprising:

receiving speech in a first language;

recognizing at least one phoneme within the speech with reference to a list of phonemes in the first language, to yield a recognized portion;

determining, based on the recognized portion and a phoneme distribution for the first language, an estimate of intelligibility;

if the estimate of intelligibility is below a threshold, applying a phoneme distribution for a second language to determine an estimated language; and

communicating to a user at least one of the estimate of intelligibility and the estimated language.

16. The non-transitory computer-readable storage medium of claim 15 , the instructions further comprising if the estimate of intelligibility is above the threshold determining other words within the speech using the phoneme distribution for the first language.

17. The non-transitory computer-readable storage medium of claim 15 , the instructions further comprising if the estimate of intelligibility is below the threshold requesting the user repeat the speech.

18. The non-transitory computer-readable storage medium of claim 15 , wherein the phoneme distribution for the first language is a domain-specific phoneme distribution.

19. The non-transitory computer-readable storage medium of claim 18 , the instructions further comprising if the estimate of intelligibility is above the threshold assigning a domain to the speech, wherein the domain is associated with the domain-specific phoneme distribution.

20. The non-transitory computer-readable storage medium of claim 18 , wherein the estimate of intelligibility is based on a number of words within the speech found in the domain-specific phoneme distribution.

Assignments (7)
CORRECTIVE ASSIGNMENT TO CORRECT THE REMOVAL OF 7529667, 8095363, 11/169547, US0207236, US0207237, US0207235 AND 11/231452 PREVIOUSLY RECORDED ON REEL 034590 FRAME 0045. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Aug 24, 2018
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: AT&T ALEX HOLDINGS, LLC
Reel/Frame 046733/0932 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE/ASSIGNOR NAME INCORRECT ASSIGNMENT PREVIOUSLY RECORDED AT REEL: 034590 FRAME: 0045. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Jun 23, 2017
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 042962/0290 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T ALEX HOLDINGS, LLC
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041495/0903 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 10, 2014
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: AT&T ALEX HOLDINGS, LLC
Reel/Frame 034590/0045 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 12, 2014
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 033727/0547 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 12, 2014
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 033731/0184 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 10, 2014
From: COHEN, HARVEY S.; GOLDBERG, RANDY G.; ROSEN, KENNETH H.
To: AT&T CORP.
Reel/Frame 033706/0288 →