IP Library Granted Patent US 8,010,360
Granted Patent B2
US 8,010,360 · App. 12/638,604 · Granted Aug 30, 2011

System and method for latency reduction for automatic speech recognition using partial multi-pass results

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,010,360
App. No.
12/638,604
Granted
Aug 30, 2011
Kind
B2
Abstract

A system and method is provided for reducing latency for automatic speech recognition. In one embodiment, intermediate results produced by multiple search passes are used to update a display of transcribed text.

Claims (13)

1. A method for reducing latency in automatic speech recognition, the method comprising:

transcribing speech data using a first automatic speech recognition pass, which operates at a first transcription rate near real time, to produce first transcription data;

displaying the first transcription data on a display unit;

determining, via a processor of a computing device, how many additional automatic speech recognition transcription passes are needed, wherein the additional automatic speech recognition transcription passes are slower and more accurate than the first automatic speech recognition pass, to produce at least a second transcription data; and

displaying an indication of how many additional automatic speech recognition transcription passes are forthcoming that will update the first transcription data displayed on the display unit.

2. The method of claim 1 , wherein the first automatic speech recognition pass operates at real time.

3. The method of claim 1 , wherein the first automatic speech recognition pass operates at greater than real time.

4. The method of claim 1 , wherein displaying the first transcription data further comprises displaying an indicator that signifies that more accurate transcription data is being generated.

5. The method of claim 1 , wherein the first transcription data is displayed in at least one of a different color and a different shade as compared to the indication.

6. The method of claim 1 , wherein portions of the first transcription data having a relatively lower confidence score are distinctly displayed as compared to other portions of the first transcription data having a relatively higher confidence score.

7. The method of claim 6 , wherein the portions of the first transcription data having the relatively lower confidence score are displayed in a darker shade as compared to the other portions of the first transcription data having the relatively higher confidence score.

8. The method of claim 6 , wherein the portions of the first transcription data having the relatively lower confidence score enable a user to listen to corresponding portions of the speech data.

9. The method of claim 1 , wherein the additional automatic speech recognition transcription passes of the speech data are normalized automatic speech recognition passes.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041512/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 26, 2016
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 038529/0164 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 26, 2016
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 038529/0240 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 29, 2016
From: BACCHIANI, MICHIEL ADRIAAN UNICO; AMENTO, BRIAN SCOTT
To: AT&T CORP.
Reel/Frame 038127/0689 →