IP Library Granted Patent US 8,521,510
Granted Patent B2
US 8,521,510 · App. 11/469,016 · Granted Aug 27, 2013

Method and system for providing an automated web transcription service

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,521,510
App. No.
11/469,016
Granted
Aug 27, 2013
Kind
B2
Abstract

A system, method and computer readable medium that provides an automated web transcription service is disclosed. The method may include receiving input speech from a user using a communications network, recognizing the received input speech, understanding the recognized speech, transcribing the understood speech to text, storing the transcribed text in a database, receiving a request via a web page to display the transcribed text, retrieving transcribed text from the database, and displaying the transcribed text to the requester using the web page.

Claims (68)

1. A method comprising:

verifying, via a processor, an identity of a first user;

when the identity of the first user is verified:

identifying a template for a domain associated with the first user;

receiving input speech from the first user, the input speech comprising a substantive portion and an instructional portion, the instructional portion related to navigation between fields in the template;

transcribing the substantive portion of the input speech to text based on the domain, to yield transcribed text; and

storing the transcribed text in a database;

upon receiving a first request via a web page-from a second user to display the transcribed text:

retrieving the transcribed text from the database; and

displaying the transcribed text to the second user; and

upon receiving a second request from the second user to play a dictation for a particular word in the transcribed text, playing the dictation of the particular word.

2. The method of claim 1 , further comprising:

recognizing and understanding system commands spoken by the second user, wherein the system commands facilitate navigation of the second user through the transcribed text and a correction of the transcribed text.

3. The method of claim 1 , further comprising:

prompting the second user to correct the transcribed text;

receiving correction inputs for the transcribed text from the second user; and

correcting the transcribed text according to the correction inputs.

4. The method of claim 3 , further comprising:

calculating an accuracy of the transcribed text based on the correction inputs; and

adjusting a cost to the first user for the transcribed text based on the accuracy of the transcribed text.

5. The method of claim 1 , further comprising:

prompting the second user to select one of printing the transcribed text, sending the transcribed text to another party, and saving the transcribed text to a memory.

6. The method of claim 1 , wherein transcribing the substantive portion of the input speech further comprises transcribing the substantive portion into a predefined document template.

7. The method of claim 1 , wherein transcribing the substantive portion of the input speech further comprises highlighting the transcribed text based on recognition confidence levels.

8. The method of claim 1 , further comprising:

learning from correction data and new data to yield learned correction data; and

adapting the method based on the learned correction data and new data.

9. A computer-readable storage device having instructions stored which, when executed by a computing device, cause the computing device to perform operations comprising:

verifying, via a processor, an identity of a first user;

when the identity of the first user is verified:

identifying a template for a domain associated with the first user;

receiving input speech from the first user, the input speech comprising a substantive portion and an instructional portion, the instructional portion related to navigation between fields in the template;

transcribing the substantive portion of the input speech to text based on the domain, to yield transcribed text; and

storing the transcribed text in a database;

upon receiving a first request via a web page-from a second user to display the transcribed text:

retrieving the transcribed text from the database; and

displaying the transcribed text to the second user; and

upon receiving a second request from the second user to play a dictation for a particular word in the transcribed text, playing the dictation of the particular word.

10. The computer-readable storage device of claim 9 , wherein the telephone communications network is one of a computer network, an internet, and a telephone network.

11. The computer-readable storage device of claim 9 , wherein the computer-readable device has additional instructions stored which result in the operations further comprising:

recognizing and understanding system commands spoken by the second user, wherein the system commands facilitate navigation of the second user through the transcribed text and a correction of the transcribed text.

12. The computer-readable storage device of claim 9 , wherein the computer-readable device has additional instructions stored which result in the operations further comprising:

prompting the second user to correct the transcribed text;

receiving correction inputs for the transcribed text from the second user; and

correcting the transcribed text according to the correction inputs.

13. The computer-readable storage device of claim 12 , wherein the computer-readable device has additional instructions stored which result in the operations further comprising:

calculating an accuracy of the transcribed text based on the correction inputs; and

adjusting a cost to the first user for the transcribed text based on the accuracy.

14. The computer-readable storage device of claim 9 , wherein the computer-readable device has additional instructions stored which result in the operations further comprising:

prompting the second user to select one of printing the transcribed text, sending the transcribed text to another party, and saving the transcribed text to a memory.

15. The computer-readable storage device of claim 9 , wherein transcribing the substantive portion of the input speech further comprises transcribing the input speech into a predefined document template.

16. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed by the processor, result in the processor performing operations comprising:

verifying, via a processor, an identity of a first user;

when the identity of the first user is verified:

identifying a template for a domain associated with the first user;

receiving input speech from the first user, the input speech comprising a substantive portion and an instructional portion, the instructional portion related to navigation between fields in the template;

transcribing the substantive portion of the input speech to text based on the domain, to yield transcribed text; and

storing the transcribed text in a database;

upon receiving a first request via a web page-from a second user to display the transcribed text:

retrieving the transcribed text from the database; and

displaying the transcribed text to the second user; and

upon receiving a second request from the second user to play a dictation for a particular word in the transcribed text, playing the dictation of the particular word.

17. The system of claim 16 , wherein the telephone communications network is one of a computer network, an internet, and a telephone network.

18. The system of claim 16 , the computer-readable storage medium having additional instructions stored which result in the operations further comprising proving navigation to the second user through the transcribed text and a correction of the transcribed text.

19. The system of claim 16 , the computer-readable storage medium having additional instructions stored which result in the operations further comprising prompting the second user to correct the transcribed text, receiving correction inputs for the transcribed text from the second user, and correcting the transcribed text according to the correction inputs.

20. The system of claim 16 , wherein an automated web transcription service unit prompts the second user to select one of printing the transcribed text, sending the transcribed text to another party, and saving the transcribed text to a memory.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 9, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065533/0389 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041512/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2016
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 038275/0238 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2016
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 038275/0310 →