IP Library Granted Patent US 7,860,717
Granted Patent B2
US 7,860,717 · App. 10/951,291 · Granted Dec 28, 2010

System and method for customizing speech recognition input and output

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,860,717
App. No.
10/951,291
Granted
Dec 28, 2010
Kind
B2
Abstract

A system and method may be disclosed for facilitating the site-specific customization of automated speech recognition systems by providing a customization client for site-specific individuals to update and modify language model input files and post processor input files. In customizing the input files, the customization client may provide a graphical user interface for facilitating the inclusion of words specific to a particular site. The customization client may also be configured to provide the user with a series of formatting rules for controlling the appearance and format of a document transcribed by an automated speech recognition system.

Claims (29)

1. A method for customizing input data for a language model in a speech recognition system, the language model having a predetermined data format for receiving words and/or phrases to be added to the language model, the method comprising:

providing a graphical user interface (GUI) that allows a user to specify a plurality of words and/or phrases in a first format independent of the predetermined data format of the language model; and

automatically converting the specified plurality of words and/or phrases in the first format, via at least one computer system, to language model input data in a second format compatible with the predetermined data format of the language model.

2. The method according to claim 1 , wherein providing the GUI comprises formatting the plurality of words and/or phrases into flat files.

3. The method according to claim 1 , wherein the automatically converting the specified plurality of words and/or phrases comprises normalizing the specified plurality of words and/or phrases for use by the language model.

4. The method according to claim 1 , wherein the automatically converting the specified plurality of words and/or phrases comprises filtering the specified plurality of words and/or phrases for use by the language model.

5. The method according to claim 1 , wherein providing the GUI comprises displaying words and/or phrases for possible selection by the user via the GUI.

6. The method according to claim 5 , wherein the displaying the words and/or phrases comprises normalizing the words and/or phrases for display to the user.

7. The method according to claim 5 , wherein the displaying the words and/or phrases comprises filtering the words and/or phrases for display to the user.

8. The method according to claim 1 , wherein providing the GUI comprises importing words and/or phrases for possible selection by the user via the GUI.

9. The method according to claim 1 , wherein providing the GUI comprises providing buttons representing predetermined options for the user to select via the GUI.

10. The method according to claim 1 , further comprising adding the specified plurality of words and/or phrases to the language model using the language model input data in the second format.

11. A method for customizing input data for a post processor in a speech recognition system, the post processor having a predetermined data format for receiving formatting rules to be used in formatting text documents, the method comprising:

providing a graphical user interface (GUI) that allows a user to specify formatting rules in a first format independent of the predetermined data format of the post processor; and

automatically converting the specified formatting rules in the first format, via at least one computer system, to post processor input data in a second format compatible with the predetermined data format of the post processor.

12. The method according to claim 11 , wherein providing the GUI comprises providing buttons representing predetermined options for the user to select via the GUI.

13. The method according to claim 11 , further comprising configuring the post processor to format text documents in accordance with the specified formatting rules using the post processor input data in the second format.

14. A system for customizing input data for a language model in a speech recognition system, the language model having a predetermined data format for receiving words and/or phrases to be added to the language model, the system comprising at least one computer system configured to:

provide a graphical user interface (GUI) that allows a user to specify a plurality of words and/or phrases in a first format independent of the predetermined data format of the language model; and

automatically convert the specified plurality of words and/or phrases in the first format to language model input data in a second format compatible with the predetermined data format of the language model.

15. At least one computer-readable storage medium encoded with a plurality of computer-executable instructions that, when executed, perform a method for customizing input data for a language model in a speech recognition system, the language model having a predetermined data format for receiving words and/or phrases to be added to the language model, the method comprising:

providing a graphical user interface (GUI) that allows a user to specify a plurality of words and/or phrases in a first format independent of the predetermined data format of the language model; and

automatically converting the specified plurality of words and/or phrases in the first format to language model input data in a second format compatible with the predetermined data format of the language model.

16. A system for customizing input data for a post processor in a speech recognition system, the post processor having a predetermined data format for receiving formatting rules to be used in formatting text documents, the system comprising at least one computer system configured to:

provide a graphical user interface (GUI) that allows a user to specify formatting rules in a first format independent of the predetermined data format of the post processor; and

automatically convert the specified formatting rules in the first format to post processor input data in a second format compatible with the predetermined data format of the post processor.

17. At least one computer-readable storage medium encoded with a plurality of computer-executable instructions that, when executed, perform a method for customizing input data for a post processor in a speech recognition system, the post processor having a predetermined data format for receiving formatting rules to be used in formatting text documents, the method comprising:

providing a graphical user interface (GUI) that allows a user to specify formatting rules in a first format independent of the predetermined data format of the post processor; and

automatically converting the specified formatting rules in the first format to post processor input data in a second format compatible with the predetermined data format of the post processor.

Assignments (8)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 9, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065531/0665 →
PATENT RELEASE (REEL:017435/FRAME:0199) Recorded May 20, 2016
From: MORGAN STANLEY SENIOR FUNDING, INC., AS ADMINISTRATIVE AGENT
To: NUANCE COMMUNICATIONS, INC., AS GRANTOR; ART ADVANCED RECOGNITION TECHNOLOGIES, INC., A DELAWARE CORPORATION, AS GRANTOR; SPEECHWORKS INTERNATIONAL, INC., A DELAWARE CORPORATION, AS GRANTOR; TELELOGUE, INC., A DELAWARE CORPORATION, AS GRANTOR; DSP, INC., D/B/A DIAMOND EQUIPMENT, A MAINE CORPORATON, AS GRANTOR; SCANSOFT, INC., A DELAWARE CORPORATION, AS GRANTOR; DICTAPHONE CORPORATION, A DELAWARE CORPORATION, AS GRANTOR
Reel/Frame 038770/0824 →
PATENT RELEASE (REEL:018160/FRAME:0909) Recorded May 20, 2016
From: MORGAN STANLEY SENIOR FUNDING, INC., AS ADMINISTRATIVE AGENT
To: NUANCE COMMUNICATIONS, INC., AS GRANTOR; ART ADVANCED RECOGNITION TECHNOLOGIES, INC., A DELAWARE CORPORATION, AS GRANTOR; SPEECHWORKS INTERNATIONAL, INC., A DELAWARE CORPORATION, AS GRANTOR; TELELOGUE, INC., A DELAWARE CORPORATION, AS GRANTOR; DSP, INC., D/B/A DIAMOND EQUIPMENT, A MAINE CORPORATON, AS GRANTOR; HUMAN CAPITAL RESOURCES, INC., A DELAWARE CORPORATION, AS GRANTOR; INSTITIT KATALIZA IMENI G.K. BORESKOVA SIBIRSKOGO OTDELENIA ROSSIISKOI AKADEMII NAUK, AS GRANTOR; NOKIA CORPORATION, AS GRANTOR; MITSUBISH DENKI KABUSHIKI KAISHA, AS GRANTOR; STRYKER LEIBINGER GMBH & CO., KG, AS GRANTOR; NORTHROP GRUMMAN CORPORATION, A DELAWARE CORPORATION, AS GRANTOR; SCANSOFT, INC., A DELAWARE CORPORATION, AS GRANTOR; DICTAPHONE CORPORATION, A DELAWARE CORPORATION, AS GRANTOR
Reel/Frame 038770/0869 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 9, 2013
From: DICTAPHONE CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 029596/0836 →
MERGER Recorded Sep 13, 2012
From: DICTAPHONE CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 028952/0397 →
SECURITY AGREEMENT Recorded Aug 24, 2006
From: NUANCE COMMUNICATIONS, INC.
To: USB AG. STAMFORD BRANCH
Reel/Frame 018160/0909 →
SECURITY AGREEMENT Recorded Apr 7, 2006
From: NUANCE COMMUNICATIONS, INC.
To: USB AG, STAMFORD BRANCH
Reel/Frame 017435/0199 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 19, 2005
From: UHRBACH, AMY J.; FRANKEL, ALAN; CARRIER, JILL; SANTISTEBAN, ANA; COTE, WILLIAM F.
To: DICTAPHONE CORPORATION
Reel/Frame 015606/0497 →