IP Library Granted Patent US 7,840,409
Granted Patent B2
US 7,840,409 · App. 11/679,284 · Granted Nov 23, 2010

Ordering recognition results produced by an automatic speech recognition engine for a multimodal application

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,840,409
App. No.
11/679,284
Granted
Nov 23, 2010
Kind
B2
Abstract

Ordering recognition results produced by an automatic speech recognition (‘ASR’) engine for a multimodal application implemented with a grammar of the multimodal application in the ASR engine, with the multimodal application operating in a multimodal browser on a multimodal device supporting multiple modes of interaction including a voice mode and one or more non-voice modes, the multimodal application operatively coupled to the ASR engine through a VoiceXML interpreter, includes: receiving, in the VoiceXML interpreter from the multimodal application, a voice utterance; determining, by the VoiceXML interpreter using the ASR engine, a plurality of recognition results in dependence upon the voice utterance and the grammar; determining, by the VoiceXML interpreter according to semantic interpretation scripts of the grammar, a weight for each recognition result; and sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result.

Claims (32)

1. A computer-implemented method of ordering recognition results produced by an automatic speech recognition (‘ASR’) engine for a multimodal application, the method implemented with a grammar of the multimodal application in the ASR engine, with the multimodal application operating in a multimodal browser on a multimodal device supporting multiple modes of interaction including a voice mode and one or more non-voice modes, the multimodal application operatively coupled to the ASR engine through a VoiceXML interpreter, the method comprising:

receiving, in the VoiceXML interpreter from the multimodal application, a voice utterance;

determining, by the VoiceXML interpreter using the ASR engine, a plurality of recognition results in dependence upon the voice utterance and the grammar;

determining, by the VoiceXML interpreter according to semantic interpretation scripts of the grammar, a weight for each recognition result; and

sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result.

2. The method of claim 1 wherein sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result comprises sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result in accordance with a sorting attribute for the grammar defining a sorting scheme, the sorting attribute specified by the multimodal application using a VoiceXML <grammar> element.

3. The method of claim 2 wherein the sorting attribute specifies sorting the plurality of recognition results according to a value of the weight for each recognition result.

4. The method of claim 2 wherein the sorting attribute specifies sorting the plurality of recognition results according to an ECMAScript script.

5. The method of claim 1 wherein the semantic interpretation scripts of the grammar statically define the weight for each recognition result.

6. The method of claim 1 wherein the semantic interpretation scripts of the grammar dynamically define the weight for each recognition result.

7. Apparatus for implementing a method of ordering recognition results produced by an automatic speech recognition (‘ASR’) engine for a multimodal application, the method implemented with a grammar of the multimodal application in the ASR engine, with the multimodal application operating in a multimodal browser on a multimodal device supporting multiple modes of interaction including a voice mode and one or more non-voice modes, the multimodal application operatively coupled to the ASR engine through a VoiceXML interpreter, the apparatus comprising:

at least one computer processor; and

a computer memory operatively coupled to the at least one computer processor, the computer memory storing computer program instructions which, when executed by the at least one computer processor, cause performance of the method, the method comprising:

receiving, in the VoiceXML interpreter from the multimodal application, a voice utterance;

determining, by the VoiceXML interpreter using the ASR engine, a plurality of recognition results in dependence upon the voice utterance and the grammar;

determining, by the VoiceXML interpreter according to semantic interpretation scripts of the grammar, a weight for each recognition result; and

sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result.

8. The apparatus of claim 7 wherein sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result comprises sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result in accordance with a sorting attribute for the grammar defining a sorting scheme, the sorting attribute specified by the multimodal application using a VoiceXML <grammar> element.

9. The apparatus of claim 8 wherein the sorting attribute specifies sorting the plurality of recognition results according to a value of the weight for each recognition result.

10. The apparatus of claim 8 wherein the sorting attribute specifies sorting the plurality of recognition results according to an ECMAScript script.

11. The apparatus of claim 7 wherein the semantic interpretation scripts of the grammar statically define the weight for each recognition result.

12. The apparatus of claim 7 wherein the semantic interpretation scripts of the grammar dynamically define the weight for each recognition result.

13. A computer program product for ordering recognition results produced by an automatic speech recognition (‘ASR’) engine for a multimodal application, with the multimodal application operating in a multimodal browser on a multimodal device supporting multiple modes of interaction including a voice mode and one or more non-voice modes, the multimodal application operatively coupled to the ASR engine through a VoiceXML interpreter, the computer program product disposed upon a recordable computer-readable medium, the computer program product comprising computer program instructions which, when executed, cause performance of a method comprising:

receiving, in the VoiceXML interpreter from the multimodal application, a voice utterance;

determining, by the VoiceXML interpreter using the ASR engine, a plurality of recognition results in dependence upon the voice utterance and a grammar of the multimodal application;

determining, by the VoiceXML interpreter according to semantic interpretation scripts of the grammar, a weight for each recognition result; and

sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result.

14. The computer program product of claim 13 wherein the semantic interpretation scripts of the grammar statically define the weight for each recognition result.

15. The computer program product of claim 13 wherein the semantic interpretation scripts of the grammar dynamically define the weight for each recognition result.

16. The computer program product of claim 13 wherein sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result comprises sorting, by the VoiceXML interpreter, the plurality of recognition results in dependence upon the weight for each recognition result in accordance with a sorting attribute for the grammar defining a sorting scheme, the sorting attribute specified by the multimodal application using a VoiceXML <grammar> element.

17. The computer program product of claim 16 wherein the sorting attribute specifies sorting the plurality of recognition results according to a value of the weight for each recognition result.

18. The computer program product of claim 16 wherein the sorting attribute specifies sorting the plurality of recognition results according to an ECMAScript script.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 9, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065532/0152 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 13, 2009
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 022689/0317 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ATTORNEY DOCKET NUMBER PREVIOUSLY RECORDED ON REEL 019076 FRAME 0963. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 5, 2007
From: ATIVANICHAYAPHONG, SOONTHORN; CROSS, CHARLES W., JR.; JABLOKOV, IGOR R.; MCCOBB, GERALD M.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 019118/0867 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2007
From: ATIVANICHAYAPHONG, SOONTHORN; CROSS, CHARLES W., JR.; JABLOKOV, IGOR R.; MCCOBB, GERALD M.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 019076/0963 →