IP Library Granted Patent US 8,521,534
Granted Patent B2
US 8,521,534 · App. 13/612,014 · Granted Aug 27, 2013

Dynamically extending the speech prompts of a multimodal application

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,521,534
App. No.
13/612,014
Granted
Aug 27, 2013
Kind
B2
Abstract

A prompt generation engine operates to dynamically extend prompts of a multimodal application. The prompt generation engine receives a media file having a metadata container. The prompt generation engine operates on a multimodal device that supports a voice mode and a non-voice mode for interacting with the multimodal device. The prompt generation engine retrieves from the metadata container a speech prompt related to content stored in the media file for inclusion in the multimodal application. The prompt generation engine modifies the multimodal application to include the speech prompt.

Claims (30)

1. A method of dynamically extending the speech prompts of a multimodal application, the method comprising:

receiving, by a prompt generation engine, a media file having a metadata container, wherein the prompt generation engine operates on a multimodal device that supports a voice mode and a non-voice mode for interacting with the multimodal device;

retrieving, by the prompt generation engine from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application; and

modifying, by the prompt generation engine, the multimodal application to include the speech prompt.

2. The method of claim 1 wherein retrieving, by the prompt generation engine, from the metadata container a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprises retrieving a text string prompt for execution by a text to speech engine.

3. The method of claim 1 wherein retrieving, by the prompt generation engine, from the metadata container a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprises retrieving an audio prompt to be played by the multimodal device.

4. The method of claim 1 wherein retrieving, by the prompt generation engine, from the metadata container a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprises identifying a tag for prompts in the metadata container.

5. The method of claim 4 wherein identifying a tag for prompts in the metadata container further comprises identifying a frame for prompts in an ID3 container of an MPEG media file.

6. The method of claim 1 wherein modifying, by the prompt generation engine, the multimodal application to include the speech prompt further comprises updating a prompt document with the retrieved speech prompt.

7. A multimodal device that supports multiple modes for interacting with the multimodal device, the multimodal device comprising:

a computer processor;

a computer memory operatively coupled to the computer processor, the computer memory having disposed within it computer program instructions configured to:

receive a media file having a metadata container;

retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in a multimodal application; and

modify the multimodal application to include the speech prompt.

8. The multimodal device of claim 7 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to retrieve a text string prompt for execution by a text to speech engine.

9. The multimodal device of claim 7 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to retrieve an audio prompt to be played by the multimodal device.

10. The multimodal device of claim 7 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to identify a tag for prompts in the metadata container.

11. The multimodal device of claim 10 wherein computer program instructions configured to identify a tag for prompts in the metadata container further comprise computer program instructions configured to identify a frame for prompts in an ID3 container of an MPEG media file.

12. The multimodal device of claim 7 wherein computer program instructions configured to modify the multimodal application to include the speech prompt further comprise computer program instructions configured to update a prompt document with the retrieved speech prompt.

13. A computer program product for dynamically extending speech prompts of a multimodal application, the computer program product comprising:

a recordable media having computer program instructions stored therein, the computer program instructions configured to:

receive a media file having a metadata container;

retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application; and

modify the multimodal application to include the speech prompt.

14. The computer program product of claim 13 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to retrieve a text string prompt for execution by a text to speech engine.

15. The computer program product of claim 13 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to retrieve an audio prompt to be played by the multimodal device.

16. The computer program product of claim 13 wherein computer program instructions configured to retrieve, from the metadata container, a speech prompt related to content stored in the media file for inclusion in the multimodal application further comprise computer program instructions configured to identify a tag for prompts in the metadata container.

17. The computer program product of claim 16 wherein computer program instructions configured to identify a tag for prompts in the metadata container further comprise computer program instructions configured to identify a frame for prompts in an ID3 container of an MPEG media file.

18. The computer program product of claim 13 wherein computer program instructions configured to modify the multimodal application to include the speech prompt further comprise computer program instructions configured to update a prompt document with the retrieved speech prompt.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 1, 2013
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 030323/0965 →