IP Library Granted Patent US 7,548,859
Granted Patent B2
US 7,548,859 · App. 11/196,112 · Granted Jun 16, 2009

Method and system for assisting users in interacting with multi-modal dialog systems

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,548,859
App. No.
11/196,112
Granted
Jun 16, 2009
Kind
B2
Abstract

A method and system for assisting a user in interacting with a multi-modal dialog system ( 104 ) is provided. The method includes interpreting a “What Can I Do? (WCID)” question from a user in a turn of the dialog. A multi-modal grammar ( 212 ) is generated, based on the current context of the dialog. One or more user multi-modal utterances are generated, based on the WCID question and the multi-modal grammar. One or more user multi-modal utterances are conveyed to the user.

Claims (22)

1. A method for assisting a user in interacting with a multi-modal dialog apparatus, the method comprising:

interpreting by a processor of the multi-modal dialog apparatus a What Can I Do (WCID) type of question from the user that is received at a user interface device coupled to the multi-modal dialog apparatus;

generating by the processor one or more user multi-modal utterances as sequences of words that a user may provide in the multi-modal inputs in a next turn of the dialog based on the WCID question and a multi-modal grammar, wherein the multi-modal grammar is based on a current context of a multi-modal dialog; and

conveying the one or more user multi-modal utterances to the user at a user interface device coupled to the multi-modal dialog apparatus.

2. The method according to claim 1 further comprising de-referencing the one or more user multi-modal utterances using a visual context of the multi-modal dialog.

3. The method according to claim 2 further comprising storing the visual context of the multi-modal dialog.

4. The method according to claim 1 , wherein generating the one or more user multi-modal utterances comprises generating the multi-modal grammar based on the current context of the multi-modal dialog.

5. The method according to claim 1 further comprising ranking the one or more user multi-modal utterances based on one or more of a group consisting of the current context of the multi-modal dialog, a user model and modality availability.

6. A multi-modal dialog apparatus comprising:

a user interface device;

a processor; and

a memory, wherein the memory stores programming instructions organized into functional groups that control the processor, the functional groups comprising

a multi-modal input fusion (MMIF) component, the multi-modal input fusion component accepting a What Can I Do (WCID) type of question from the user through the user interface device;

a dialog manager, the dialog manager generating a multi-modal grammar based on a current context of a multi-modal dialog;

a multi-modal utterance generator, the multi-modal utterance generator generating one or more user multi-modal utterances through the user interface device, as sequences of words that a user may provide in the multi-modal inputs in a next turn of the dialog based on the question and the multi-modal grammar.

7. The multi-modal dialog apparatus of claim 6 wherein the functional groups further comprise a multi-modal utterance ranker, the multi-modal utterance ranker ranking the one or more user multi-modal utterances based on one or more of a group consisting of a current context of the multi-modal dialog, a user model and modality availability.

8. The multi-modal dialog apparatus of claim 6 wherein the functional groups further comprise a visual context manager, the visual context manager storing a visual context of the multi-modal dialog.

9. The multi-modal dialog apparatus of claim 8 , wherein the multi-modal utterance generator generates the one or more user multi-modal utterances based on the visual context of the multi-modal dialog.

10. The multi-modal dialog apparatus of claim 6 wherein the functional groups further comprise an output generator, the output generator conveying the one or more user multi-modal utterances to the user.

11. The multi-modal dialog apparatus of claim 6 , wherein the MMIF component maintains and updates the modality availability.

12. The multi-modal dialog apparatus of claim 6 , wherein the dialog manager interprets the question from the user.

13. The multi-modal dialog apparatus of claim 6 , wherein the dialog manager maintains and updates the current context of the dialog.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 24, 2014
From: MOTOROLA MOBILITY LLC
To: GOOGLE TECHNOLOGY HOLDINGS LLC
Reel/Frame 034419/0001 →
CHANGE OF NAME Recorded Oct 2, 2012
From: MOTOROLA MOBILITY, INC.
To: MOTOROLA MOBILITY LLC
Reel/Frame 029216/0282 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 13, 2010
From: MOTOROLA, INC
To: MOTOROLA MOBILITY, INC
Reel/Frame 025673/0558 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 13, 2006
From: THOMPSON, WILLIAM K.; LEE, HANG S.; GUPTA, ANURAG K.
To: MOTOROLA, INC.
Reel/Frame 017333/0374 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 18, 2006
From: THOMPSON, WILLIAM K.; GUPTA, ANURAG K.; LEE, HANG S.
To: MOTOROLA, INC.
Reel/Frame 017495/0851 →