IP Library Granted Patent US 8,355,916
Granted Patent B2
US 8,355,916 · App. 13/485,574 · Granted Jan 15, 2013

Systems and methods for extracting meaning from multimodal inputs using finite-state devices

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,355,916
App. No.
13/485,574
Granted
Jan 15, 2013
Kind
B2
Abstract

Multimodal utterances contain a number of different modes. These modes can include speech, gestures, and pen, haptic, and gaze inputs, and the like. This invention use recognition results from one or more of these modes to provide compensation to the recognition process of one or more other ones of these modes. In various exemplary embodiments, a multimodal recognition system inputs one or more recognition lattices from one or more of these modes, and generates one or more models to be used by one or more mode recognizers to recognize the one or more other modes. In one exemplary embodiment, a gesture recognizer inputs a gesture input and outputs a gesture recognition lattice to a multimodal parser. The multimodal parser generates a language model and outputs it to an automatic speech recognition system, which uses the received language model to recognize the speech input that corresponds to the recognized gesture input.

Claims (34)

1. An apparatus for recognizing an utterance comprising:

means for receiving an utterance, the utterance comprising a first portion having a first mode and a second portion having a second mode;

means for generating a first mode recognition lattice and a first finite-state transducer, based on the first mode;

means for relating the first portion of the utterance to the second portion of the utterance based on the first finite-state transducer; and

means for generating a second finite-state transducer, the second finite-state transducer comprising a gesture and speech recognition model finite-state transducer, based on the first mode recognition lattice and the first finite-state transducer.

2. The apparatus of claim 1 , further comprising:

means for recognizing the second mode based on the second finite-state transducer.

3. The apparatus of claim 1 wherein the first mode is a gesture mode.

4. The apparatus of claim 1 wherein the second mode is a speech mode.

5. The apparatus of claim 1 wherein the generating the first mode recognition lattice is further based on a first mode feature lattice.

6. An apparatus for recognizing an utterance comprising:

means for receiving an utterance comprising a plurality of modes;

means for relating a first portion of the utterance comprising a first mode to a second portion of the utterance comprising a second mode, based on a first finite-state transducer; and

means for generating a second finite-state transducer, the second finite-state transducer comprising a gesture and speech recognition model finite-state transducer, based on a first mode recognition lattice and the first finite-state transducer.

7. The apparatus of claim 6 further comprising:

means for generating the first mode recognition lattice and the first finite-state transducer, based on the first mode.

8. The apparatus of claim 7 further comprising:

means for recognizing the second mode based on the second finite-state transducer.

9. The apparatus of claim 7 wherein the first mode is a gesture mode.

10. The apparatus of claim 7 wherein the second mode is a speech mode.

11. The apparatus of claim 7 wherein the generating the first mode recognition lattice is further based on a first mode feature lattice.

12. The apparatus of claim 11 further comprising:

means for generating the first mode feature lattice based on the utterance.

13. A apparatus for extracting meaning from multimodal inputs comprising:

means for receiving a multimodal input;

means for relating a first portion of the multimodal input comprising a first mode to a second portion of the multimodal input comprising a second mode, based on a first finite-state transducer; and

means for generating a second finite-state transducer, the second finite-state transducer comprising a gesture and speech recognition model finite-state transducer, based on a first mode recognition lattice and the first finite-state transducer.

14. The apparatus of claim 13 further comprising:

means for generating the first mode recognition lattice and the first finite-state transducer, based on the first mode.

15. The apparatus of claim 14 wherein the first mode is a gesture mode.

16. The apparatus of claim 14 wherein the second mode is a speech mode.

17. The apparatus of claim 14 wherein the generating the first mode recognition lattice is further based on a first mode feature lattice.

18. The apparatus of claim 17 further comprising:

means for generating the first mode feature lattice based on the multimodal input.

Assignments (18)
RELEASE OF SECURITY INTEREST Recorded Sep 4, 2025
From: RUNWAY GROWTH FINANCE CORP., AS AGENT
To: INTERACTIONS CORPORATION; INTERACTIONS LLC
Reel/Frame 072802/0931 →
CORRECTIVE ASSIGNMENT TO CORRECT THE THE APPLICATION NUMBER PREVIOUSLY RECORDED AT REEL: 060445 FRAME: 0733. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Feb 1, 2023
From: INTERACTIONS LLC; INTERACTIONS CORPORATION
To: RUNWAY GROWTH FINANCE CORP.
Reel/Frame 062919/0063 →
RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY RECORDED AT REEL/FRAME: 036100/0925 Recorded Jul 1, 2022
From: SILICON VALLEY BANK
To: INTERACTIONS LLC
Reel/Frame 060559/0576 →
RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY RECORDED AT REEL/FRAME: 049388/0082 Recorded Jun 30, 2022
From: SILICON VALLEY BANK
To: INTERACTIONS LLC
Reel/Frame 060558/0474 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jun 27, 2022
From: INTERACTIONS LLC; INTERACTIONS CORPORATION
To: RUNWAY GROWTH FINANCE CORP.
Reel/Frame 060445/0733 →
TERMINATION AND RELEASE OF SECURITY INTEREST IN INTELLECTUAL PROPERTY Recorded May 23, 2022
From: ORIX GROWTH CAPITAL, LLC
To: INTERACTIONS CORPORATION; INTERACTIONS LLC
Reel/Frame 061749/0825 →
RELEASE OF SECURITY INTEREST Recorded May 18, 2020
From: BEARCUB ACQUISITIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 052693/0866 →
INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jun 5, 2019
From: INTERACTIONS LLC
To: SILICON VALLEY BANK
Reel/Frame 049388/0082 →
ASSIGNMENT OF IP SECURITY AGREEMENT Recorded Nov 17, 2017
From: ARES VENTURE FINANCE, L.P.
To: BEARCUB ACQUISITIONS LLC
Reel/Frame 044481/0034 →
CORRECTIVE ASSIGNMENT TO CORRECT THE CHANGE PATENT 7146987 TO 7149687 PREVIOUSLY RECORDED ON REEL 036009 FRAME 0349. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Nov 17, 2015
From: INTERACTIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 037134/0712 →
FIRST AMENDMENT TO INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Jul 13, 2015
From: INTERACTIONS LLC
To: SILICON VALLEY BANK
Reel/Frame 036100/0925 →
SECURITY INTEREST Recorded Jun 23, 2015
From: INTERACTIONS LLC
To: ARES VENTURE FINANCE, L.P.
Reel/Frame 036009/0349 →
SECURITY INTEREST Recorded Dec 19, 2014
From: INTERACTIONS LLC
To: ORIX VENTURES, LLC
Reel/Frame 034677/0768 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 16, 2014
From: AT&T ALEX HOLDINGS, LLC
To: INTERACTIONS LLC
Reel/Frame 034642/0640 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 10, 2014
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: AT&T ALEX HOLDINGS, LLC
Reel/Frame 034467/0822 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 8, 2014
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 034429/0467 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 8, 2014
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 034429/0474 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 31, 2012
From: BANGALORE, SRINIVAS; JOHNSTON, MICHAEL J.
To: AT&T CORP.
Reel/Frame 028299/0466 →