IP Library Granted Patent US 9,430,467
Granted Patent B2
US 9,430,467 · App. 14/974,529 · Granted Aug 30, 2016

Mobile speech-to-speech interpretation system

Inventors: Farzad Ehsani (Sunnyvale, CA); Demitrios Master (Cupertino, CA); Elaine Drom Zuber (Cupertino, CA)
Assignee: Nant Holdings IP, LLC
G06F17/289G06F17/28G06F17/2818G06F17/2854G10L13/00G10L15/005G10L15/02G10L15/26G10L15/30
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,430,467
App. No.
14/974,529
Granted
Aug 30, 2016
Kind
B2
Abstract

Interpretation from a first language to a second language via one or more communication devices is performed through a communication network (e.g. phone network or the internet) using a server for performing recognition and interpretation tasks, comprising the steps of: receiving an input speech utterance in a first language on a first mobile communication device; conditioning said input speech utterance; first transmitting said conditioned input speech utterance to a server; recognizing said first transmitted speech utterance to generate one or more recognition results; interpreting said recognition results to generate one or more interpretation results in an interlingua; mapping the interlingua to a second language in a first selected format; second transmitting said interpretation results in the first selected format to a second mobile communication device; and presenting said interpretation results in a second selected format on said second communication device.

Claims (29)

1. A computer translation system comprising

A language input device that detects a first language signal and a device location; and

An interpretation engine coupled with the language input device and configured to:

obtain the first language signal and the device location;

activate a language model, including a mobile interference model, based on the first language signal and at least in part based on the device location;

update utterance probabilities of the language model based the device location;

extract a conceptual interlingua result set from the first language signal according to the updated language model and by applying the mobile interference model;

map the conceptual interlingua result set to a second language representation of a second language; and

cause an output device to present an output interpretation according to the second language derived from the second language representation.

2. The system of claim 1 , wherein the language input device comprises the interpretation engine.

3. The system of claim 1 , wherein the language input device comprises a mobile device.

4. The system of claim 3 , wherein the output device is a second mobile device at a different location that the language input device.

5. The system of claim 1 , wherein the output device is the language input device.

6. The system of claim 1 , wherein the device location corresponds to a user position.

7. The system of claim 1 , wherein the interpretation engine is further configured to boost scores for results in the conceptual interlingua result set for an entity in proximity to the language input device based on the device location.

8. The system of claim 7 , wherein the entity includes at least one of the following: a street name and a landmark.

9. The system of claim 7 , wherein the entity is named in the first language signal.

10. The system of claim 1 , wherein the output interpretation includes at least one of the following: text, a spoken language according to the second language, audio, a still image, a video, and a sign language.

11. The system of claim 1 , wherein the activated language model relates to a nurse.

12. The system of claim 1 , wherein the activated language model relates to a pharmacist.

13. The system of claim 1 , wherein the activated language model relates to a tour.

14. The system of claim 1 , wherein the activated language model relates to a sign language.

15. The system of claim 1 , wherein the first language signal comprises a voice signal.

16. The system of claim 1 , wherein the activated language model comprises an “n-best list”.

17. The system of claim 1 , wherein the activated language model comprises a reject list of at least some recognition results.

18. The system of claim 1 , wherein the activated language model comprises a domain-specific language model.

19. The system of claim 18 , wherein the domain-specific language model comprises a user domain selection.

20. The system of claim 1 , wherein the activated language model comprises an empirically derived model.

21. The system of claim 1 , wherein the interpretation engine is further configured to update the utterance probabilities of the language model based on a detected move of the device location.

Assignments (3)
MERGER Recorded Feb 4, 2016
From: FLUENTIAL, INC.
To: FLUENTIAL,LLC
Reel/Frame 037662/0691 →
NUNC PRO TUNC ASSIGNMENT Recorded Feb 4, 2016
From: FLUENTIAL, LLC
To: NANT HOLDINGS IP, LLC
Reel/Frame 037670/0024 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 2, 2016
From: EHSANI, FARZAD; MASTER, DEMITIRIOS; ZUBER, ELAINE
To: FLUENTIAL, INC.
Reel/Frame 037650/0237 →
Continuity (5)
Continuation 14326283 · Jul 8, 2014
Continuation 13934194 · Jul 2, 2013
Continuation 12351793 · Jan 9, 2009
Provisional Application 61020112 · Jan 9, 2008
Related Publication 20160103825A1 · Apr 14, 2016