IP Library › Granted Patent US 9,600,474
Granted Patent B2
US 9,600,474 · App. 14/075,018 · Granted Mar 21, 2017

User interface for realtime language translation

Inventors: Alexander J. Cuthbert (Oakland, CA); Sunny Goyal (Mountain View, CA); Matthew Gaba (San Francisco, CA); Joshua J. Estelle (San Francisco, CA); Masakazu Seno (Cupertino, CA)
Assignee: Google Inc.
G06F17/289G10L25/48G10L15/26
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,600,474
App. No.
14/075,018
Granted
Mar 21, 2017
Kind
B2
Abstract

A language translation application on a user device includes a user interface that provides relevant textual and graphical feedback mechanisms associated with various states of voice input and translated speech.

Claims (39)

1. A computer-implemented method comprising:

displaying a graphical user interface for a language translation application on a user device, the graphical user interface comprising a first graphical representation identifying a source language, a second graphical representation identifying a target language, and a third graphical representation indicating the user device operating in a listening mode that is arranged adjacent to both the first graphical representation identifying the source language and the second graphical representation identifying the target language;

animating, in response to a request to initiate listening for an utterance in the source language, the third graphical representation indicating the listening mode while the language translation application prepares to listen for the source language;

highlighting, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode while also highlighting the first graphical representation identifying the source language to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode indicating that the language translation application is prepared to receive voice input in the source language;

receiving an utterance spoken in the source language for translation into the target language while the third graphical representation indicating the listening mode and the first graphical representation identifying the source language are highlighted to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode;

replacing, in response to the language translation application preparing an output of a translation of the utterance into the target language, the third graphical representation indicating the listening mode on the graphical user interface with a fourth graphical representation indicating a translation transcription mode on the graphical user interface; and

highlighting, in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode while also highlighting the second graphical representation identifying the target language to create a visual correspondence between the second graphical representation identifying the target language and the fourth graphical representation indicating the translation transcription mode indicating that the language translation application is outputting a translation of the utterance in the target language.

2. The method of claim 1 , further comprising, animating, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode.

3. The method of claim 2 , wherein animating the third graphical representation indicating the listening mode comprises animating a graphical representation of a microphone while a microphone of the user device is receiving an audio signal.

4. The method of claim 1 , further comprising animating in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode.

5. The method of claim 1 , further comprising animating, in response to a request to initiate listening for an utterance in the source language, the first graphical representation identifying the source language.

6. The method of claim 1 , wherein the third graphical representation indicating the listening mode comprises a microphone icon.

7. The method of claim 1 , wherein the fourth graphical representation indicating the translation transcription mode comprises a speaker icon.

8. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

displaying a graphical user interface for a language translation application on a user device, the graphical user interface comprising a first graphical representation identifying a source language, a second graphical representation identifying a target language, and a third graphical representation indicating the user device operating in a listening mode that is arranged adjacent to both the first graphical representation identifying the source language and the second graphical representation identifying the target language;

animating, in response to a request to initiate listening for an utterance in the source language, the third graphical representation indicating the listening mode while the language translation application prepares to listen for the source language;

highlighting, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode while also highlighting the first graphical representation identifying the source language to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode indicating that the language translation application is prepared to receive voice input in the source language;

receiving an utterance spoken in the source language for translation into the target language while the third graphical representation indicating the listening mode and the first graphical representation identifying the source language are highlighted to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode;

replacing, in response to the language translation application preparing an output of a translation of the utterance into the target language, the third graphical representation indicating the listening mode on the graphical user interface with a fourth graphical representation indicating a translation transcription mode on the graphical user interface; and

highlighting, in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode while also highlighting the second graphical representation identifying the target language to create a visual correspondence between the second graphical representation identifying the target language and the fourth graphical representation indicating the translation transcription mode indicating that the language translation application is outputting a translation of the utterance in the target language.

9. The system of claim 8 , wherein the operations further comprise, animating, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode.

10. The system of claim 9 , wherein animating the third graphical representation indicating the listening mode comprises animating a graphical representation of a microphone while a microphone of the user device is receiving an audio signal.

11. The system of claim 8 , wherein the operations further comprise animating in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode.

12. The system of claim 8 , wherein the operations further comprise animating, in response to a request to initiate listening for an utterance in the source language, the first graphical representation identifying the source language.

13. The system of claim 8 , wherein the third graphical representation indicating the listening mode comprises a microphone icon.

14. The system of claim 8 , wherein the fourth graphical representation indicating the translation transcription mode comprises a speaker icon.

15. A non-transitory computer-readable medium storing software comprising instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:

displaying a graphical user interface for a language translation application on a user device, the graphical user interface comprising a first graphical representation identifying a source language, a second graphical representation identifying a target language, and a third graphical representation indicating the user device operating in a listening mode that is arranged adjacent to both the first graphical representation identifying the source language and the second graphical representation identifying the target language;

animating, in response to a request to initiate listening for an utterance in the source language, the third graphical representation indicating the listening mode while the language translation application prepares to listen for the source language;

highlighting, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode while also highlighting the first graphical representation identifying the source language to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode indicating that the language translation application is prepared to receive voice input in the source language;

receiving an utterance spoken in the source language for translation into the target language while the third graphical representation indicating the listening mode and the first graphical representation identifying the source language are highlighted to create a visual correspondence between the first graphical representation identifying the source language and the third graphical representation indicating the listening mode;

replacing, in response to the language translation application preparing an output of a translation of the utterance into the target language, the third graphical representation indicating the listening mode on the graphical user interface with a fourth graphical representation indicating a translation transcription mode on the graphical user interface; and

highlighting, in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode while also highlighting the second graphical representation identifying the target language to create a visual correspondence between the second graphical representation identifying the target language and the fourth graphical representation indicating the translation transcription mode indicating that the language translation application is outputting a translation of the utterance in the target language.

16. The computer-readable medium of claim 15 , wherein the operations further comprise, animating, in response to the language translation application completing preparations to listen for the source language, the third graphical representation indicating the listening mode.

17. The computer-readable medium of claim 16 , wherein animating the third graphical representation indicating the listening mode comprises animating a graphical representation of a microphone while a microphone of the user device is receiving an audio signal.

18. The computer-readable medium of claim 15 , wherein the operations further comprise animating in response to the language translation application completing preparations to output the translation of the transcription into the target language, the fourth graphical representation indicating the translation transcription mode.

19. The computer-readable medium of claim 15 , wherein the operations further comprise animating, in response to a request to initiate listening for an utterance in the source language, the first graphical representation identifying the source language.

20. The computer-readable medium of claim 15 , wherein the third graphical representation indicating the listening mode comprises a microphone icon.

Assignments (2)
CHANGE OF NAME Recorded Oct 2, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044097/0658 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 27, 2013
From: CUTHBERT, ALEXANDER J.; GOYAL, SUNNY; GABA, MATTHEW; ESTELLE, JOSHUA J.; SENO, MASAKAZU
To: GOOGLE INC.
Reel/Frame 031687/0483 →
Continuity (1)
Related Publication 20150134322A1 · May 14, 2015