IP Library Granted Patent US 10,157,040
Granted Patent B2
US 10,157,040 · App. 14/988,408 · Granted Dec 18, 2018

Multi-modal input on an electronic device

Inventors: Brandon M. Ballinger (San Francisco, CA); Johan Schalkwyk (Scarsdale, NY); Michael H. Cohen (Portola Valley, CA); William J. Byrne (Davis, CA); Gudmundur Hafsteinsson (Los Gatos, CA); Michael J. LeBeau (New York, NY)
Assignee: Google LLC
G06F3/167G06F3/04886G06F17/277G06F17/289G10L15/005G10L15/18G10L15/183G10L15/22G10L15/26G10L15/265G10L15/30G10L15/197G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,157,040
App. No.
14/988,408
Filed
Jan 5, 2016
Granted
Dec 18, 2018
Kind
B2
Art Unit
2658
USPC
704/8
Abstract

A computer-implemented input-method editor process includes receiving a request from a user for an application-independent input method editor having written and spoken input capabilities, identifying that the user is about to provide spoken input to the application-independent input method editor, and receiving a spoken input from the user. The spoken input corresponds to input to an application and is converted to text that represents the spoken input. The text is provided as input to the application.

Claims (58)

1. A computer-implemented method, comprising:

displaying, on an electronic display of a computing device, a user interface that includes a voice input control and a virtual keyboard, wherein the virtual keyboard includes a plurality of character entry keys displayed within a first region of the user interface and the voice input control comprises a graphical representation of a microphone;

in response to receiving data that indicates user interaction with the voice input control of the user interface:

(i) removing the virtual keyboard from the user interface so as to disable an ability of the computing device to receive typed text input; and

(ii) enabling a voice input mode of the computing device, including displaying, in place of the virtual keyboard, a visual indication that the computing device is enabled to receive voice input; and

detecting, by the computing device, an utterance as a result of the computing device being enabled to receive voice input.

2. The computer-implemented method of claim 1 , further comprising receiving data that indicates user interaction with a text field displayed on the electronic display of the computing device,

wherein the user interface that includes the voice input control and the virtual keyboard is displayed in response to receiving the data that indicates user interaction with the text field.

3. The computer-implemented method of claim 2 , wherein the user interface is a user interface for a multi-modal input method editor that enables the computing device to receive voice input and typed input.

4. The computer-implemented method of claim 3 , further comprising:

displaying, on the electronic display of the computing device and before receiving the data that indicates user interaction with the text field, a user interface for an application that provides the text field; and

in response to receiving the data that indicates user interaction with the text field, displaying the user interface for the multi-modal input method editor, including the voice input control and the virtual keyboard, over a first portion of the user interface for the application that provides the text field, such that the first portion of the user interface for the application is not visible while the user interface for the multi-modal input method editor is displayed,

wherein a second portion of the user interface for the application remains visible while the user interface for the multi-modal input method editor is displayed, wherein the second portion of the user interface for the application includes the text field.

5. The computer-implemented method of claim 4 , further comprising displaying the user interface for the multi-modal input method editor on the electronic display of the computing device below the second portion of the user interface for the application.

6. The computer-implemented method of claim 2 , wherein the text field is provided in a web page presented by the computing device.

7. The computer-implemented method of claim 1 , further comprising, in response to receiving the data that indicates user interaction with the voice input control of the user interface, removing the voice input control from display in the user interface.

8. The computer-implemented method of claim 1 , wherein displaying the indication that the computing device is enabled to receive voice input comprises displaying a second graphical representation of a microphone within the first region of the user interface, wherein at least a location or size of the second graphical representation of the microphone is different from a location or size of the graphical representation of the microphone for the voice input control.

9. The computer-implemented method of claim 1 , further comprising:

in response to receiving the data that indicates user interaction with the voice input control of the user interface, displaying, within the first region of the user interface, a voice input cancellation control; and

in response to receiving data that indicates user interaction with the voice input cancellation control, disabling a capability of the computing device to receive voice input.

10. The computer-implemented method of claim 1 , wherein the user interface is a user interface for an application-independent input method editor that is enabled to receive input directed to multiple different applications on the computing device.

11. The computer-implemented method of claim 1 , wherein the virtual keyboard is a QWERTY keyboard.

12. One or more non-transitory computer-readable media having instructions stored thereon that, when executed by one or more processors, cause performance of operations comprising:

displaying, on an electronic display of a computing device, a user interface that includes a voice input control and a virtual keyboard, wherein the virtual keyboard comprises a plurality of keys including character entry keys, wherein the plurality of keys are arranged in a plurality of rows that each includes a respective subset of the plurality of keys of the virtual keyboard, wherein a particular row of the plurality of rows includes the voice input control and at least one key of the plurality of keys of the virtual keyboard;

in response to receiving data that indicates user interaction with the voice input control of the user interface:

(i) removing the virtual keyboard from the user interface so as to disable an ability of the computing device to receive typed text input; and

(ii) enabling a voice input mode of the computing device, including displaying, in place of the virtual keyboard, a visual indication that the computing device is enabled to receive voice input; and

detecting, by the computing device, an utterance as a result of the computing device being enabled to receive voice input.

13. A computing device, comprising:

an electronic display;

one or more processors; and

one or more computer-readable media having instructions stored thereon that, when executed by the one or more processors, cause performance of operations comprising:

displaying, on the electronic display of the computing device, a user interface that includes a voice input control and a virtual keyboard, wherein the virtual keyboard includes a plurality of character entry keys displayed within a first region of the user interface;

in response to receiving data that indicates user interaction with the voice input control of the user interface that includes the voice input control and the virtual keyboard:

(i) enabling the computing device to receive voice input, including transitioning the computing device from a text entry mode to a voice input mode where the device is enabled to receive voice input rather than typed text input, and

(ii) replacing the virtual keyboard in the user interface with a visual indication that the computing device is enabled to receive voice input rather than typed text input, including displaying the visual indication in place of the plurality of character entry keys within the first region of the user interface while the computing device is enabled to receive voice input; and

detecting, by the computing device, an utterance as a result of the computing device being enabled to receive voice input.

14. A computer-implemented method, comprising:

displaying, on an electronic display of a computing device, a user interface that includes a voice input control and a virtual keyboard, wherein the virtual keyboard includes a plurality of character entry keys displayed within a first region of the user interface;

in response to receiving data that indicates user interaction with the voice input control of the user interface that includes the voice input control and the virtual keyboard:

(i) enabling the computing device to receive voice input, including transitioning the computing device from a text entry mode to a voice input mode where the device is enabled to receive voice input rather than typed text input, and

(ii) replacing the virtual keyboard in the user interface with a visual indication that the computing device is enabled to receive voice input rather than typed text input, including displaying the visual indication in place of the plurality of character entry keys within the first region of the user interface while the computing device is enabled to receive voice input;

detecting, by the computing device, an utterance as a result of the computing device being enabled to receive voice input;

providing, by the computing device and to a server system, audio data that corresponds to the utterance; and

providing, by the computing device and for display in the text field, a transcription of the utterance received from the server system.

15. The computer-implemented method of claim 14 , further comprising:

determining a context associated with the utterance;

providing, by the computing device and to the server system, data that indicates the context associated with the utterance; and

receiving the transcription at the computing device and from the server system, wherein the transcription was generated by the server system based at least in part on the data that indicates the context associated with the utterance.

16. The computer-implemented method of claim 14 , further comprising:

receiving, by the computing device and from the server system, a plurality of candidate transcriptions of the utterance;

in response to receiving the plurality of candidate transcriptions of the utterance, displaying on the electronic display of the computing device at least some of the plurality of candidate transcriptions; and

identifying that user input selected a first candidate transcription among the at least some of the plurality of candidate transcriptions,

wherein providing the transcription of the utterance for display in the text field comprises providing the first candidate transcription for display in the text field based on having identified that the user input selected the first candidate transcription from among the at least some of the plurality of candidate transcriptions.

17. The computer-implemented method of claim 14 , further comprising:

after providing to the server system the audio data that corresponds to the utterance, displaying within the first region of the user interface an indication that a transcription process is in progress to generate a transcription of the utterance;

receiving, by the computing device and from the server system, the transcription of the utterance; and

in response to receiving the transcription of the utterance from the server system, ceasing to display the indication that the transcription process is in progress.

Assignments (2)
CHANGE OF NAME Recorded Oct 5, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044129/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 11, 2016
From: BYRNE, WILLIAM J.; BALLINGER, BRANDON M.; SCHALKWYK, JOHAN; COHEN, MICHAEL H.; HAFSTEINSSON, GUDMUNDUR; LEBEAU, MICHAEL J.
To: GOOGLE INC.
Reel/Frame 037709/0039 →
Continuity (6)
Continuation 14299837 · Jun 9, 2014
Continuation 13249172 · Sep 29, 2011
Continuation 12977003 · Dec 22, 2010
Provisional Application 61330219 · Apr 30, 2010
Provisional Application 61289968 · Dec 23, 2009
Related Publication 20160132293A1 · May 12, 2016
Cited By (35)
US 12,197,699 US 12,210,730 US 12,223,228 US 12,236,952 US 12,242,702 US 12,244,755 US 12,256,128 US 12,260,059 US 12,262,089 US 12,265,364 US 12,265,696 US 12,267,622 US 12,301,979 US 12,302,035 US 12,348,663 US 12,368,946 US 12,379,827 US 12,381,924 US 12,386,585 US 12,422,976 US 12,449,961 US 12,452,389 US 12,504,944 US 12,526,361 US 12,541,338 US 12,563,299 US 12,578,200 US 12,578,837 US 12,591,329 US 12,615,491 US 12,620,155 US 12,676,927 US 12,681,587 US 12,732,665 US 12,737,049