IP Library Granted Patent US 10,839,799
Granted Patent B2
US 10,839,799 · App. 15/987,509 · Granted Nov 17, 2020

Developer voice actions system

Inventors: Bo Wang (San Jose, CA); Sunil Vemuri (Pleasanton, CA); Nitin Mangesh Shetti (Sunnyvale, CA); Pravir Kumar Gupta (Los Altos, CA); Scott B Huffman (Portola Valley, CA); Javier Alejandro Rey (San Francisco, CA); Jeffrey A. Boortz (San Francisco, CA)
Assignee: GOOGLE LLC
G10L15/22G06F3/167G10L15/1815G10L15/19G10L2015/0638G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,839,799
App. No.
15/987,509
Granted
Nov 17, 2020
Kind
B2
Abstract

Methods, systems, and apparatus for receiving data identifying an application and a voice command trigger term, validating the received data, inducting the received data to generate an intent that specifies the application, the voice command trigger term, and one or more other voice command trigger terms that are determined based at least on the voice command trigger term, and storing the intent at a contextual intent database, wherein the contextual intent database comprises one or more other intents.

Claims (41)

1. A computer-implemented method, comprising:

receiving, by a voice action service system, a spoken utterance provided at a computing device of a user, the spoken utterance including a voice command trigger phrase;

processing, by the voice action service system, the spoken utterance to determine an intent associated with the voice command trigger phrase, wherein the intent specifies an activity to be performed via the computing device of the user;

identifying, by the voice action service system, at least a first application and a second application that are each capable of satisfying the intent, wherein identifying at least the first application and the second application is based on determining at least the first application and the second application are associated with the intent in one or more databases;

selecting, by the voice action service system, the first application over the second application,

wherein selecting the first application over the second application is based at least in part on recency of usage of the first application by the user for the activity specified by the intent, the usage being a past usage of the first application by the user for the activity specified by the intent, the past usage being prior to receiving the spoken utterance including the voice command trigger phrase, and

wherein selecting the first application over the second application based at least in part on the recency of usage of the first application by the user comprises determining that the first application was most recently selected, by the user, in response to receiving a prior spoken utterance including the voice command trigger phrase, the prior spoken utterance being received prior to receiving the spoken utterance including the voice command trigger phrase;

providing, to the computing device of the user by the voice action service system and responsive to the spoken utterance, an indication of the selected first application;

receiving, by the voice action service system, an additional spoken utterance at the computing device of the user, the additional spoken utterance including a confirmation of the selected first application; and

in response to receiving the additional spoken utterance, executing the first application to satisfy the intent, wherein executing the first application to satisfy the intent comprises causing the selected first application to perform the activity specified by the intent.

2. The computer-implemented method of claim 1 , wherein selecting the first application over the second application is further based at least in part on a strength of relationship score between the first application and at least one of: the voice command trigger phrase and the intent.

3. The computer-implemented method of claim 1 , wherein selecting the first application over the second application is further based at least in part on the first application being executed on the computing device of the user when the spoken utterance is received.

4. The computer-implemented method of claim 1 , wherein providing the indication of the selected first application comprises providing an audible indication of the selected first application.

5. The computer-implemented method of claim 1 , wherein processing the spoken utterance to determine the intent comprises:

performing, by the voice action service system, speech recognition on the spoken utterance to obtain a transcription of the spoken utterance; and

determining, by the voice action service system, at least a portion of the transcription includes the voice command trigger phrase and that the voice command trigger phrase matches the intent.

6. The method of claim 1 , wherein the activity specified by the intent is a ride booking activity.

7. A system, comprising:

at least one processor; and

at least one memory comprising instructions that when executed, cause the at least one processor to:

receive a spoken utterance provided at a computing device of a user, the spoken utterance including a voice command trigger phrase;

process the spoken utterance to determine the spoken utterance includes the voice command trigger phrase;

identify at least a first application and a second application, wherein identifying at least the first application and the second application is based on determining at least the first application and the second application are mapped to the voice command trigger phrase in one or more databases;

select the first application over the second application,

wherein selecting the first application over the second application is based at least in part on recency of usage of the first application by the user, the usage being a past usage prior to receiving the spoken utterance including the voice command trigger phrase, the past usage being prior to receiving the spoken utterance including the voice command trigger phrase, and

wherein the instructions to select the first application over the second application based at least in part on the recency of usage of the first application by the user comprise instructions to determine that the first application was most recently selected, by the user, in response to receiving a prior spoken utterance including the voice command trigger phrase, the prior spoken utterance being received prior to receiving the spoken utterance including the voice command trigger phrase;

provide, to the computing device of the user and responsive to the spoken utterance, an indication of the selected first application; and

receive an additional spoken utterance at the computing device of the user, the additional spoken utterance including a confirmation of the selected first application; and

in response to receiving the additional spoken utterance, execute the selected first application to satisfy the intent, wherein the instructions to execute the selected first application to satisfy the intent comprise instructions to cause the selected first application to perform the activity specified by the intent.

8. The system of claim 7 , wherein the instructions to select the first application over the second application further comprise instructions to select the first application over the second application based at least in part on a strength of relationship score between the first application and the voice command trigger phrase.

9. The system of claim 7 , wherein the instructions to select the first application over the second application further comprise instructions to select the first application over the second application based at least in part on the first application being executed on the computing device of the user when the spoken utterance is received.

10. The system of claim 7 , wherein the instructions to provide the indication of the selected first application further comprise instructions to provide an audible indication of the selected first application.

11. A non-transitory computer-readable storage medium comprising instructions that, when executed, cause at least one processor to:

receive a spoken utterance provided at a computing device of a user, the spoken utterance including a voice command trigger phrase;

process the spoken utterance to determine an intent associated with the voice command trigger phrase, wherein the intent specifies an activity to be performed via the computing device of the user;

identify at least a first application and a second application that are each capable of satisfying the intent, wherein identifying at least the first application and the second application is based on determining at least the first application and the second application are associated with the intent in one or more databases;

select the first application over the second application,

wherein selecting the first application over the second application is based at least in part on recency of usage of the first application by the user for the activity specified by the intent, the usage being a past usage of the first application by the user for the activity specified by the intent, the past usage being prior to receiving the spoken utterance including the voice command trigger phrase, and

wherein the instructions to select the first application over the second application based at least in part on the recency of usage of the first application by the user further cause the at least one processor to determine that the first application was most recently selected, by the user, in response to receiving a prior spoken utterance including the voice command trigger phrase, the prior spoken utterance being received prior to receiving the spoken utterance including the voice command trigger phrase; and

provide, to the computing device of the user and responsive to the spoken utterance, an indication of the selected first application; and

subsequent to providing the indication of the selected first application, and in response to receiving the spoken utterance, execute the first application to satisfy the intent without requiring any additional user input from the user, wherein the instructions to execute the selected first application to satisfy the intent comprise instructions to cause the selected first application to perform the activity specified by the intent.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2020
From: WANG, BO; VEMURI, SUNIL; SHETTI, NITIN MANGESH; GUPTA, PRAVIR KUMAR; HUFFMAN, SCOTT B.; REY, JAVIER ALEJANDRO; BOORTZ, JEFFREY A.
To: GOOGLE INC.
Reel/Frame 053928/0618 →
CHANGE OF NAME Recorded Sep 30, 2020
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 053932/0620 →
Continuity (3)
Continuation 15258084 · Sep 7, 2016
Continuation 14693330 · Apr 22, 2015
Related Publication 20180374480A1 · Dec 27, 2018
Cited By (1)
US 12,451,127