IP Library Granted Patent US 11,017,769
Granted Patent B2
US 11,017,769 · App. 16/374,343 · Granted May 25, 2021

Customized voice action system

Inventor: Pedro J. Moreno Mengibar (Jersey City, NJ)
Assignee: GOOGLE LLC
G10L15/22G06F40/242G06Q30/0256G06Q30/0275G06Q30/0277G06Q30/08G10L15/08G10L15/197G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,017,769
App. No.
16/374,343
Granted
May 25, 2021
Kind
B2
Abstract

Systems, methods, and computer-readable media that may be used to modify a voice action system to include voice actions provided by advertisers or users are provided. One method includes receiving electronic voice action bids from advertisers to modify the voice action system to include a specific voice action (e.g., a triggering phrase and an action). One or more bids may be selected. The method includes, for each of the selected bids, modifying data associated with the voice action system to include the voice action associated with the bid, such that the action associated with the respective voice action is performed when voice input from a user is received that the voice action system determines to correspond to the triggering phrase associated with the respective voice action.

Claims (52)

1. A system to command a computing device, comprising:

one or more servers comprising one or more processors and memory;

a voice action system executed by the one or more servers to:

receive, via a network from a first provider device, a first voice action that includes a first plurality of triggering phrases and a first action;

identify a second voice action that includes a second triggering phrase for a second action different from the first action, the second triggering phrase different from each of the first plurality of triggering phrases;

select, based on a predetermined criteria, the first voice action from a set comprising the first voice action and the second voice action;

increase, via modification of a voice action language model and responsive to the selection of the first voice action based on the predetermined criteria, a likelihood that a voice input matches at least one of the first plurality of triggering phrases relative to the second triggering phrase; and

configure the first voice action selected based on the predetermined criteria to perform, via the computing device, the first action of the first voice action responsive to the voice input matching the at least one of the first plurality of triggering phrases of the first voice action.

2. The system of claim 1 , wherein the second action is a default action.

3. The system of claim 1 , wherein the second action is performed responsive to voice input not matching the first plurality of triggering phrases.

4. The system of claim 1 , comprising:

the voice action system to receive, from the computing device comprising a digital assistant, the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action; and

the voice action system to cause, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action, the computing device to perform the first action.

5. The system of claim 1 , comprising:

the voice action system to cause the computing device to perform the first action comprising a series of actions.

6. The system of claim 1 , comprising:

the voice action system to cause the computing device to perform the first action comprising searching a directory, identifying a match in the directory, initiating a call to a phone number corresponding to the match, or presenting a plurality of options to the computing device.

7. The system of claim 1 , comprising:

the voice action system to receive from a digital assistant, the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action; and

the voice action system to cause, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action, the computing device to perform the first action comprising displaying a website in a browser application.

8. The system of claim 1 , comprising:

the voice action system to receive from a digital assistant, second voice input that does not correspond to any of the first plurality of triggering phrases; and

the voice action system to cause, responsive to receipt of the second voice input that does not correspond to any of the first plurality of triggering phrases, the computing device to perform the second action comprising a default action.

9. The system of claim 1 , comprising:

the voice action system to configure the first voice action to cause the computing device to perform the first action comprising at least two of displaying a website in a browser application, calling a telephone number, displaying a location via a mapping application, downloading a data file, presenting media content, sending an electronic message to a recipient device, performing a search via a search engine, or launching a chat application.

10. The system of claim 1 , comprising the voice action system to:

cause, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action based on the likelihood, the computing device to display a location via a mapping application executed by the computing device.

11. A method of commanding a computing device, comprising:

receiving, by one or more servers via a network from a first provider device, a first voice action that includes a first plurality of triggering phrases and a first action;

identifying, by the one or more servers, a second voice action that includes a second triggering phrase for a second action different from the first action, the second triggering phrase different from each of the first plurality of triggering phrases;

selecting, based on a predetermined criteria, the first voice action from a set comprising the first voice action and the second voice action;

increasing, by the one or more servers and responsive to the selection of the first voice action based on the predetermined criteria, via modification of a voice action language model, a likelihood that a voice input matches at least one of the first plurality of triggering phrases relative to the second triggering phrase; and

configuring, by the one or more servers, the first voice action selected based on the predetermined criteria to perform, via the computing device, the first action of the first voice action responsive to the voice input matching the at least one of the first plurality of triggering phrases of the first voice action.

12. The method of claim 11 , wherein the second action is a default action.

13. The method of claim 11 , wherein the second action is performed responsive to voice input not matching the first plurality of triggering phrases.

14. The method of claim 11 , comprising:

receiving, from the computing device comprising a digital assistant, the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action; and

causing, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action, the computing device to perform the first action.

15. The method of claim 11 , comprising:

causing the computing device to perform the first action comprising a series of actions.

16. The method of claim 11 , comprising:

causing the computing device to perform the first action comprising searching a directory, identifying a match in the directory, initiating a call to a phone number corresponding to the match, or presenting a plurality of options to the computing device.

17. The method of claim 11 , comprising:

receiving from a digital assistant, the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action; and

causing, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action, the computing device to perform the first action comprising displaying a website in a browser application.

18. The method of claim 11 , comprising:

receiving from a digital assistant, second voice input that does not correspond to any of the first plurality of triggering phrases; and

causing, responsive to receipt of the second voice input that does not correspond to any of the first plurality of triggering phrases, the computing device to perform the second action comprising a default action.

19. The method of claim 11 , comprising:

configuring the first voice action to cause the computing device to perform the first action comprising at least two of displaying a website in a browser application, calling a telephone number, displaying a location via a mapping application, downloading a data file, presenting media content, sending an electronic message to a recipient device, performing a search via a search engine, or launching a chat application.

20. The method of claim 11 , comprising:

causing, responsive to receipt of the voice input corresponding to the at least one of the first plurality of triggering phrases of the first voice action based on the likelihood, the computing device to display a location via a mapping application executed by the computing device.

Assignments (2)
CHANGE OF NAME Recorded Apr 4, 2019
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 050245/0004 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 4, 2019
From: MENGIBAR, PEDRO J. MORENO
To: GOOGLE INC.
Reel/Frame 048797/0377 →
Continuity (4)
Continuation 15638285 · Jun 29, 2017
Continuation 15054301 · Feb 26, 2016
Continuation 13478803 · May 23, 2012
Related Publication 20190228771A1 · Jul 25, 2019