IP Library Granted Patent US 10,490,190
Granted Patent B2
US 10,490,190 · App. 16/203,521 · Granted Nov 26, 2019

Task initiation using sensor dependent context long-tail voice commands

Inventors: Yuzhao Ni (London, GB); Bo Wang (San Jose, CA); Barnaby James (Los Gatos, CA); Pravir Gupta (Los Gatos, CA); David Schairer (San Jose, CA)
Assignee: GOOGLE LLC
G10L15/22G10L15/063G10L15/1815G10L15/1822G06F3/167G06F16/243G06F16/3338G06F17/289G10L15/183G10L2015/088G10L2015/223G10L2015/225
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,490,190
App. No.
16/203,521
Granted
Nov 26, 2019
Kind
B2
Abstract

In various implementations, upon receiving a given voice command from a user, a voice-based trigger may be selected from a library of voice-based triggers previously used across a population of users. The library may include association(s) between each voice-based trigger and responsive action(s) previously performed in response to the voice-based trigger. The selecting may be based on a measure of similarity between the given voice command and the selected voice-based trigger. One or more responsive actions associated with the selected voice-based trigger in the library may be determined. Based on the one or more responsive actions, current responsive action(s) may be performed by a target client device selected based on sensor-dependent context. Feedback associated with performance of the current responsive action(s) may be received from the user and used to alter a strength of an association between the selected voice-based trigger and the one or more responsive actions.

Claims (47)

1. A method comprising:

receiving, at a first client device, a given voice command from a user, wherein the given voice command is ambiguous as to what responsive action it is meant to invoke;

disambiguating the given voice command to identify a given responsive action to perform, wherein the disambiguating comprises:

identifying a context of the user detected using one or more sensors of the first client device or another client device;

selecting, from a library of voice-based triggers previously used across a population of users, a selected voice-based trigger, wherein the library includes one or more associations between each respective voice-based trigger of the library of voice-based triggers and one or more responsive actions previously invoked automatically at one or more other client devices operated by one or more other users of the population of users in response to the respective voice-based trigger, and wherein the selecting is based at least in part on a measure of similarity between the given voice command and the selected voice-based trigger;

determining a plurality of candidate responsive actions associated with the selected voice-based trigger in the library, wherein each candidate responsive action of the plurality of candidate responsive actions is associated with a different client device of a plurality of client devices that are controllable by the user;

based on the context of the user, selecting a target client device from the plurality of client devices; and

based on the selected target device, selecting the given responsive action from the plurality of candidate responsive actions associated with the selected voice-based trigger in the library; and

performing, at the target client device, the given responsive action.

2. The method of claim 1 , wherein the selected voice-based trigger includes one or more terms or tokens that are analogous to one or more terms or tokens in the given voice command.

3. The method of claim 1 , further comprising:

receiving, from the user, feedback associated with the performance of the given responsive action; and

altering a strength of an association between the selected voice-based trigger and the given responsive action based on the feedback;

wherein the altering comprises weakening or strengthening the association between the selected voice-based trigger and the given responsive action based on the feedback.

4. The method of claim 1 , wherein the plurality of candidate responsive actions include altering an air conditioning parameter in a building and altering an air conditioning parameter in a vehicle, and the context of the user indicates whether the user is in the building or in the vehicle.

5. The method of claim 1 , wherein the target client device is the first client device.

6. The method of claim 1 , wherein the target client device is different than the first client device.

7. The method of claim 1 , wherein the measure of similarity is a measure of syntactic or semantic similarity.

8. A system comprising one or more processors and memory operably coupled with the one or more processors, wherein the memory stores instructions that, in response to execution of the instructions by the one or more processors, cause the one or more processors to operate an interactive assistant module configured to perform the following operations:

receiving, at a first client device, a given voice command from a user, wherein the given voice command is ambiguous as to what responsive action it is meant to invoke;

disambiguating the given voice command to identify a given responsive action to perform, wherein the disambiguating comprises:

identifying a context of the user detected using one or more sensors of the first client device or another client device;

selecting, from a library of voice-based triggers previously used across a population of users, a selected voice-based trigger, wherein the library includes one or more associations between each respective voice-based trigger of the library of voice-based triggers and one or more responsive actions previously invoked automatically at one or more other client devices operated by one or more other users of the population of users in response to the respective voice-based trigger, and wherein the selecting is based at least in part on a measure of similarity between the given voice command and the selected voice-based trigger;

determining a plurality of candidate responsive actions associated with the selected voice-based trigger in the library, wherein each candidate responsive action of the plurality of candidate responsive actions is associated with a different client device of a plurality of client devices that are controllable by the user;

based on the context of the user, selecting a target client device from the plurality of client devices;

based on the selected target device, selecting the given responsive action from the plurality of candidate responsive actions associated with the selected voice-based trigger in the library; and

performing, at a target client device controlled by the user, the given responsive action.

9. The system of claim 8 , wherein the selected voice-based trigger includes one or more terms or tokens that are analogous to one or more terms or tokens in the given voice command.

10. The system of claim 8 , further comprising:

receiving, from the user, feedback associated with the performance of the given responsive action; and

altering a strength of an association between the selected voice-based trigger and the given responsive action based on the feedback;

wherein the altering comprises weakening or strengthening the association between the selected voice-based trigger and the given responsive action based on the feedback.

11. The system of claim 8 , wherein the plurality of responsive actions include altering an air conditioning parameter in a building and altering an air conditioning parameter in a vehicle, and the context of the user indicates whether the user is in the building or in the vehicle.

12. The system of claim 8 , wherein the target client device is the first client device.

13. The system of claim 8 , wherein the target client device is different than the first client device.

14. The system of claim 8 , wherein the measure of similarity is a measure of syntactic or semantic similarity.

15. At least one non-transitory computer-readable medium comprising instructions that, in response to execution of the instructions by one or more processors, cause the one or more processors to perform the following operations:

receiving, at a first client device, a given voice command from a user, wherein the given voice command is ambiguous as to what responsive action it is meant to invoke;

disambiguating the given voice command to identify a given responsive action to perform, wherein the disambiguating comprises:

identifying a context of the user detected using one or more sensors of the first client device or another client device;

selecting, from a library of voice-based triggers previously used across a population of users, a selected voice-based trigger, wherein the library includes one or more associations between each respective voice-based trigger of the library of voice-based triggers and one or more responsive actions previously invoked automatically at one or more other client devices operated by one or more other users of the population of users in response to the respective voice-based trigger, and wherein the selecting is based at least in part on a measure of similarity between the given voice command and the selected voice-based trigger;

determining a plurality of candidate responsive actions associated with the selected voice-based trigger in the library, wherein each candidate responsive action of the plurality of candidate responsive actions is associated with a different client device of a plurality of client devices that are controllable by the user;

based on the context of the user, selecting a target client device from the plurality of client devices;

based on the selected target device, selecting the given responsive action from the plurality of candidate responsive actions associated with the selected voice-based trigger in the library; and

performing, at a target client device controlled by the user, the given responsive action.

16. The at least one non-transitory computer-readable medium of claim 15 , wherein the plurality of candidate responsive actions include altering an air conditioning parameter in a building and altering an air conditioning parameter in a vehicle, and the context of the user indicates whether the user is in the building or in the vehicle.

17. The at least one non-transitory computer-readable medium of claim 15 , wherein the target client device is the first client device.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 2, 2019
From: NI, YUZHAO; WANG, BO; JAMES, BARNABY; GUPTA, PRAVIR; SCHAIRER, DAVID
To: GOOGLE INC.
Reel/Frame 048772/0402 →
CHANGE OF NAME Recorded Apr 2, 2019
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 048773/0530 →
Continuity (2)
Continuation 15284473 · Oct 3, 2016
Related Publication 20190096406A1 · Mar 28, 2019
Cited By (1)
US 12,315,510