IP Library › Granted Patent US 10,860,639
Granted Patent B2
US 10,860,639 · App. 15/267,463 · Granted Dec 8, 2020

Query response using media consumption history

Inventor: Matthew Sharifi (Kilchberg, CH)
Assignee: Google LLC
G06F16/487G06F16/245G06F16/2455G06F16/24578G06F16/433G06F16/435G06F16/437G06F16/489G06F16/685G06F16/7834G06F16/955G06F16/9535G06Q30/02G06Q30/0631
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,860,639
App. No.
15/267,463
Granted
Dec 8, 2020
Kind
B2
Abstract

Methods, systems, and apparatus for receiving a natural language query of a user, and environmental data, identifying a media item based on the environmental data, determining an entity type based on the natural language query, selecting an entity associated with the media item that matches the entity type, selecting, from a media consumption database that identifies media items that have been indicated as consumed by the user, one or more media items that have been indicated as consumed by the user and that are associated with the selected entity, and providing a response to the query based on selecting the one or more media items that have been indicated as consumed by the user and that are associated with the selected entity.

Claims (97)

1. A computer-implemented method comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data by:

identifying the particular media item based on detecting a match between one or more features of the environmental audio data associated with the content playing in the environment of the user and one or more features of the particular media item; and

determining that the particular media item is identified in a media consumption database that identifies media items that are identified as having been previously consumed by the user; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the identified previous consumption of the particular media item.

2. The computer-implemented method of claim 1 , wherein providing the response to the natural language query submitted by the user comprises:

accessing, at the media consumption database, information that specifies contextual information associated with the previous consumption of the particular media item by the user; and

providing the response to the natural language query submitted by the user that includes at least a portion of the contextual information associated with the previous consumption of the particular media item by the user.

3. A computer-implemented method comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item, wherein the natural language query submitted by the user specifies an entity type; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the previous consumption of the particular media item of the specified entity type that is determined based at least on the environmental audio data associated with the content playing in the environment of the user.

4. The computer-implemented method of claim 3 , wherein providing the response to the natural language query submitted by the user that identifies the context of the previous consumption of the particular media item of the specified entity type comprises:

obtaining a transcription of the natural language query submitted by the user;

comparing the transcription of the natural language query submitted by the user to one or more keyword phrases that are each associated with an entity type; and

selecting the entity type based on determining that the transcription of the natural language query submitted by the user contains a particular keyword phrase that is associated with the selected entity type.

5. The computer-implemented method of claim 1 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item for output in a first region of a user interface displayed at the user device of the user; and

providing information that indicates one or more characteristics of the particular media item in a second region of the user interface displayed at the user device of the user.

6. The computer-implemented method of claim 1 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item in a first region of the user device of the user; and

providing one or more search results identified in response to the natural language query submitted by the user in a second region of the user interface displayed at the user device of the user.

7. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data by:

identifying the particular media item based on detecting a match between one or more features of the environmental audio data associated with the content playing in the environment of the user and one or more features of the particular media item; and

determining that the particular media item is identified in a media consumption database that identifies media items that are identified as having been previously consumed by the user; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the identified previous consumption of the particular media item.

8. The system of claim 7 , wherein providing the response to the natural language query submitted by the user comprises:

accessing, at the media consumption database, information that specifies contextual information associated with the previous consumption of the particular media item by the user; and

providing the response to the natural language query submitted by the user that includes at least a portion of the contextual information associated with the previous consumption of the particular media item by the user.

9. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item, wherein the natural language query submitted by the user specifies an entity type; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the previous consumption of the particular media item of the specified entity type that is determined based at least on the environmental audio data associated with the content playing in the environment of the user.

10. The system of claim 9 , wherein providing the response to the natural language query submitted by the user that identifies the context of the previous consumption of the particular media item of the specified entity type comprises:

obtaining a transcription of the natural language query submitted by the user;

comparing the transcription of the natural language query submitted by the user to one or more keyword phrases that are each associated with an entity type; and

selecting the entity type based on determining that the transcription of the natural language query submitted by the user contains a particular keyword phrase that is associated with the selected entity type.

11. The system of claim 7 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item for output in a first region of a user interface displayed at the user device of the user; and

providing information that indicates one or more characteristics of the particular media item in a second region of the user interface displayed at the user device of the user.

12. The system of claim 7 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item in a first region of the user device of the user; and

providing one or more search results identified in response to the natural language query submitted by the user in a second region of the user interface displayed at the user device of the user.

13. A computer-readable storage device encoded with a computer program, the program comprising instructions that, if executed by one or more computers, cause the one or more computers to perform operations comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data by:

identifying the particular media item based on detecting a match between one or more features of the environmental audio data associated with the content playing in the environment of the user and one or more features of the particular media item; and

determining that the particular media item is identified in a media consumption database that identifies media items that are identified as having been previously consumed by the user; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the identified previous consumption of the particular media item.

14. The computer-readable storage device of claim 13 , wherein providing the response to the natural language query submitted by the user comprises:

accessing, at the media consumption database, information that specifies contextual information associated with the previous consumption of the particular media item by the user; and

providing the response to the natural language query submitted by the user that includes at least a portion of the contextual information associated with the previous consumption of the particular media item by the user.

15. A computer-readable storage device encoded with a computer program, the program comprising instructions that, if executed by one or more computers, cause the one or more computers to perform operations comprising:

receiving waveform data generated by a user device of a user, the waveform data comprising:

an utterance detected by the user device, the utterance corresponding to a natural language query submitted by the user that requests information relating to a context of a prior consumption of a media item, wherein the natural language query submitted by the user specifies an entity type; and

environmental audio data detected by the user device within a threshold amount of time before or after detecting the utterance, the environmental audio data associated with content playing in an environment of the user;

processing the waveform data to separate the utterance corresponding to the natural language query from the environmental audio data by:

detecting voice activity of the user in a portion of the waveform data; and

extracting the portion of the waveform data that includes the detected voice activity of the user to separate the utterance corresponding to the natural language query from the environmental audio data;

identifying a previous consumption of a particular media item based on the environmental audio data; and

providing, to the user device, a response to the natural language query submitted by the user that identifies a context of the previous consumption of the particular media item of the specified entity type that is determined based at least on the environmental audio data associated with the content playing in the environment of the user.

16. The computer-readable storage device of claim 13 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item for output in a first region of a user interface displayed at the user device of the user; and

providing information that indicates one or more characteristics of the particular media item in a second region of the user interface displayed at the user device of the user.

17. The computer-readable storage device of claim 13 , wherein providing the response to the natural language query submitted by the user comprises:

providing information that identifies the context of the previous consumption of the particular media item in a first region of the user device of the user; and

providing one or more search results identified in response to the natural language query submitted by the user in a second region of the user interface displayed at the user device of the user.

Assignments (2)
CHANGE OF NAME Recorded Oct 5, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044129/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2016
From: SHARIFI, MATTHEW
To: GOOGLE INC.
Reel/Frame 039902/0673 →
Continuity (4)
Continuation 14217940 · Mar 18, 2014
Continuation 14047708 · Oct 7, 2013
Provisional Application 61866234 · Aug 15, 2013
Related Publication 20170004132A1 · Jan 5, 2017