IP Library Granted Patent US 9,520,130
Granted Patent B2
US 9,520,130 · App. 15/001,894 · Granted Dec 13, 2016

Providing pre-computed hotword models

Inventor: Matthew Sharifi (Kilchberg, CH)
Assignee: Google Inc.
G10L15/22G10L15/063G10L15/08G10L15/265G10L15/18G10L2015/0631G10L2015/0638G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,520,130
App. No.
15/001,894
Filed
Jan 20, 2016
Granted
Dec 13, 2016
Kind
B2
Art Unit
2676
USPC
704/244
Abstract

Methods, systems, and apparatus, including computer programs encoded on a computer storage medium, for obtaining, for each of multiple words or sub-words, audio data corresponding to multiple users speaking the word or sub-word; training, for each of the multiple words or sub-words, a pre-computed hotword model for the word or sub-word based on the audio data for the word or sub-word; receiving a candidate hotword from a computing device; identifying one or more pre-computed hotword models that correspond to the candidate hotword; and providing the identified, pre-computed hotword models to the computing device.

Claims (43)

1. A computer-implemented method comprising:

receiving, by a computing device, a candidate hotword;

in response to receiving the candidate hotword, generating a prompt that requests that a user confirm whether to designate the candidate hotword as a hotword;

receiving a confirmation that the user confirms that the candidate hotword is to be designated as a hotword; and

in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, processing an utterance beginning with the candidate hotword as a voice command.

2. The method of claim 1 , wherein generating a prompt that requests that a user confirm whether to designate the candidate hotword as a hotword comprises:

providing a graphical user interface on a display of the computing device, wherein the graphical user interface includes a transcription of the candidate hotword and one or more graphical elements that can be selected to confirm that the candidate hotword represented by the transcription is to be designated as a hotword.

3. The method of claim 1 , wherein a hotword comprises a sequence of one or more words that causes the computing device to process an utterance beginning with the hotword as a voice command.

4. The method of claim 1 , wherein receiving the candidate hotword comprises receiving audio data corresponding to an utterance of the candidate hotword.

5. The method of claim 1 , comprising in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, obtaining a hotword model for the candidate hotword.

6. The method of claim 5 , wherein obtaining the hotword model for the candidate hotword comprises:

training one or more hotword models for the candidate hotword based on audio data.

7. The method of claim 6 , wherein training one or more hotword models for the candidate hotword based on audio data comprises:

obtaining audio data corresponding to multiple users speaking a word or a sub-word;

identifying a subset of the audio data that corresponds to the multiple users speaking the candidate hotword; and

training the one or more hotwords models using the subset of the audio data.

8. The method of claim 7 , wherein identifying the subset of the audio data is based on a geographic location of the computing device.

9. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

receiving, by a computing device, a candidate hotword;

in response to receiving the candidate hotword, generating a prompt that requests that a user confirm whether to designate the candidate hotword as a hotword;

receiving a confirmation that the user confirms that the candidate hotword is to be designated as a hotword; and

in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, processing an utterance beginning with the candidate hotword as a voice command.

10. The system of claim 9 , wherein a hotword comprises a sequence of one or more words that causes the computing device to process an utterance beginning with the hotword as a voice command.

11. The system of claim 9 , wherein receiving the candidate hotword comprises receiving audio data corresponding to an utterance of the candidate hotword.

12. The system of claim 9 , the operations comprising in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, obtaining a hotword model for the candidate hotword.

13. The system of claim 12 , wherein obtaining the hotword model for the candidate hotword comprises:

training one or more hotword models for the candidate hotword based on audio data.

14. The system of claim 13 , wherein training one or more hotword models for the candidate hotword based on audio data comprises:

obtaining audio data corresponding to multiple users speaking a word or a sub-word;

identifying a subset of the audio data that corresponds to the multiple users speaking the candidate hotword; and

training the one or more hotwords models using the subset of the audio data.

15. The system of claim 14 , wherein identifying the subset of the audio data is based on a geographic location of the computing device.

16. A non-transitory computer-readable medium storing software comprising instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:

receiving, by a computing device, a candidate hotword;

in response to receiving the candidate hotword, generating a prompt that requests that a user confirm whether to designate the candidate hotword as a hotword;

receiving a confirmation that the user confirms that the candidate hotword is to be designated as a hotword; and

in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, processing an utterance beginning with the candidate hotword as a voice command.

17. The medium of claim 16 , wherein a hotword comprises a sequence of one or more words that causes the computing device to process an utterance beginning with the hotword as a voice command.

18. The medium of claim 16 , wherein receiving the candidate hotword comprises receiving audio data corresponding to an utterance of the candidate hotword.

19. The medium of claim 16 , the operations comprising in response to receiving the confirmation that the user confirms that the candidate hotword is to be designated as a hotword, obtaining a hotword model for the candidate hotword.

20. The medium of claim 19 , wherein obtaining the hotword model for the candidate hotword comprises:

training one or more hotword models for the candidate hotword based on audio data.

Assignments (2)
CHANGE OF NAME Recorded Oct 5, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044129/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 21, 2016
From: SHARIFI, MATTHEW
To: GOOGLE INC.
Reel/Frame 037542/0626 →
Continuity (2)
Continuation 14340833 · Jul 25, 2014
Related Publication 20160140961A1 · May 19, 2016