IP Library › Granted Patent US 11,087,743
Granted Patent B2
US 11,087,743 · App. 16/682,716 · Granted Aug 10, 2021

Multi-user authentication on a device

Inventors: Ignacio Lopez Moreno (New York, NY); Diego Melendo Casado (Mountain View, CA)
Assignee: GOOGLE LLC
G10L15/08G06F16/636G06F21/32G06K9/00362G10L15/07G10L15/22G10L17/00G10L17/06G10L15/26G10L2015/088
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,087,743
App. No.
16/682,716
Filed
Nov 13, 2019
Granted
Aug 10, 2021
Kind
B2
Art Unit
2658
USPC
704/244
Abstract

In some implementations, an utterance is determined to include a particular user speaking a hotword based at least on a first set of samples of the particular user speaking the hotword. In response to determining that an utterance includes a particular user speaking a hotword based at least on a first set of samples of the particular user speaking the hotword, at least a portion of the utterance is stored as a new sample. A second set of samples of the particular user speaking the utterance is obtained, where the second set of samples includes the new sample and less than all the samples in the first set of samples. A second utterance is determined to include the particular user speaking the hotword based at least on the second set of samples of the user speaking the hotword.

Claims (51)

1. A computer-implemented method comprising:

determining that an utterance captured by a first speech-enabled device includes a particular user speaking a hotword based at least on a first hotword detection model generated from a first set of samples of the particular user speaking the hotword;

in response to determining that the utterance includes the particular user speaking the hotword based at least on the first hotword detection model generated from the first set of samples of the particular user speaking the hotword, selecting at least a portion of the utterance as a new sample for inclusion in a second set of samples of the particular user speaking the hotword; and

providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples.

2. The method of claim 1 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

determining that the second speech-enabled device is used by the particular user.

3. The method of claim 1 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

providing the new sample to the second speech-enabled device to generate the second hotword detection model.

4. The method of claim 3 , wherein providing the new sample to the second speech-enabled device to train the second hotword detection model comprises:

determining the second speech-enable device does not already store the new sample.

5. The method of claim 4 , wherein determining the second speech-enable device does not already store the new sample comprises:

receiving an identifier of a set of samples that the second speech-enable device has stored; and

determining that the set of samples does not include the new sample.

6. The method of claim 4 , wherein determining the second speech-enable device does not already store the new sample comprises:

receiving, for each sample stored by the second speech-enabled device, an identifier of the sample; and

determining that none of the identifiers received match an identifier of the new sample.

7. The method of claim 1 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples is in response to selecting at least the portion of the utterance as the new sample for inclusion in the second set of samples of the particular user speaking the hotword.

8. The method of claim 1 , wherein selecting at least a portion of the utterance as a new sample for inclusion in a second set of samples of the particular user speaking the hotword comprises:

determining that a type of the first speech-enabled device matches a type of the second speech-enabled device; and

in response to determining that the type of the first speech-enabled device matches the type of the second speech-enabled device, selecting at least the portion of the utterance as the new sample for inclusion in the second set of samples of the particular user speaking the hotword.

9. A system comprising:

one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

determining that an utterance captured by a first speech-enabled device includes a particular user speaking a hotword based at least on a first hotword detection model generated from a first set of samples of the particular user speaking the hotword;

in response to determining that the utterance includes the particular user speaking the hotword based at least on the first hotword detection model generated from the first set of samples of the particular user speaking the hotword, selecting at least a portion of the utterance as a new sample for inclusion in a second set of samples of the particular user speaking the hotword; and

providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples.

10. The system of claim 9 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

determining that the second speech-enabled device is used by the particular user.

11. The system of claim 9 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

providing the new sample to the second speech-enabled device to generate the second hotword detection model.

12. The system of claim 11 , wherein providing the new sample to the second speech-enabled device to generate the second hotword detection model comprises:

determining the second speech-enable device does not already store the new sample.

13. The system of claim 12 , wherein determining the second speech-enable device does not already store the new sample comprises:

receiving an identifier of a set of samples that the second speech-enable device has stored; and

determining that the set of samples does not include the new sample.

14. The system of claim 12 , wherein determining the second speech-enable device does not already store the new sample comprises:

receiving, for each sample stored by the second speech-enabled device, an identifier of the sample; and

determining that none of the identifiers received match an identifier of the new sample.

15. The system of claim 9 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples is in response to selecting at least the portion of the utterance as the new sample for inclusion in the second set of samples of the particular user speaking the hotword.

16. The system of claim 9 , wherein selecting at least a portion of the utterance as a new sample for inclusion in a second set of samples of the particular user speaking the hotword comprises:

determining that a type of the first speech-enabled device matches a type of the second speech-enabled device; and

in response to determining that the type of the first speech-enabled device matches the type of the second speech-enabled device, selecting at least the portion of the utterance as the new sample for inclusion in the second set of samples of the particular user speaking the hotword.

17. A non-transitory computer-readable medium storing software comprising instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:

determining that an utterance captured by a first speech-enabled device includes a particular user speaking a hotword based at least on a first hotword detection model generated from a first set of samples of the particular user speaking the hotword;

in response to determining that the utterance includes the particular user speaking the hotword based at least on the first hotword detection model generated from the first set of samples of the particular user speaking the hotword, selecting at least a portion of the utterance as a new sample for inclusion in a second set of samples of the particular user speaking the hotword; and

providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples.

18. The medium of claim 17 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

determining that the second speech-enabled device is used by the particular user.

19. The medium of claim 17 , wherein providing the second set of samples of the particular user speaking the hotword to a second speech-enabled device, that is used by the particular user, to train a second hotword detection model from the second set of samples comprises:

providing the new sample to the second speech-enabled device to generate the second hotword detection model.

20. The medium of claim 19 , wherein providing the new sample to the second speech-enabled device to generate the second hotword detection model comprises:

determining the second speech-enable device does not already store the new sample.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 13, 2019
From: LOPEZ MORENO, IGNACIO; MELENDO CASADO, DIEGO
To: GOOGLE LLC
Reel/Frame 051002/0068 →
Continuity (4)
Continuation 15956493 · Apr 18, 2018
Provisional Application 62567372 · Oct 3, 2017
Provisional Application 62488000 · Apr 20, 2017
Related Publication 20200082812A1 · Mar 12, 2020
Cited By (1)
US 12,609,112