IP Library Granted Patent US 11,403,065
Granted Patent B2
US 11,403,065 · App. 17/136,069 · Granted Aug 2, 2022

User interface customization based on speaker characteristics

Inventors: Eugene Weinstein (New York, NY); Ignacio L. Moreno (New York, NY)
Assignee: Google LLC
G06F3/167G06F3/04817G06F9/451G06F40/109
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,403,065
App. No.
17/136,069
Granted
Aug 2, 2022
Kind
B2
Abstract

Characteristics of a speaker are estimated using speech processing and machine learning. The characteristics of the speaker are used to automatically customize a user interface of a client device for the speaker.

Claims (34)

1. A computer-implemented method when executed on data processing hardware causes the data processing hardware to perform operations comprising:

receiving audio data corresponding to an utterance captured by a microphone of a user device, the user device shared by one or more users corresponding to a child age classification and one or more other users corresponding to an adult age classification;

generating an utterance feature vector derived from the received audio data corresponding to the utterance;

generating, using a neural network configured to receive the utterance feature vector as input, a probability distribution indicating a likelihood that the utterance was spoken by one of the one or more users corresponding to the child age classification;

determining whether the likelihood that the utterance was spoken by the one of the one or more users corresponding to the child age classification satisfies a threshold;

when the likelihood that the utterance was spoken by the one of the one or more users corresponding to the child age classification satisfies the threshold, generating a safe mode user interface that is associated with the child age classification, the safe mode user interface only allowing access to a selected set of child-safe applications; and

providing the safe mode user interface for display on a screen in communication with the data processing hardware.

2. The computer-implemented method of claim 1 , wherein the utterance is used to access a user profile.

3. The computer-implemented method of claim 1 , wherein the data processing hardware resides on the user device.

4. The computer-implemented method of claim 1 , wherein the data processing hardware resides on a server in communication with the user device via a network.

5. The computer-implemented method of claim 1 , wherein a button on the user device is pressed before the microphone of the user device captures the utterance.

6. The computer-implemented method of claim 1 , wherein the utterance comprises a voice command preceded by a hotword.

7. The computer-implemented method of claim 1 , wherein the safe mode user interface includes characteristics targeted to the one or more users corresponding to the child age classification that would not be present in a user interface that would be generated if the utterance was spoken by one of the one or more other users corresponding to the adult age classification.

8. The computer-implemented method of claim 1 , wherein the user device comprises a digital assistant device.

9. The computer-implemented method of claim 1 , wherein each user among the one or more users corresponding to the child age classification and the one or more other users corresponding to the adult age classification that share the user device establish a respective user profile that includes preferences for operating the user device and applications of interest to the respective user.

10. The computer-implemented method of claim 9 , wherein each respective user profile is associated with a respective voice signature for the respective user.

11. A system comprising:

data processing hardware; and

memory hardware in communication with the data processing hardware and storing instructions that when executed by the data processing hardware causes the data processing hardware to perform operations comprising:

receiving audio data corresponding to an utterance captured by a microphone of a user device, the user device shared by one or more users corresponding to a child age classification and one or more other users corresponding to an adult age classification;

generating an utterance feature vector derived from the received audio data corresponding to the utterance;

generating, using a neural network configured to receive the utterance feature vector as input, a probability distribution indicating a likelihood that the utterance was spoken by one of the one or more users corresponding to the child age classification;

determining whether the likelihood that the utterance was spoken by the one of the one or more users corresponding to the child age classification satisfies a threshold;

when the likelihood that the utterance was spoken by the one of the one or more users corresponding to the child age classification satisfies the threshold, generating a safe mode user interface that is associated with the child age classification, the safe mode user interface only allowing access to a selected set of child-safe applications; and

providing the safe mode user interface for display on a screen in communication with the data processing hardware.

12. The system of claim 11 , wherein the utterance is used to access a user profile.

13. The system of claim 11 , wherein the data processing hardware resides on the user device.

14. The system of claim 11 , wherein the data processing hardware resides on a server in communication with the user device via a network.

15. The system of claim 11 , wherein a button on the user device is pressed before the microphone of the user device captures the utterance.

16. The system of claim 11 , wherein the utterance comprises a voice command preceded by a hotword.

17. The system of claim 11 , wherein the safe mode user interface includes characteristics targeted to the one or more users corresponding to the child age classification that would not be present in a user interface that would be generated if the utterance was spoken by one of the one or more other users corresponding to the adult age classification.

18. The system of claim 11 , wherein the user device comprises a digital assistant device.

19. The system of claim 11 , wherein each user among the one or more users corresponding to the child age classification and the one or more other users corresponding to the adult age classification that share the user device establish a respective user profile that includes preferences for operating the user device and applications of interest to the respective user.

20. The system of claim 19 , wherein each respective user profile is associated with a respective voice signature for the respective user.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 29, 2020
From: WEINSTEIN, EUGENE; MORENO, IGNACIO L.
To: GOOGLE INC.
Reel/Frame 054760/0543 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 29, 2020
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 054864/0253 →
Continuity (3)
Continuation 15230891 · Aug 8, 2016
Continuation 14096608 · Dec 4, 2013
Related Publication 20210117153A1 · Apr 22, 2021