IP Library Granted Patent US 10,013,971
Granted Patent B1
US 10,013,971 · App. 15/394,104 · Granted Jul 3, 2018

Automated speech pronunciation attribution

Inventors: Justin Lewis (Marina del Rey, CA); Lisa Takehana (San Bruno, CA)
Assignee: Google LLC
G10L13/02G10L15/02G10L15/22G10L25/51H04L67/18H04L67/24H04L67/306
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,013,971
App. No.
15/394,104
Granted
Jul 3, 2018
Kind
B1
Abstract

Methods, systems, and apparatus for determining candidate user profiles as being associated with a shared device, and identifying, from the candidate user profiles, candidate pronunciation attributes associated with at least one of the candidate user profiles determined to be associated with the shared device. The methods, systems, and apparatus are also for receiving, at the shared device, a spoken utterance; determining a received pronunciation attribute based on received audio data corresponding to the spoken utterance; comparing the received pronunciation attribute to at least one of the candidate pronunciation attributes; and selecting a particular pronunciation attribute from the candidate pronunciation attributes based on a result of the comparison of the received pronunciation attribute to at least one of the candidate pronunciation attributes. With the methods, systems, and apparatus, the particular pronunciation attribute, selected from the candidate pronunciation attributes, is provided for outputting audio associated with the spoken utterance.

Claims (55)

1. A computer-implemented method comprising:

determining candidate user profiles that are associated with a shared digital assistant device;

determining that a mobile computing device that is associated with a particular candidate user profile, from among the candidate user profiles that are associated with the shared digital assistant device, is indicated as being proximate to the shared digital assistant device;

identifying, from the candidate user profiles, candidate pronunciation attributes associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

receiving, at the shared digital assistant device, a spoken utterance;

determining a received pronunciation attribute based on received audio data corresponding to the spoken utterance;

comparing the received pronunciation attribute to at least one of the candidate pronunciation attributes that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

selecting a particular pronunciation attribute from the candidate pronunciation attributes that that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device based on a result of the comparison of the received pronunciation attribute to at least one of the candidate pronunciation attributes; and

providing the particular pronunciation attribute, selected from the candidate pronunciation attributes, for outputting audio associated with the spoken utterance.

2. The computer-implemented method of claim 1 , wherein determining candidate user profiles that are associated with a shared digital assistant device comprises:

determining a relationship between each of a plurality of user profiles and the shared digital assistant device;

determining, for each user profile, whether the relationship is indicative of an association between the user profile and the shared digital assistant device; and

identifying, for each user profile having a relationship indicative of an association with the shared digital assistant device, the user profile as being one of the candidate user profiles associated with the shared digital assistant device.

3. The computer-implemented method of claim 2 , wherein, for each of the plurality of user profiles, the relationship comprises a record of whether the user profile has been logged-in to the shared digital assistant device or whether at least one user device associated with the user profile has communicated with the shared digital assistant device.

4. The computer-implemented method of claim 2 , wherein, for each of the plurality of user profiles, the relationship comprises a geographical proximity of at least one user device associated with the user profile to the shared digital assistant device.

5. The computer-implemented method of claim 2 , wherein, for each of the plurality of user profiles, the relationship comprises a social connectivity, the social connectivity being based on at least one social connectivity metric.

6. The computer-implemented method of claim 1 , wherein each user profile of the candidate user profiles comprises one or more pronunciation attributes associated with a canonical identifier, the canonical identifier representing a particular pronunciation.

7. The computer-implemented method of claim 1 , further comprising:

providing an audio response to the spoken utterance, the audio response comprising the particular pronunciation selected from the candidate pronunciation attributes.

8. A system comprising one or more computers and one or more storage devices storing instructions that are operable, when executed by the one or more computers, to cause the one or more computers to perform operations comprising:

determining candidate user profiles that are associated with a shared digital assistant device;

determining that a mobile computing device that is associated with a particular candidate user profile, from among the candidate user profiles that are associated with the shared digital assistant device, is indicated as being proximate to the shared digital assistant device;

identifying, from the candidate user profiles, candidate pronunciation attributes associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

receiving, at the shared digital assistant device, a spoken utterance;

determining a received pronunciation attribute based on received audio data corresponding to the spoken utterance;

comparing the received pronunciation attribute to at least one of the candidate pronunciation attributes that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

selecting a particular pronunciation attribute from the candidate pronunciation attributes that that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device based on a result of the comparison of the received pronunciation attribute to at least one of the candidate pronunciation attributes; and

providing the particular pronunciation attribute, selected from the candidate pronunciation attributes, for outputting audio associated with the spoken utterance.

9. The system of claim 8 , wherein determining candidate user profiles that are associated with a shared digital assistant device comprises:

determining a relationship between each of a plurality of user profiles and the shared digital assistant device;

determining, for each user profile, whether the relationship is indicative of an association between the user profile and the shared digital assistant device; and

identifying, for each user profile having a relationship indicative of an association with the shared digital assistant device, the user profile as being one of the candidate user profiles associated with the shared digital assistant device.

10. The system of claim 9 , wherein, for each of the plurality of user profiles, the relationship comprises a record of whether the user profile has been logged-in to the shared digital assistant device or whether at least one user device associated with the user profile has communicated with the shared digital assistant device.

11. The system of claim 9 , wherein, for each of the plurality of user profiles, the relationship comprises a geographical proximity of at least one user device associated with the user profile to the shared digital assistant device.

12. The system of claim 9 , wherein, for each of the plurality of user profiles, the relationship comprises a social connectivity, the social connectivity being based on at least one social connectivity metric.

13. The system of claim 8 , wherein each user profile of the candidate user profiles comprises one or more pronunciation attributes associated with a canonical identifier, the canonical identifier representing a particular pronunciation.

14. The system of claim 8 , further comprising:

providing an audio response to the spoken utterance, the audio response comprising the particular pronunciation selected from the candidate pronunciation attributes.

15. A computer-readable storage device storing instructions executable by one or more computers which, upon such execution, cause the one or more computers to perform operations comprising:

determining candidate user profiles that are associated with a shared digital assistant device;

determining that a mobile computing device that is associated with a particular candidate user profile, from among the candidate user profiles that are associated with the shared digital assistant device, is indicated as being proximate to the shared digital assistant device;

identifying, from the candidate user profiles, candidate pronunciation attributes associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

receiving, at the shared digital assistant device, a spoken utterance;

determining a received pronunciation attribute based on received audio data corresponding to the spoken utterance;

comparing the received pronunciation attribute to at least one of the candidate pronunciation attributes that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device;

selecting a particular pronunciation attribute from the candidate pronunciation attributes that that are associated with the particular candidate user profile that is associated with the mobile computing device that is indicated as being proximate to the shared digital assistant device based on a result of the comparison of the received pronunciation attribute to at least one of the candidate pronunciation attributes; and

providing the particular pronunciation attribute, selected from the candidate pronunciation attributes, for outputting audio associated with the spoken utterance.

16. The computer-readable storage device of claim 15 , wherein determining candidate user profiles that are associated with a shared digital assistant device comprises:

determining a relationship between each of a plurality of user profiles and the shared digital assistant device;

determining, for each user profile, whether the relationship is indicative of an association between the user profile and the shared digital assistant device; and

identifying, for each user profile having a relationship indicative of an association with the shared digital assistant device, the user profile as being one of the candidate user profiles associated with the shared digital assistant device.

17. The computer-readable storage device of claim 16 , wherein, for each of the plurality of user profiles, the relationship comprises a record of whether the user profile has been logged-in to the shared digital assistant device or whether at least one user device associated with the user profile has communicated with the shared digital assistant device.

18. The computer-readable storage device of claim 16 , wherein, for each of the plurality of user profiles, the relationship comprises a geographical proximity of at least one user device associated with the user profile to the shared digital assistant device.

19. The computer-readable storage device of claim 16 , wherein, for each of the plurality of user profiles, the relationship comprises a social connectivity, the social connectivity being based on at least one social connectivity metric.

20. The computer-readable storage device of claim 15 , wherein each user profile of the candidate user profiles comprises one or more pronunciation attributes associated with a canonical identifier, the canonical identifier representing a particular pronunciation.

Assignments (2)
CHANGE OF NAME Recorded Oct 20, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044567/0001 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 21, 2017
From: LEWIS, JUSTIN; TAKEHANA, LISA
To: GOOGLE INC.
Reel/Frame 041314/0934 →