IP Library › Granted Patent US 11,734,326
Granted Patent B2
US 11,734,326 · App. 16/923,511 · Granted Aug 22, 2023

Profile disambiguation

Inventors: Rebecca Joy Lopdrup Miller (Seattle, WA); Dick Clarence Hardt (Seattle, WA); Joseph Jessup (Seattle, WA); Yu Bao (Issaquah, WA); Gonzalo Alvarez Barrio (Seattle, WA); Liron Torres (Redmond, WA)
Assignee: Amazon Technologies, Inc.
G06F16/335G10L15/18G10L15/26G10L17/06
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,734,326
App. No.
16/923,511
Granted
Aug 22, 2023
Kind
B2
Abstract

Techniques for disambiguating which profile, of multiple profiles, is to be used to respond to a user input are described. A device located in a communal space (e.g., a hotel room or suite of rooms, conference room, hospital room, etc.) may be associated with a device profile and a user profile of a user presently occupying the communal space. When the user inputs a command to the device (either by text or speech), a system associated with the device determines the profiles (e.g., a device profile and a user profile) associated with the device. The system determines one or more policies associated with the device. The one or more policies may correspond to rules for disambiguating which profile to use to execute with respect to the user input. Using the one or more policies, the system determines which profile is to be used, and causes a speechlet component to execute using information specific to the determined profile.

Claims (68)

1. A computer-implemented method comprising:

receiving, from a first device that is associated with a first profile, first data;

determining that the first data is associated with a second profile that is associated with a second device different from the first device;

based at least in part on the first data being associated with the second profile, determining to temporarily associate the second profile with the first device;

storing second data associating the second profile with the first device;

receiving input audio data corresponding to an utterance detected by the first device;

determining, using the second data, that the second first profile is associated with the first device;

using the second profile to determine the utterance was spoken by a user associated with the second profile;

performing speech processing using the input audio data to determine a command to be executed based at least in part on the second profile; and

causing the command to be executed.

2. The computer-implemented method of claim 1 , wherein using the second profile to determine the utterance was spoken by a user associated with the second profile comprises:

processing the input audio data with respect to third data, stored in the second profile, representing a voice of the user; and

determining, based at least in part on the processing, that the voice is represented in the input audio data.

3. The computer-implemented method of claim 1 , wherein:

performing the speech processing comprises determining natural language understanding (NLU) results data corresponding to the input audio data; and

the method further comprises determining, based at least in part on the NLU results data, that the command is to be executed based at least in part on the second profile.

4. The computer-implemented method of claim 1 , further comprising, prior to receiving the first data:

determining third data corresponding to a request for permission to associate the second profile with the first device; and

causing output of the third data.

5. The computer-implemented method of claim 4 , wherein causing output of the third data comprises causing the first second device to output the third data.

6. The computer-implemented method of claim 4 , further comprising, prior to determining the third data:

receiving second input audio data;

performing speech processing on the second input audio data to determine a request to perform a second command; and

determining that further permission data is needed to perform the second command.

7. The computer-implemented method of claim 1 , wherein:

the first data is received based at least in part on the user being located in a space associated with the first device.

8. The computer-implemented method of claim 7 , further comprising:

determining the user is no longer located in the space associated with the first device; and

causing the second data to be deleted.

9. The computer-implemented method of claim 1 , wherein:

the second data includes time data indicating a time of disassociation of the second profile from the first device; and

the method further comprises using the time data to determine that the input audio data is associated with a first time prior to the time of disassociation.

10. The computer-implemented method of claim 1 , wherein the first profile corresponds to a communal profile.

11. A system comprising:

at least one processor; and

at least one memory comprising instructions that, when executed by the at least one processor, cause the system to:

receive, from a first device that is associated with a first profile, first data;

determine that the first data is associated with a second profile that is associated with a second device different from the first device;

based at least in part on the first data being associated with the second profile, determine to temporarily associate the second profile with the first device;

store second data associating the second first profile with the first device;

receive input audio data corresponding to an utterance detected by the first device;

determine, using the second data, that the second first profile is associated with the first device;

use the second profile to determine the utterance was spoken by a user associated with the second profile;

perform speech processing using the input audio data to determine a command to be executed based at least in part on the second profile; and

cause the command to be executed.

12. The system of claim 11 , wherein the instructions that cause the system to use the second profile to determine the utterance was spoken by a user associated with the second profile comprise instructions that, when executed by the at least one processor, cause the system to:

process the input audio data with respect to third data, stored in the second profile, representing a voice of the user; and

determine, based at least in part on processing of the input audio data, that the voice is represented in the input audio data.

13. The system of claim 11 , wherein:

the instructions that cause the system to perform the speech processing comprise instructions that, when executed by the at least one processor, cause the system to determine natural language understanding (NLU) results data corresponding to the input audio data; and

the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to determine, based at least in part on the NLU results data, that the command is to be executed based at least in part on the second profile.

14. The system of claim 11 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to, prior to receipt of the first data:

determine third data corresponding to a request for permission to associate the second profile with the first device; and

cause output of the third data.

15. The system of claim 14 , wherein the instructions that cause the system to cause output of the third data comprises instructions that, when executed by the at least one processor, cause the system to cause the first device to output the third data.

16. The system of claim 14 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to, prior to determination of the third data:

receive second input audio data;

perform speech processing on the second input audio data to determine a request to perform a second command; and

determine that further permission data is needed to perform the second command.

17. The system of claim 11 , wherein:

the first data is received based at least in part on the user being located in a space associated with the first device.

18. The system of claim 17 , wherein the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to:

determine the user is no longer located in the space associated with the first device; and

cause the second data to be deleted.

19. The system of claim 11 , wherein:

the second data includes time data indicating a time of disassociation of the second profile from the first device; and

the at least one memory further comprises instructions that, when executed by the at least one processor, further cause the system to use the time data to determine that the input audio data is associated with a first time prior to the time of disassociation.

20. The system of claim 11 , wherein the first profile corresponds to a communal profile.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 8, 2020
From: MILLER, REBECCA JOY LOPDRUP; HARDT, DICK CLARENCE; JESSUP, JOSEPH; BAO, YU; BARRIO, GONZALO ALVAREZ; TORRES, LIRON
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 053154/0014 →
Continuity (2)
Continuation 15997068 · Jun 4, 2018
Related Publication 20200342011A1 · Oct 29, 2020
Cited By (1)
US 12,688,855