IP Library Granted Patent US 10,755,723
Granted Patent B1
US 10,755,723 · App. 15/914,721 · Granted Aug 25, 2020

Shared audio functionality based on device grouping

Inventors: Albert M Scalise (San Jose, CA); Tony David (San Jose, CA)
Assignee: AMAZON TECHNOLOGIES, INC.
G10L21/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,755,723
App. No.
15/914,721
Granted
Aug 25, 2020
Kind
B1
Abstract

Techniques are described for shared audio functionality between multiple computing devices, based on identifying computing devices in a device set. The devices may provide audio output, audio input, or both audio output and input. The devices may be organized into one or more device sets based on location, supported functions, or other criteria. The shared audio functionality may enable a voice command received at one device to be employed for controlling audio output or other operations of other device(s) in the device set. Shared audio functionality between devices may also enable synchronized audio output through using multiple devices.

Claims (55)

1. A computer-implemented method comprising:

receiving, at a first computing device, a voice command to output audio content using a target device set;

determining the target device set comprises a second computing device and a third computing device;

determining the first computing device, the second computing device, and the third computing device are included in a second device set that enables shared audio functionality among the first computing device, the second computing device, and the third computing device;

receiving the audio content to be outputted by the second computing device and the third computing device; and

sending, over a network to the second computing device and the third computing device, a command to output the audio content by the second computing device within a period of time as when the third computing device outputs the audio content.

2. The method of claim 1 , further comprising:

storing the audio content; and

sending the audio content to the second computing device and the third computing device using the network, wherein the network comprises a wireless peer-to-peer network.

3. The method of claim 1 , further comprising:

receiving the audio content to be outputted from a fourth computing device; and

sending the audio content to the second computing device and the third computing device using the network, wherein the network comprises a wireless peer-to-peer network.

4. The method of claim 1 , further comprising:

outputting the audio content using an audio output function of the second computing device and the third computing device substantially synchronized based on synchronization of a first clock of the second computing device to a second clock of the third computing device.

5. The method of claim 1 , further comprising:

accessing a section of device set information delineated by a first set of one or more metadata tags, the first set of one or more metadata tags describing functions of the second computing device and the third computing device selected from the target device set to present the audio content.

6. A system comprising:

a microphone;

one or more memories storing computer-executable instructions; and

one or more hardware processors configured to execute the computer-executable instructions to:

receive, at a first computing device, a voice command from the microphone;

identify a target device set;

determine the target device set comprises a second computing device and a third computing device;

determine the first computing device, the second computing device, and the third computing device are included in a second device set that enables shared audio functionality among the first computing device, the second computing device, and the third computing device;

receive audio content to be outputted by the second computing device and the third computing device; and

send, over a network to the second computing device and the third computing device, a command to output the audio content by the second computing device within a period of time as when the third computing device outputs the audio content.

7. The system of claim 6 , wherein the audio content is received from a content service.

8. The system of claim 6 , wherein the audio content is received from a fourth computing device.

9. The system of claim 6 , wherein the network comprises a wireless peer-to-peer network.

10. The system of claim 6 , wherein:

device set information includes a section associated with the target device set, the section delineated by a first set of one or more metadata tags; and

the section associated with the target device set includes a second set of one or more metadata tags that describes functions of the second computing device and the third computing device to present the audio content.

11. The system of claim 10 , wherein one or more metadata tags in the second set of one or more metadata tags include an attribute indicating one or more audio output functions of the second computing device and the third computing device.

12. The system of claim 6 , the one or more hardware processors further configured to execute the computer-executable instructions to:

perform speech recognition to determine an utterance included in the voice command, the utterance including a description of the audio content to be presented; and

wherein the voice command is determined using the utterance.

13. The system of claim 6 , wherein the one or more hardware processors are further configured to execute the computer-executable instructions to send a request for the audio content.

14. One or more non-transitory computer-readable media storing instructions which, when executed by at least one processor of one or more servers, instruct the at least one processor to perform actions comprising:

receiving, at a first computing device, a voice command to output audio content using a target device set;

determining the target device set comprises the first computing device and a second computing device;

determining the first computing device and the second computing device are included in a second device set that enables shared audio functionality among the first computing device and the second computing device;

receiving the audio content to be outputted by the first computing device and the second computing device; and

sending, over a network to the second computing device, a command to output the audio content by the second computing device within a period of time as when the first computing device outputs the audio content.

15. The one or more non-transitory computer-readable media of claim 14 , further storing instructions, which when executed by the at least one processor of the one or more servers, instruct the at least one processor to perform actions comprising:

receiving the audio content from one or more of a content service or a third computing device.

16. The one or more non-transitory computer-readable media of claim 14 , further storing instructions, which when executed by the at least one processor of the one or more servers, instruct the at least one processor to perform actions comprising:

outputting the audio content using an audio output function of the second computing device and the first computing device substantially synchronized based on synchronization of a first clock of the first computing device to a second clock of the second computing device.

17. The one or more non-transitory computer-readable media of claim 14 , further storing instructions which, when executed by the at least one processor of the one or more servers, instruct the at least one processor to perform actions comprising:

accessing a section of device set information delineated by a first set of one or more metadata tags, the first set of the one or more metadata tags describing functions of the second computing device.

18. The one or more non-transitory computer-readable media of claim 14 , wherein the first computing device and the second computing device present the audio content serially, such that the first computing device and the second computing device present the audio content during non-overlapping time periods.

19. The one or more non-transitory computer-readable media of claim 14 , further storing instructions which, when executed by the at least one processor of the one or more servers, instruct the at least one processor to perform actions comprising:

sending a request for the audio content.

20. The one or more non-transitory computer-readable media of claim 14 , further storing instructions which, when executed by the at least one processor of the one or more servers, instruct the at least one processor to perform the actions comprising:

performing speech recognition to determine an utterance included in the voice command, the utterance including a description of the audio content to be presented; and

wherein the voice command is determined using the utterance.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 7, 2018
From: SCALISE, ALBERT M.; DAVID, TONY
To: AMAZON TECHNOLOGIES, INC.
Reel/Frame 045137/0480 →
Continuity (1)
Continuation 14227227 · Mar 27, 2014
Cited By (1)
US 12,334,073