IP Library Granted Patent US 10,867,604
Granted Patent B2
US 10,867,604 · App. 16/271,560 · Granted Dec 15, 2020

Devices, systems, and methods for distributed voice processing

Inventors: Connor Kristopher Smith (New Hudson, MI); John Tolomei (Renton, WA); Betty Lee (Cambridge, MA)
Assignee: Sonos, Inc.
G10L15/22G10L15/08G10L15/30H04R1/406H04R3/005G10L2015/088G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,867,604
App. No.
16/271,560
Granted
Dec 15, 2020
Kind
B2
Abstract

Systems and methods for distributed voice processing are disclosed herein. In one example, the method includes detecting sound via a microphone array of a first playback device and analyzing, via a first wake-word engine of the first playback device, the detected sound. The first playback device may transmit data associated with the detected sound to a second playback device over a local area network. A second wake-word engine of the second playback device may analyze the transmitted data associated with the detected sound. The method may further include identifying that the detected sound contains either a first wake word or a second wake word based on the analysis via the first and second wake-word engines, respectively. Based on the identification, sound data corresponding to the detected sound may be transmitted over a wide area network to a remote computing device associated with a particular voice assistant service.

Claims (51)

1. A method comprising:

detecting sound via a microphone array of a first playback device;

transmitting data associated with the detected sound from the first playback device to a second playback device over a local area network;

analyzing, via a wake word engine of the second playback device, the transmitted data associated with the detected sound for identification of a wake word;

identifying that the detected sound contains the wake word based on the analysis via the wake word engine;

based on the identification, transmitting sound data corresponding to the detected sound from the second playback device to a remote computing device over a wide area network, wherein the remote computing device is associated with a particular voice assistant service;

receiving via the second playback device a response from the remote computing device, wherein the response is based on the detected sound;

transmitting a message from the second playback device to the first playback device over the local area network, wherein the message is based on the response from the remote computing device and includes instructions to perform an action; and

performing the action via the first playback device.

2. The method of claim 1 , wherein the action is a first action and the method further comprises performing a second action via the second playback device, wherein the second action is based on the response from the remote computing device.

3. The method of claim 1 , further comprising disabling a wake word engine of the first playback device in response to the identification of the wake word via the wake word engine of the second playback device.

4. The method of claim 3 , further comprising enabling a wake word engine of the first playback device after the second playback device receives the response from the remote computing device.

5. The method of claim 4 , wherein the wake word is a second wake word, and wherein the wake word engine of the first playback device is configured to detect a first wake word that is different than the second wake word.

6. The method of claim 1 , wherein the first playback device is configured to communicate with the remote computing device associated with the particular voice assistant service.

7. The method of claim 1 , wherein the remote computing device is a first remote computing device and the voice assistant service is a first voice assistant service, and wherein the first playback device is configured to detect a wake word associated with a second voice assistant service different than the first voice assistant service.

8. A first playback device comprising:

one or more processors;

a computer-readable medium storing instructions that, when executed by the one or more processors, cause the first playback device to perform operations comprising:

receiving, from a second playback device over a local area network, data associated with sound detected via a microphone array of the second playback device;

analyzing, via a wake word engine of the first playback device, the data associated with the detected sound for identification of a wake word;

identifying that the detected sound contains the wake word based on the analysis via the wake word engine;

based on the identification, transmitting sound data corresponding to the detected sound to a remote computing device over a wide area network, wherein the remote computing device is associated with a particular voice assistant service;

receiving a response from the remote computing device, wherein the response is based on the detected sound; and

transmitting a message to the second playback device over the local area network, wherein the message is based on the response from the remote computing device and includes instructions for the second playback device to perform an action.

9. The first playback device of claim 8 , wherein the action is a first action and the operations further comprise performing a second action via the first playback device, wherein the second action is based on the response from the remote computing device.

10. The first playback device of claim 8 , wherein the operations further comprise disabling a wake word engine of the second playback device in response to the identification of the wake word via the wake word engine of the first playback device.

11. The first playback device of claim 10 , wherein the operations further comprise enabling the wake word engine of the second playback device after the first playback device receives the response from the remote computing device.

12. The first playback device of claim 11 , wherein the wake word is a first wake word, and wherein the wake word engine of the second playback device is configured to detect a second wake word that is different than the first wake word.

13. The first playback device of claim 8 , wherein the second playback device is configured to communicate with the remote computing device associated with the particular voice assistant service.

14. The first playback device of claim 8 , wherein the remote computing device is a first remote computing device and the voice assistant service is a first voice assistant service, and wherein the second playback device is configured to detect a wake word associated with a second voice assistant service different than the first voice assistant service.

15. A system, comprising:

a first playback device comprising:

one or more processors;

a microphone array; and

a first computer-readable medium storing instructions that, when executed by the one or more processors, cause the first playback device to perform first operations, the first operations comprising:

detecting sound via the microphone array;

transmitting data associated with the detected sound to a second playback device over a local area network;

the second playback device comprising:

one or more processors; and

a second computer-readable medium storing instructions that, when executed by the one or more processors, cause the second playback device to perform second operations, the second operations comprising:

analyzing, via a wake word engine of the second playback device, the transmitted data associated with the detected sound from the first playback device for identification of a wake word;

identifying that the detected sound contains the wake word based on the analysis via the wake word engine;

based on the identification, transmitting sound data corresponding to the detected sound to a remote computing device over a wide area network, wherein the remote computing device is associated with a particular voice assistant service;

receiving a response from the remote computing device, wherein the response is based on the detected sound; and

transmitting a message to the first playback device over the local area network, wherein the message is based on the response from the remote computing device and includes instructions to perform an action,

wherein the first computer-readable medium of the first playback device causes the first playback device to perform the action from the instructions received from the second playback device.

16. The system of claim 15 , wherein the action is a first action and the second operations further comprise performing a second action via the second playback device, wherein the second action is based on the response from the remote computing device.

17. The system of claim 15 , wherein the second operations further comprise disabling a wake word engine of the first playback device in response to the identification of the wake word via the wake word engine of the second playback device.

18. The system of claim 17 , wherein the second operations further comprise enabling the wake word engine of the first playback device after the second playback device receives the response from the remote computing device.

19. The system of claim 15 , wherein the first playback device is configured to communicate with the remote computing device associated with the particular voice assistant service.

20. The system of claim 15 , wherein the remote computing device is a first remote computing device and the voice assistant service is a first voice assistant service, and wherein the first playback device is configured to detect a wake word associated with a second voice assistant service different than the first voice assistant service.

Assignments (2)
SECURITY AGREEMENT Recorded Oct 15, 2021
From: SONOS, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 058123/0206 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 12, 2019
From: SMITH, CONNOR KRISTOPHER; TOLOMEI, JOHN; LEE, BETTY
To: SONOS, INC.
Reel/Frame 048576/0994 →
Continuity (1)
Related Publication 20200258513A1 · Aug 13, 2020
Cited By (30)
US 12,192,713 US 12,211,490 US 12,212,945 US 12,217,748 US 12,230,291 US 12,236,932 US 12,277,368 US 12,279,096 US 12,283,269 US 12,288,558 US 12,322,390 US 12,327,549 US 12,327,556 US 12,360,734 US 12,375,052 US 12,387,716 US 12,406,676 US 12,424,220 US 12,438,977 US 12,505,832 US 12,505,841 US 12,513,466 US 12,513,479 US 12,518,755 US 12,518,756 US 12,579,978 US 12,626,717 US 12,640,148 US 12,699,543 US 12,711,962