IP Library › Granted Patent US 11,769,505
Granted Patent B2
US 11,769,505 · App. 17/658,717 · Granted Sep 26, 2023

Echo of tone interferance cancellation using two acoustic echo cancellers

Inventor: Saeed Bagheri Sereshki (Goleta, CA)
Assignee: Sonos, Inc.
G10L15/22G10K11/1785G10L15/08G10L21/0208G10L21/0232G10L25/78G10L2015/088G10L2015/223G10L2021/02085H04M3/53H04S7/301
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,769,505
App. No.
17/658,717
Granted
Sep 26, 2023
Kind
B2
Abstract

Example techniques involve systems with multiple acoustic echo cancellers. An example implementation captures first audio within an acoustic environment and detecting, within the captured first audio content, a wake-word. In response to the wake-word and before playing an acknowledgement tone, the implementation activates (a) a first sound canceller when one or more speakers are playing back audio content or (b) a second sound canceller when the one or more speakers are idle. In response to the wake-word and after activating either (a) the first sound canceller or (b) the second sound canceller, the implementation outputs the acknowledgement tone via the one or more speakers. The implementation captures second audio within the acoustic environment and cancelling the acoustic echo of the acknowledgement tone from the captured second audio using the activated sound canceller.

Claims (78)

1. A playback device comprising:

a first acoustic echo canceller (AEC);

a second AEC;

one or more audio transducers;

one or more microphones;

at least one processor; and

data storage including instructions that are executable by the at least one processor such that the playback device is configured to:

play back first audio content via the one or more audio transducers;

during playback of the first audio content, capture first microphone data via the one or more microphones;

cancel echo of the first audio content from the captured first microphone data via the first AEC;

receive data representing second audio content for playback;

after receipt of at least a portion of the data representing the second audio content, activate the second AEC in place of the first AEC to cancel echo from playback;

play back the second audio content via the one or more audio transducers;

during playback of the second audio content, capture second microphone data via the one or more microphones; and

cancel echo of the second audio content from the captured second microphone data via the second AEC.

2. The playback device of claim 1 , wherein the first audio content is a first type of content, and wherein the data storage further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

determine that the second audio content is a second type of content, and wherein the instructions that are executable by the at least one processor such that the playback device is configured to activate the second AEC in place of the first AEC comprise instructions that are executable by the at least one processor such that the playback device is configured to:

based on the determination that the second audio content is the second type of content, activate the second AEC.

3. The playback device of claim 2 , wherein the first type of content has sound in a first frequency range, and wherein the second type of content has sound in a second frequency range that is wider than the first frequency range.

4. The playback device of claim 3 , wherein the first type of content is a tone.

5. The playback device of claim 2 , wherein the data storage further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

receive data representing third audio content for playback;

determine that the third audio content is the first type of content;

based on the determination that the second audio content is the second type of content, activate the first AEC in place of the second AEC to cancel echo from playback;

play back the third audio content via the one or more audio transducers;

during playback of the third audio content, capture third microphone data via the one or more microphones; and

cancel echo of the third audio content from the captured third microphone data via the first AEC.

6. The playback device of claim 1 , wherein a portion of the second microphone data represents a voice input, and wherein the data storage further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

after cancellation of the echo of the second audio content from the captured second microphone data, send data representing the voice input to a voice assistant for processing.

7. The playback device of claim 6 , wherein the data storage further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

while the first AEC and the second AEC are deactivated, capture additional microphone data via the one or more microphones;

detect a wake word for the voice assistant in the additional microphone data; and

based on detection of the wake word, activate the first AEC.

8. The playback device of claim 1 , wherein the data storage further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

when the second audio content stops playing, activate the first AEC in place of the second AEC to cancel echo from playback.

9. The playback device of claim 1 , further comprising a network interface, wherein the instructions that are executable by the at least one processor such that the playback device is configured to receive data representing second audio content for playback comprise instructions that are executable by the at least one processor such that the playback device is configured to:

receive, via the network interface, a data stream representing the second audio content.

10. The playback device of claim 1 , further comprising a housing, the housing carrying the one or more audio transducers, the one or more microphones, and the at least one processor.

11. A tangible, non-transitory computer-readable medium comprising instructions that are executable by at least one processor such that a playback device is configured to:

play back first audio content via one or more audio transducers;

during playback of the first audio content, capture first microphone data via one or more microphones;

cancel echo of the first audio content from the captured first microphone data via a first acoustic echo canceller (AEC);

receive data representing second audio content for playback;

after receipt of at least a portion of the data representing the second audio content, activate a second AEC in place of the first AEC to cancel echo from playback;

play back the second audio content via the one or more audio transducers;

during playback of the second audio content, capture second microphone data via the one or more microphones; and

cancel echo of the second audio content from the captured second microphone data via the second AEC.

12. The non-transitory computer-readable medium of claim 11 , wherein the first audio content is a first type of content, and wherein the non-transitory computer-readable medium further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

determine that the second audio content is a second type of content, and wherein the instructions that are executable by the at least one processor such that the playback device is configured to activate the second AEC in place of the first AEC comprise instructions that are executable by the at least one processor such that the playback device is configured to:

based on the determination that the second audio content is the second type of content, activate the second AEC.

13. The non-transitory computer-readable medium of claim 12 , wherein the first type of content has sound in a first frequency range, and wherein the second type of content has sound in a second frequency range that is wider than the first frequency range.

14. The non-transitory computer-readable medium of claim 13 , wherein the first type of content is a tone.

15. The non-transitory computer-readable medium of claim 12 , wherein the non-transitory computer-readable medium further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

receive data representing third audio content for playback;

determine that the third audio content is the first type of content;

based on the determination that the second audio content is the second type of content, activate the first AEC in place of the second AEC to cancel echo from playback;

play back the third audio content via the one or more audio transducers;

during playback of the third audio content, capture third microphone data via the one or more microphones; and

cancel echo of the third audio content from the captured third microphone data via the first AEC.

16. The non-transitory computer-readable medium of claim 11 , wherein a portion of the second microphone data represents a voice input, and wherein the non-transitory computer-readable medium further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

after cancellation of the echo of the second audio content from the captured second microphone data, send data representing the voice input to a voice assistant for processing.

17. The non-transitory computer-readable medium of claim 16 , wherein the non-transitory computer-readable medium further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

while the first AEC and the second AEC are deactivated, capture additional microphone data via the one or more microphones;

detect a wake word for the voice assistant in the additional microphone data; and

based on detection of the wake word, activate the first AEC.

18. The non-transitory computer-readable medium of claim 11 , wherein the non-transitory computer-readable medium further comprises instructions that are executable by the at least one processor such that the playback device is configured to:

when the second audio content stops playing, activate the first AEC in place of the second AEC to cancel echo from playback.

19. The non-transitory computer-readable medium of claim 11 , wherein the instructions that are executable by the at least one processor such that the playback device is configured to receive data representing second audio content for playback comprise instructions that are executable by the at least one processor such that the playback device is configured to:

receive, via a network interface, a data stream representing the second audio content.

20. A method to be performed by a playback device, the method comprising:

playing back first audio content via one or more audio transducers;

while playing back the first audio content, capturing first microphone data via one or more microphones;

cancelling echo of the first audio content from the captured first microphone data via a first acoustic echo canceller (AEC);

receiving data representing second audio content for playback;

after receiving at least a portion of the data representing the second audio content, activating a second AEC in place of the first AEC to cancel echo from playback;

playing back the second audio content via the one or more audio transducers;

while playing back the second audio content, capturing second microphone data via the one or more microphones; and

cancelling echo of the second audio content from the captured second microphone data via the second AEC.

Assignments (2)
SECURITY INTEREST Recorded Jan 30, 2026
From: SONOS, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 074533/0615 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 11, 2022
From: SERESHKI, SAEED BAGHERI
To: SONOS, INC.
Reel/Frame 059561/0041 →
Continuity (3)
Continuation 16845946 · Apr 10, 2020
Continuation 15718521 · Sep 28, 2017
Related Publication 20220383846A1 · Dec 1, 2022
Cited By (27)
US 12,192,713 US 12,210,801 US 12,211,490 US 12,231,859 US 12,236,932 US 12,288,558 US 12,314,633 US 12,322,390 US 12,340,802 US 12,360,734 US 12,374,334 US 12,375,052 US 12,424,220 US 12,438,977 US 12,462,802 US 12,498,899 US 12,505,832 US 12,513,479 US 12,518,755 US 12,518,756 US 12,562,167 US 12,578,779 US 12,579,978 US 12,699,543 US 12,711,962 US 12,732,547 US 12,749,486