IP Library Granted Patent US 10,097,939
Granted Patent B2
US 10,097,939 · App. 15/438,741 · Granted Oct 9, 2018

Compensation for speaker nonlinearities

Inventors: Timothy W. Sheen (Brighton, MA); Simon Jarvis (Cambridge, MA); Romi Kadri (Boston, MA); Yean-Nian Willy Chen (Santa Barbara, CA)
Assignee: SONOS, INC.
H04R29/007H04R3/04H04R2227/003H04R2227/005H04R2227/007
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,097,939
App. No.
15/438,741
Granted
Oct 9, 2018
Kind
B2
Abstract

A first signal may be received indicative of audio to be played by a speaker. A second signal may be received which comprises (i) a voice input received by a microphone and (ii) at least a portion of the audio played by the speaker at a same time that the microphone receives the voice input. Based on the first signal, nonlinearities output by the speaker which played the audio may be determined. At least the nonlinearities from the second signal may be removed to output a third signal comprising substantially the voice input received at the microphone.

Claims (33)

1. An audio system comprising:

a playback device comprising a speaker, the playback device disposed at a first location; and

a network microphone device disposed at a second location, the network microphone device being displaceable relative to the playback device, the network microphone device comprising:

a microphone;

a processor; and

memory storing instructions executable by the processor to cause the processor to:

receive a first signal indicative of audio to be played back via the speaker of the playback device and a second signal that comprises (i) a voice input received via the microphone and (ii) at least a portion of the audio played by the speaker of the playback device at a same time that the microphone receives the voice input; and

perform self-sound suppression on at least one of the first signal and the second signal, wherein performing self-sound suppression comprises:

based on the first signal, determining nonlinearities output via the speaker of the playback device by inputting a representation of the first signal into a model configured to output an indication of a frequency response that changes over time, wherein at least a portion of the frequency response is indicative of nonlinear audio effects, and wherein the nonlinear audio effects comprise an intermodulation distortion; and

removing at least a portion of the determined nonlinearities from the second signal to output a third signal comprising substantially the voice input received at the microphone.

2. The audio system of claim 1 , wherein the model is based on measurement of a position of a moving component of the speaker.

3. The audio system of claim 1 , wherein removing at least the nonlinearities from the second signal to output a third signal comprises determining a compensated audio signal based on the first signal and the nonlinear audio effects output by the speaker of the playback device, wherein the compensated audio signal characterizes how the audio played by the speaker sounds at the microphone.

4. The audio system of claim 1 , wherein removing at least the nonlinearities from the second signal to output a third signal comprises applying a transfer function to the first audio signal wherein the transfer function is a relative frequency response between a fourth signal indicative of second audio to be played by the speaker of the playback device and a fifth audio signal received at the microphone when the second audio is played.

5. The audio system of claim 1 , wherein the microphone is located within a given distance from speaker of the playback device, wherein at the given distance the microphone detects the audio played by the speaker of the playback device.

6. The audio system of claim 1 , further comprising computer instructions for converting the voice input in the third signal into text.

7. The audio system of claim 1 , wherein the first signal is tapped from a signal processing pathway associated with the speaker of the playback device after a time varying filter is applied to the first signal.

8. A method comprising:

receiving a first signal indicative of audio to be played back via a speaker of a playback device disposed at a first location and a second signal that comprises (i) a voice input received via a microphone of a network microphone device disposed at a second location, the network microphone device being displaceable relative to the playback device, and (ii) at least a portion of the audio played by the speaker at a same time that the microphone receives the voice input; and

performing self-sound suppression on at least one of the first signal and the second signal, wherein performing self-sound suppression comprises:

based on the first signal, determining nonlinearities output via the speaker of the playback device by inputting a representation of the first signal into a model configured to output an indication of a frequency response that changes over time, wherein at least a portion of the frequency response is indicative of nonlinear audio effects, and wherein the nonlinear audio effects comprise an intermodulation distortion; and

removing at least a portion of the determined nonlinearities from the second signal to output a third signal comprising substantially the voice input received at the microphone of the network microphone device.

9. The method of claim 8 , wherein the model is based on measurement of a position of a moving component of the speaker.

10. The method of claim 8 , wherein removing at least the nonlinearities from the second signal to output a third signal comprises determining a compensated audio signal based on the first signal and the nonlinear audio effects output by the speaker, wherein the compensated audio signal characterizes how the audio played by the speaker sounds at the microphone.

11. The method of claim 8 , wherein removing at least the nonlinearities from the second signal to output a third signal comprises applying a transfer function to the first audio signal wherein the transfer function is a relative frequency response between a fourth signal indicative of second audio to be played by the speaker and a fifth audio signal received at the microphone when the second audio is played.

12. The method of claim 8 , wherein the microphone is acoustically proximate to the speaker.

13. The method of claim 8 , further comprising converting the voice input in the third signal into text.

14. The method of claim 8 , wherein the first signal is tapped from a signal processing pathway associated with the speaker after a time varying filter is applied to the first signal.

15. A tangible non-transitory computer readable storage medium including instructions for execution by a processor, the instructions, when executed, cause the processor to implement a method comprising:

receiving a first signal indicative of audio to be played back via a speaker of a playback device disposed at a first location and a second signal that comprises (i) a voice input received via a microphone of a network microphone device disposed at a second location, the network microphone device being displaceable relative to the playback device, and (ii) at least a portion of the audio played by the speaker at a same time that the microphone receives the voice input; and

performing self-sound suppression on at least one of the first signal and the second signal, wherein performing self-sound suppression comprises:

based on the first signal, determining nonlinearities output via the speaker of the playback device by inputting a representation of the first signal into a model configured to output an indication of a frequency response that changes over time, wherein at least a portion of the frequency response is indicative of nonlinear audio effects, and wherein the nonlinear audio effects comprise an intermodulation distortion; and

removing at least a portion of the determined nonlinearities from the second signal to output a third signal comprising substantially the voice input received at the microphone of the network microphone device.

16. The tangible non-transitory computer readable storage medium of claim 15 , further comprising computer instructions to obtain acoustics of an environment in which the speaker is located; and apply the acoustics to the third signal comprising substantially the voice input received at the microphone.

Assignments (4)
RELEASE OF SECURITY INTEREST Recorded Oct 18, 2021
From: JPMORGAN CHASE BANK, N.A.
To: SONOS, INC.
Reel/Frame 058213/0597 →
SECURITY AGREEMENT Recorded Oct 15, 2021
From: SONOS, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 058123/0206 →
SECURITY INTEREST Recorded Aug 30, 2018
From: SONOS, INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 046991/0433 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 1, 2017
From: SHEEN, TIMOTHY W.; JARVIS, SIMON; KADRI, ROMI; CHEN, YEAN-NIAN WILLY
To: SONOS, INC.
Reel/Frame 041423/0238 →
Continuity (2)
Provisional Application 62298433 · Feb 22, 2016
Related Publication 20170245079A1 · Aug 24, 2017
Cited By (1)
US 12,659,663