IP Library Granted Patent US 11,825,290
Granted Patent B2
US 11,825,290 · App. 18/308,016 · Granted Nov 21, 2023

Media playback based on sensor data

Inventors: Jonathon Reilly (Cambridge, MA); Niels Van Erven (Santa Barbara, CA)
Assignee: Sonos, Inc.
H04S7/303G01S3/023G01S3/80G06F3/165H04R3/04H04R27/00H04R29/001H04R29/008H04S7/302H04R2227/005H04R2430/01H04S7/301
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,825,290
App. No.
18/308,016
Granted
Nov 21, 2023
Kind
B2
Abstract

Example techniques relate to playback based on acoustic signals in a system including a first network device and a second network device. A first network device may detect a presence of a user using a camera and/or infrared sensors. The first network device sends, in response to detecting the presence of the user, a particular signal via the first network interface. The second network device receives data corresponding to the particular signal and plays back an audio output corresponding to the particular signal.

Claims (61)

1. A system comprising a first network device and a second network device, the first network device comprising:

a first audio transducer;

a microphone;

one or more sensors;

a first network interface;

at least one first processor; and

at least one first non-transitory computer-readable medium storing program instructions that are executable by the at least one first processor such that the first network device is configured to perform first functions comprising:

detecting a human presence in sensor data received via at least one sensor of the one or more sensors;

causing the second network device to play back a first audio output corresponding to a first signal, wherein causing the second network device to play back the first audio output comprises: based on detecting the human presence, sending the first signal via the first network interface to the second network device; and

playing back, via the first audio transducer, an audio output corresponding to a second signal, and

the second network device comprising:

a second audio transducer a second network interface;

at least one second processor; and

at least one second non-transitory computer-readable medium storing program instructions that are executable by the at least one second processor such that the second network device is configured to perform second functions comprising:

playing back the first audio output corresponding to the first signal via the second audio transducer; and

sending, via the second network interface to the first network device, the second signal.

2. The system of claim 1 , wherein the first audio transducer comprises a loudspeaker, wherein the second network device comprises an additional microphone, wherein the second signal comprises speech, and wherein playing back the audio output corresponding to the second signal comprises playing back the speech via the loudspeaker.

3. The system of claim 1 , wherein detecting the human presence comprises detecting movement indicating the human presence.

4. The system of claim 3 , wherein the one or more sensors comprise a camera, and wherein detecting the movement indicating the human presence comprises detecting the movement indicating the human presence via frames captured by the camera.

5. The system of claim 3 , wherein the one or more sensors comprise an infrared sensor, and wherein detecting movement indicating the human presence comprises detecting the movement indicating the human presence in samples captured by the infrared sensor.

6. The system of claim 3 , wherein the one or more sensors comprise a sensor configured to detect heat, and wherein detecting the movement indicating the human presence comprises detecting the movement indicating the human presence in samples captured by the sensor configured to detect heat.

7. The system of claim 1 , wherein the one or more sensors comprise a depth sensor, and wherein detecting the human presence comprises detecting movement indicating the human presence in sample captured by the depth sensor.

8. The system of claim 1 , wherein sending the first signal to the second network device comprises sending the first signal to the second network device via at least one third network device.

9. The system of claim 1 , wherein the first network device comprises a lighting fixture.

10. The system of claim 9 , wherein the lighting fixture comprises a lamp.

11. The system of claim 1 , wherein the first network device comprises an enclosure configured for outdoor use, wherein the enclosure is configured to carry one or more of the first network interface, the first audio transducer, the one or more sensors, the microphone, the at least one first processor, and the at least one first non-transitory computer-readable medium.

12. The system of claim 1 , wherein the first functions comprise causing, via the first network interface, a third network device to play back a third audio output corresponding to the first signal, wherein the third audio output is different from both of (i) the first audio output and (ii) a second audio output.

13. The system of claim 1 , wherein the first functions further comprise:

receiving, via the microphone of the first network device, microphone data comprising speech; and

causing the second network device to play back, via the audio transducer of the second network device, a third audio output, wherein causing the second network device to play back the third audio output comprises:

after receiving the microphone data, sending, via the first network interface to the second network device, a third signal comprising the speech, wherein the third audio output is different from the first audio output.

14. A first network device comprising:

an audio transducer;

a microphone;

a network interface;

at least one processor; and

at least one non-transitory computer-readable medium storing program instructions that are executable by the at least one processor such that the first network device is configured to perform functions comprising:

detecting a human presence in sensor data received via at least one sensor of one or more sensors;

causing a second network device to play back a first audio output corresponding to a first signal, wherein causing the second network device to play back the first audio output comprises: based on detecting the human presence, sending the first signal via the first network interface to the second network device;

receiving, via the network interface, second data representing a second signal from the second network device; and

playing back, via the audio transducer, an audio output corresponding to the second signal.

15. The first network device of claim 14 , wherein the audio transducer comprises a loudspeaker, wherein the second network device comprises an additional microphone, wherein the second signal comprises speech, and wherein playing back the audio output corresponding to the second signal comprises playing back the speech via the loudspeaker.

16. The first network device of claim 14 , wherein detecting the human presence comprises detecting movement indicating the human presence.

17. The first network device of claim 16 , wherein the one or more sensors comprise a camera, and wherein detecting the movement indicating the human presence comprises detecting the movement indicating the human presence via frames captured by the camera.

18. The first network device of claim 16 , wherein the one or more sensors comprise an infrared sensor, and wherein detecting movement indicating the human presence comprises detecting the movement indicating the human presence in samples captured by the infrared sensor.

19. The first network device of claim 16 , wherein the one or more sensors comprise a sensor configured to detect heat, and wherein detecting the movement indicating the human presence comprises detecting the movement indicating the human presence in samples captured by the sensor configured to detect heat.

20. The first network device of claim 16 , wherein the one or more sensors comprise a depth sensor, and wherein detecting the movement indicating the human presence comprises detecting the movement indicating the human presence in sample captured by the depth sensor.

21. The first network device of claim 16 , wherein sending the first signal to the second network device comprises sending the first signal to the second network device via at least one third network device.

22. The first network device of claim 16 , wherein the first network device comprises a lighting fixture.

23. The first network device of claim 22 , wherein the lighting fixture comprises a lamp.

24. The first network device of claim 16 , wherein the first network device comprises an enclosure configured for outdoor use, wherein the enclosure is configured to carry one or more of the network interface, the audio transducer, the one or more sensors, the microphone, the at least one processor, and the at least one non-transitory computer-readable medium.

25. The first network device of claim 16 , wherein the functions comprise causing, via the first network interface, a third network device to play back a third audio output corresponding to the first signal, wherein the third audio output is different from both of (i) the first audio output and (ii) a second audio output.

26. The first network device of claim 16 , wherein the functions further comprise:

receiving, via the microphone of the first network device, microphone data comprising speech; and

causing the second network device to play back, via the audio transducer of the second network device, a third audio output, wherein causing the second network device to play back the third audio output comprises:

after receiving the microphone data, sending, via the first network interface to the second network device, a third signal comprising the speech, wherein the third audio output is different from the first audio output.

27. A tangible, non-transitory computer-readable medium comprising program instructions that are executable by at least one processor to configure a first network device to:

detecting a human presence in sensor data received via at least one sensor of one or more sensors carried by the first network device;

causing a second network device to play back a first audio output corresponding to a first signal, wherein causing the second network device to play back the first audio output comprises: based on detecting the human presence, sending the first signal via a network interface to the second network device, wherein the network interface is carried by the first network device;

receiving, via the network interface, second data representing a second signal from the second network device; and

playing back, via an audio transducer carried by the first network device, an audio output corresponding to the second signal.

Assignments (2)
SECURITY INTEREST Recorded Jan 30, 2026
From: SONOS, INC.
To: JPMORGAN CHASE BANK, N.A., AS ADMINISTRATIVE AGENT
Reel/Frame 074533/0615 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 27, 2023
From: REILLY, JONATHON; VAN ERVEN, NIELS
To: SONOS, INC.
Reel/Frame 063464/0531 →
Continuity (10)
Continuation 17543014 · Dec 6, 2021
Continuation 17207640 · Mar 20, 2021
Continuation 17104466 · Nov 25, 2020
Continuation 16658896 · Oct 21, 2019
Continuation 15235598 · Aug 12, 2016
Continuation 15166241 · May 26, 2016
Continuation 15056553 · Feb 29, 2016
Continuation 14726921 · Jun 1, 2015
Continuation 13340126 · Dec 29, 2011
Related Publication 20230269555A1 · Aug 24, 2023
Cited By (8)
US 12,450,025 US 12,464,302 US 12,495,258 US 12,501,229 US 12,574,697 US 12,652,508 US 12,659,682 US 12,666,217