IP Library › Granted Patent US 12,219,344
Granted Patent B2
US 12,219,344 · App. 17/485,052 · Granted Feb 4, 2025

Adaptive audio centering for head tracking in spatial audio applications

Inventors: Xiaoyuan Tu (Sunnyvale, CA); Alexander Singh Alvarado (San Jose, CA)
Assignee: Apple Inc.
H04S7/304G06F3/012H04R1/32H04R5/02H04S2400/05H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,219,344
App. No.
17/485,052
Granted
Feb 4, 2025
Kind
B2
Abstract

Embodiments are disclosed for adaptive audio centering for head tracking in spatial audio applications. In an embodiment, a method comprises: obtaining first motion data from an auxiliary device communicatively coupled to a source device, the source device configured to provide spatial audio content and the auxiliary device configured to playback the spatial audio content; obtaining second motion data from one or more motion sensors of the source device; determining whether the source device and auxiliary device are in a period of mutual quiescence based on the first and second motion data; in accordance with determining that the source device and the auxiliary device are in a period of mutual quiescence, re-centering the spatial audio in a three-dimensional virtual auditory space; and rendering the 3D virtual auditory space for playback on the auxiliary device.

Claims (86)

1. A method comprising:

centering, using one or more processors, spatial audio in a three-dimensional (3D) virtual audio space;

obtaining, using the one or more processors, first sensor data from an auxiliary device communicatively coupled to a source device, the source device configured to provide the spatial audio and the auxiliary device configured to playback the spatial audio in the 3D virtual audio space;

obtaining, using the one or more processors, second sensor data from one or more motion sensors of the source device;

determining, using the one or more processors, whether the source device and auxiliary device are in a period of mutual quiescence based on the first and second sensor data;

in accordance with determining that the source device and the auxiliary device are in a period of mutual quiescence, re-centering the spatial audio in the 3D virtual auditory space following the period of mutual quiescence; and

rendering, using the one or more processors, the re-centered spatial audio in the 3D virtual auditory space.

2. The method of claim 1 , wherein the spatial audio is re-centered by zeroing out a correction angle at a rate determined by a size of the correction angle.

3. The method of claim 1 , wherein the mutual quiescence is static mutual quiescence, where both the source device and the auxiliary device are static.

4. The method of claim 1 , wherein the mutual quiescence is correlated mutual quiescence, where the first sensor data and the second sensor data are correlated.

5. The method of claim 1 , wherein the spatial audio is re-centered when the period of mutual quiescence exceeds a threshold time.

6. The method of claim 5 , further comprising:

during the period of mutual quiescence:

accumulating, using a first timer, a mutual quiescence time;

detecting, using the one or more processors, a disturbance based on the first sensor data or the second sensor data;

responsive to the detection, accumulating, using a second timer, a disturbance time;

determining, using the one or more processors, that the disturbance has ended based on the first sensor data or the second sensor data;

subtracting, using the one or more processors, the disturbance time from the mutual quiescence time;

resuming, using the one or more processors, accumulating the mutual quiescence time;

determining, using the one or more processors, whether the mutual quiescence time exceeds the threshold time; and

in accordance with the mutual quiescence time exceeding the threshold time, re-centering the spatial audio in the 3D virtual auditory space.

7. The method of claim 6 , further comprising:

during the disturbance, incrementing the second timer and decrementing the first timer; and

after the disturbance has ended, decrementing the second time and incrementing the first timer.

8. The method of claim 6 , wherein the first sensor data and the second sensor data include acceleration data and rotation rate data, and detecting a disturbance based on the first sensor data or the second sensor data, further comprises:

determining whether a source device maximum, low-pass filtered, acceleration data or an auxiliary device maximum, low-pass filtered, acceleration data meets a first threshold;

determining whether a source device maximum acceleration data or an auxiliary device maximum acceleration data meets a second threshold;

determining whether a source device maximum rotation rate or an auxiliary device maximum rotation rate meets a third threshold;

determining whether a source device maximum average rotation rate or an auxiliary device maximum average rotation rate meets a fourth threshold; and

in accordance with the first, second, third and fourth thresholds being met,

determining that a disturbance is detected.

9. The method of claim 1 , further comprising:

during the period of mutual quiescence:

accumulating, using a first timer, a first mutual quiescence time;

detecting a disturbance based on the first sensor data or the second sensor data;

responsive to the detection, accumulating, using a second timer, a disturbance time;

determining that disturbance time has exceeded a threshold time; and

resetting the first timer and the second timer to zero.

10. The method of claim 6 , further comprising:

during the period of mutual quiescence:

determining that a visual anchor was used to re-center the spatial audio; and

resetting the mutual quiescence time to zero.

11. A system comprising:

one or more processors;

memory storing instructions that when executed by the one or more processors cause the one or more processors to perform operations comprising:

centering spatial audio in a three-dimensional (3D) virtual audio space;

obtaining first sensor data from an auxiliary device communicatively coupled to a source device, the source device configured to provide the spatial audio content and the auxiliary device configured to playback the spatial audio in the 3D virtual audio space;

obtaining second sensor data from one or more motion sensors of the source device;

determining whether the source device and auxiliary device are in a period of mutual quiescence based on the first and second sensor data;

in accordance with determining that the source device and the auxiliary device are in a period of mutual quiescence, re-centering the spatial audio in the 3D virtual auditory space following the period of mutual quiescence; and

rendering the re-centered spatial audio in the 3D virtual auditory space.

12. The system of claim 11 , wherein the spatial audio is re-centered by zeroing out a correction angle at a rate determined by a size of the correction angle.

13. The system of claim 11 , wherein the mutual quiescence is static mutual quiescence, where both the source device and the auxiliary device are static.

14. The system of claim 11 , wherein the mutual quiescence is correlated mutual quiescence, where the first sensor and the second sensor are correlated.

15. The system of claim 11 , wherein the spatial audio is re-centered when the period of mutual quiescence exceeds a threshold time.

16. The system of claim 15 , the operations further comprising:

during the period of mutual quiescence:

accumulating, using a first timer, a mutual quiescence time;

detecting a disturbance based on the first sensor data or the second sensor data;

responsive to the detection, accumulating, using a second timer, a disturbance time;

determining that the disturbance has ended based on the first sensor data or the second sensor data;

subtracting the disturbance time from the mutual quiescence time;

resuming accumulating the mutual quiescence time;

determining whether the mutual quiescence time exceeds the threshold time; and

in accordance with the mutual quiescence time exceeding the threshold time, re-centering the spatial audio in the 3D virtual auditory space.

17. The system of claim 16 , the operations further comprising:

during the disturbance, incrementing the second timer and decrementing the first timer; and

after the disturbance has ended, decrementing the second time and incrementing the first timer.

18. The system of claim 16 , wherein the first sensor data and the second sensor data include acceleration data and rotation rate data, and detecting a disturbance based on the first sensor data or the second sensor data, further comprises:

determining whether a source device maximum, low-pass filtered, acceleration data or an auxiliary device maximum, low-pass filtered, acceleration data meets a first threshold;

determining whether a source device maximum acceleration data or an auxiliary device maximum acceleration data meets a second threshold;

determining whether a source device maximum rotation rate or an auxiliary device maximum rotation rate meets a third threshold;

determining whether a source device maximum average rotation rate or an auxiliary device maximum average rotation rate meets a fourth threshold; and

in accordance with the first, second, third and fourth thresholds being met,

determining that a disturbance is detected.

19. The system of claim 11 , the operations further comprising:

during the period of mutual quiescence:

accumulating, using a first timer, a first mutual quiescence time;

detecting a disturbance based on the first sensor data or the second sensor data;

responsive to the detection, accumulating, using a second timer, a disturbance time;

determining that disturbance time has exceeded a threshold time; and

resetting the first timer and the second timer to zero.

20. The system of claim 16 , the operations further comprising:

during the period of mutual quiescence:

determining that a visual anchor was used to re-center the spatial audio; and

resetting the mutual quiescence time to zero.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 18, 2021
From: TU, XIAOYUAN; SINGH ALVARADO, ALEXANDER
To: APPLE INC.
Reel/Frame 057819/0867 →
Continuity (2)
Provisional Application 63083846 · Sep 25, 2020
Related Publication 20220103965A1 · Mar 31, 2022
References Cited (56)
US 9142062B2 · Maciocci et al. · 2015 [cited by applicant]
US 9459692B1 · Li · 2016 [cited by applicant]
US 10169917B2 · Chen · 2019 [cited by examiner]
US 10339078B2 · Nair et al. · 2019 [cited by applicant]
US 11582573B2 · Tu et al. · 2023 [cited by applicant]
US 11586280B2 · Kriminger et al. · 2023 [cited by applicant]
US 11589183B2 · Tu et al. · 2023 [cited by applicant]
US 11647352B2 · Tam et al. · 2023 [cited by applicant]
US 11675423B2 · Akgul et al. · 2023 [cited by applicant]
US 20050281410A1 · Grosvenor et al. · 2005 [cited by applicant]
US 20120050493A1 · Ernst et al. · 2012 [cited by applicant]
US 20140153751A1 · Wells · 2014 [cited by applicant]
US 20150081061A1 · Aibara et al. · 2015 [cited by applicant]
US 20150193014A1 · Yamada · 2015 [cited by applicant]
US 20150302720A1 · Zhang et al. · 2015 [cited by applicant]
US 20160119731A1 · Lester, III · 2016 [cited by applicant]
US 20160262608A1 · Krueger · 2016 [cited by examiner]
US 20160269849A1 · Riggs et al. · 2016 [cited by applicant]
US 20170188895A1 · Nathan · 2017 [cited by applicant]
US 20170295446A1 · Thagadur Shivappa · 2017 [cited by applicant]
US 20180091923A1 · Satongar et al. · 2018 [cited by applicant]
US 20180125423A1 · Chang et al. · 2018 [cited by applicant]
US 20180176468A1 · Wang · 2018 [cited by examiner]
US 20180220253A1 · Kärkkäinen et al. · 2018 [cited by applicant]
US 20180242094A1 · Baek et al. · 2018 [cited by applicant]
US 20180332420A1 · Salume · 2018 [cited by examiner]
US 20180343534A1 · Norris · 2018 [cited by examiner]
US 20190121522A1 · Davis · 2019 [cited by examiner]
US 20190166435A1 · Crow · 2019 [cited by examiner]
US 20190224528A1 · Omid-Zohoor et al. · 2019 [cited by applicant]
US 20190313201A1 · Torres et al. · 2019 [cited by applicant]
US 20190313915A1 · Tzvieli et al. · 2019 [cited by applicant]
US 20190374161A1 · Ly et al. · 2019 [cited by applicant]
US 20190379995A1 · Lee et al. · 2019 [cited by applicant]
US 20200037097A1 · Torres et al. · 2020 [cited by applicant]
US 20200059749A1 · Casimiro Ericsson et al. · 2020 [cited by applicant]
US 20200169828A1 · Liu et al. · 2020 [cited by applicant]
US 20200323727A1 · Agrawal et al. · 2020 [cited by applicant]
US 20210044913A1 · Häussler · 2021 [cited by examiner]
US 20210064132A1 · Rubin et al. · 2021 [cited by applicant]
US 20210100480A1 · Kang et al. · 2021 [cited by applicant]
US 20210211825A1 · Joyner et al. · 2021 [cited by applicant]
US 20210394020A1 · Killen et al. · 2021 [cited by applicant]
US 20210396779A1 · Sarathy et al. · 2021 [cited by applicant]
US 20210397249A1 · Kriminger et al. · 2021 [cited by applicant]
US 20210397250A1 · Akgul et al. · 2021 [cited by applicant]
US 20210400414A1 · Tu et al. · 2021 [cited by applicant]
US 20210400418A1 · Tam et al. · 2021 [cited by applicant]
US 20210400419A1 · Turgut et al. · 2021 [cited by applicant]
US 20210400420A1 · Tam et al. · 2021 [cited by applicant]
US 20220103964A1 · Tu et al. · 2022 [cited by applicant]
CN 109146965 · 2019 [cited by applicant]
CN 109644317 · 2019 [cited by applicant]
CN 111149369 · 2020 [cited by applicant]
Jolliffe et al., “Principal component analysis: a review and recent developments,” Philosophical transactions of the royal society A: Mathematical, Physical and Engineering Sciences, Apr. 13, 2016, 374(2065):Feb. 2, 201… [cited by applicant]
Zhang et al., “Template Matching Based Motion Classification for Unsupervised Post-Stroke Rehabilitation,” Paper, Presented at Proceedings of the International Symposium on Bioelectronics and Bioinformations 2011, Suzho… [cited by applicant]