IP Library › Granted Patent US 12,621,629
Granted Patent B2
US 12,621,629 · App. 18/597,643 · Granted May 5, 2026

HRTF determination using a headset and in-ear devices

Inventors: Andrew Francl (Redmond, WA); Tobias Daniel Kabzinski (Aachen, DE); Hao Lu (Redmond, WA); Antje Ihlefeld (Redmond, WA); William Owen Brimijoin, II (Kirkland, WA)
Assignee: Meta Platforms Technologies, LLC
H04S7/306G02B27/0093G02B27/0172G02B2027/0138G02B2027/014H04R2499/15H04S2400/11H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,621,629
App. No.
18/597,643
Granted
May 5, 2026
Kind
B2
Abstract

Techniques for determining personalized head-related transfer functions (HRTFs) using a head-mounted device and in-ear devices include: receiving, from a sensor array of the head-mounted device, a first sound signal associated with a sound from a sound source in a local environment of a user of the head-mounted device; determining that reverberation characteristics and spectral characteristics of the sound meet predetermined criteria based on the first sound signal; determining that the sound source is stationary within a time period; determining a relative location of the sound source with respect to the user; receiving, from an in-ear device in an ear of the user, a second sound signal associated with the sound from the sound source; and determining, based on at least the second sound signal, an HRTF or one or more parameters of the HRTF associated with the relative location of the sound source for the user.

Claims (50)

1 . A method comprising:

receiving, from a sensor array of a head-mounted device, a first sound signal associated with a sound from a sound source in a local environment of a user of the head-mounted device;

determining, based on the first sound signal, that reverberation characteristics and spectral characteristics of the sound meet predetermined criteria;

determining that the sound source is stationary within a time period;

determining a relative location of the sound source with respect to the user;

receiving, from an in-ear device in an ear of the user, a second sound signal associated with the sound from the sound source; and

determining, based on at least the second sound signal, a head-related transfer function (HRTF) or one or more parameters of the HRTF associated with the relative location of the sound source for the user, the determining of the HRTF including, at least in part, determining a reference signal based on the first sound signal and the determined relative location of the sound source.

2 . The method of claim 1 , wherein determining the relative location of the sound source with respect to the user includes determining an azimuth angle of the sound source, an elevation angle of the sound source, or a combination thereof with respect to the user.

3 . The method of claim 1 , wherein determining the relative location of the sound source with respect to the user includes:

determining a direction of arrival of the sound based on the first sound signal from the senor array and locations of two or more sensors in the sensor array;

determining the relative location of the sound source with respect to the user based on images captured by one or more cameras on the head-mounted device; or

a combination thereof.

4 . The method of claim 3 , wherein determining the relative location of the sound source with respect to the user includes determining a confidence level of the determined relative location of the sound source with respect to the user.

5 . The method of claim 1 , wherein determining the HRTF or the one or more parameters of the HRTF associated with the relative location of the sound source for the user further includes:

determining the HRTF or the one or more parameters of the HRTF based on a spectrum of the reference sound signal and a spectrum of the second sound signal.

6 . The method of claim 5 , wherein determining the reference sound signal includes beamforming in a direction of the relative location of the sound source based on the first sound signal.

7 . The method of claim 1 , further comprising determining, based on data from one or more position sensors of the head-mounted device, a relative position of the torso of the user with respect to the head of the user.

8 . The method of claim 1 , further comprising saving the HRTF or the one or more parameters of the HRTF and the relative location of the sound source to a data store that stores a plurality of HRTFs for the user.

9 . The method of claim 1 , wherein the reverberation characteristics and spectral characteristics of the sound include a signal-to-noise ratio, a frequency range, a reverberation level, a reverberation time, or a combination thereof.

10 . The method of claim 1 , further comprising generating a model or a look-up table for mapping the relative location of the sound source to the one or more parameters of the HRTF.

11 . The method of claim 1 , wherein the one or more parameters of the HRTF include parameters of one or more filters or frequency scaling factors for implementing the HRTF.

12 . The method of claim 1 , further comprising performing operations of the method of claim 1 iteratively to determine HRTFs or parameters of the HRTFs associated with a plurality of sound source directions with respect to the user.

13 . The method of claim 1 , wherein the time period is greater than 10 milliseconds.

14 . A system comprising:

an in-ear device configured to generate a first sound signal associated with a sound from a sound source in a local environment of a user; and

a head-mounted device comprising:

a sensor array configured to generate a second sound signal associated with the sound; and

an audio controller configured to:

determine, based on the second sound signal, that reverberation characteristics and spectral characteristics of the sound meet predetermined criteria;

determine that the sound source is stationary within a time period;

determine a relative location of the sound source with respect to the user; and

determine, based on at least the first sound signal, a head-related transfer function (HRTF) or one or more parameters of the HRTF associated with the relative location of the sound source for the user, the determining of the HRTF including, at least in part, determining a reference signal based on the first sound signal and the determined relative location of the sound source.

15 . The system of claim 14 , wherein the audio controller is configured to determine an azimuth angle of the sound source, an elevation angle of the sound source, or a combination thereof with respect to the user.

16 . The system of claim 14 , wherein the audio controller is configured to determine the relative location of the sound source with respect to the user by performing operations including:

determining a direction of arrival of the sound based on the first sound signal from the senor array and locations of two or more sensors in the sensor array;

determining the relative location of the sound source with respect to the user based on images captured by one or more cameras on the head-mounted device; or

a combination thereof.

17 . The system of claim 14 , wherein the audio controller is configured to determine the HRTF or the one or more parameters of the HRTF associated with the relative location of the sound source for the user by performing operations including:

determining the HRTF or the one or more parameters of the HRTF based on a spectrum of the reference sound signal and a spectrum of the second sound signal.

18 . The system of claim 17 , wherein the reference sound signal is a sound signal at a center of the head of the user determined by beamforming in a direction of the relative location of the sound source based on the second sound signal.

19 . The system of claim 14 , wherein the one or more parameters of the HRTF include parameters of one or more filters or frequency scaling factors for implementing the HRTF.

20 . A system comprising:

one or more processors; and

one or more processor-readable media storing instructions which, when executed by the one or more processors, cause the one or more processors to:

receive, from a sensor array of a head-mounted device, a first sound signal associated with a sound from a sound source in a local environment of a user of the head-mounted device;

determine, based on the first sound signal, that reverberation characteristics and spectral characteristics of the sound meet predetermined criteria;

determine that the sound source is stationary within a time period;

determine a relative location of the sound source with respect to the user;

receive, from an in-ear device in an ear of the user, a second sound signal associated with the sound from the sound source; and

determine, based on at least the second sound signal, a head-related transfer function (HRTF) or one or more parameters of the HRTF associated with the relative location of the sound source for the user, the determining of the HRTF including, at least in part, determining a reference signal based on the first sound signal and the determined relative location of the sound source.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 13, 2024
From: FRANCL, ANDREW; KABZINSKI, TOBIAS DANIEL; LU, HAO; IHLEFELD, ANTJE; BRIMIJOIN, WILLIAM OWEN, II
To: META PLATFORMS TECHNOLOGIES, LLC
Reel/Frame 068588/0001 →
Continuity (2)
Provisional Application 63488895 · Mar 7, 2023
Related Publication 20240305951A1 · Sep 12, 2024
References Cited (15)
US 9681250B2 · Luo et al. · 2017 [cited by applicant]
US 10798514B2 · Reijniers et al. · 2020 [cited by applicant]
US 11082794B2 · Alon et al. · 2021 [cited by applicant]
US 11146908B2 · Mohapatra et al. · 2021 [cited by applicant]
US 11190896B1 · Satongar et al. · 2021 [cited by applicant]
US 11457325B2 · Brimijoin, II et al. · 2022 [cited by applicant]
US 20210314720A1 · Khaleghimeybodi et al. · 2021 [cited by applicant]
US 20220182772A1 · Dodds · 2022 [cited by examiner]
US 20220342213A1 · Lu · 2022 [cited by examiner]
US 20230380739A1 · Ihlefeld · 2023 [cited by examiner]
Geronazzo M., et al., “A Minimal Personalization of Dynamic Binaural Synthesis with Mixed Structural Modeling and Scattering Delay Networks,” ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signa… [cited by applicant]
Middlebrooks J.C., “Individual Differences in External-Ear Transfer Functions Reduced by Scaling in Frequency,” The Journal of the Acoustical Society of America, Sep. 1, 1999, vol. 106, No. 03, pp. 1480-1492. [cited by applicant]
Middlebrooks J.C., “Virtual Localization Improved by Scaling Nonindividualized External-Ear Transfer Functions in Frequency,” The Journal of the Acoustical Society of America, Sep. 1, 1999, vol. 106, No. 03, pp. 1493-15… [cited by applicant]
International Preliminary Report on Patentability for International Application No. PCT/US2024/018835, mailed Sep. 18, 2025, 8 pages. [cited by applicant]
International Search Report and Written Opinion for International Application No. PCT/US2024/018835, mailed Jun. 17, 2024, 12 pages. [cited by applicant]