IP Library › Granted Patent US 12,335,717
Granted Patent B2
US 12,335,717 · App. 18/108,494 · Granted Jun 17, 2025

Method and apparatus for spatial audio reproduction using directional room impulse responses interpolation

Inventors: Dae Young Jang (Daejeon, KR); Jiahong Zhao (Wollongong, AU); Xiguang Zheng (Wollongong, AU); Christian Ritz (Wollongong, AU)
Assignee: Electronics and Telecommunications Research Institute
H04S7/303H04R5/027H04S7/305H04S2400/11H04S2400/15H04S2420/11
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,335,717
App. No.
18/108,494
Granted
Jun 17, 2025
Kind
B2
Abstract

Disclosed are a method for spatial audio reproduction based on D-RIRs includes: selecting measurement points around a listener based on the location of the listener; calculating a D-RIR for the location of the listener based on D-RIRs for the measurement points around the listener; and reproducing spatial audio at the location of the listener based on the D-RIR at the location of the listener.

Claims (60)

1. A method for spatial audio reproduction based on directional room impulse responses, the method being performed by a processor that executes at least one instruction stored in memory, the method comprising:

selecting measurement points around a listener based on a location of the listener;

calculating a directional room impulse response (D-RIR) for the location of the listener based on directional room impulse responses (D-RIRs) for the measurement points around the listener; and

reproducing spatial audio at the location of the listener based on the D-RIR at the location of the listener.

2. The method of claim 1 , wherein the calculating comprises calculating the D-RIR for the location of the listener by interpolating the D-RIRs for the plurality of measurement points around the listener for the location of the listener.

3. The method of claim 2 , wherein the calculating comprises:

extracting an attenuation level, a delay, and a direction of arrival of reflection from each of the multiple D-RIRs previously measured for the plurality of measurement points around the listener; and

calculating the D-RIR for the location of the listener by interpolating the D-RIR information, extracted from the multiple D-RIRs previously measured for the plurality of measurement points around the listener, for the location of the listener.

4. The method of claim 1 , wherein the calculating comprises:

obtaining D-RIRs arriving at the measurement points around the listener from at least one sound source;

interpolating the D-RIRs, arriving at the measurement points around the listener from the at least one sound source, for the location of the listener; and

obtaining a D-RIR arriving at the location of the listener from the at least one sound source based on results of the interpolation.

5. The method of claim 4 , wherein:

the D-RIRs for the plurality of measurement points around the listener are signals obtained using ambisonic microphones; and

obtaining the D-RIRs arriving at the measurement points comprises:

detecting intervals of reflection components based on modeling using ambisonic microphones; and

calculating directions of arrival of reflections.

6. The method of claim 1 , wherein the calculating the D-RIR for the location of the listener comprises:

obtaining D-RIRs arriving at the measurement points around the listener from the at least one first sound source;

performing first interpolation on the D-RIRs, arriving at the measurement points around the listener from the at least one first sound source, for a location of a new second sound source;

performing second interpolation on D-RIRs, arriving at the measurement points around the listener from the second sound source obtained as a result of the first interpolation, for the location of the listener; and

obtaining a D-RIR, arriving at the location of the listener from the second sound source, based on results of the second interpolation.

7. The method of claim 1 , wherein the selecting comprises selecting two or more measurement points around the listener from among a plurality of virtual listener measurement points each having a D-RIR arriving from at least one sound source based on relative locations of the plurality of virtual listener measurement points and the location of the listener.

8. The method of claim 1 , wherein the selecting comprises selecting the measurement points around the listener from among a plurality of virtual listener measurement points that are distributed in a given space in a virtual reality environment having six degrees of freedom (6DoF) and each of the plurality of virtual listener measurement points have a D-RIR.

9. The method of claim 1 , wherein the calculating comprises calculating the D-RIR for the location of the listener using the D-RIRs previously obtained for the measurement points around the listener in a given space in a virtual reality environment having 6DoF.

10. A method for spatial audio reproduction based on directional room impulse responses, the method being performed by a processor that executes at least one instruction stored in memory, the method comprising:

selecting virtual listener locations as measurement points based on spatial information and at least one sound source location; and

obtaining directional room impulse responses (D-RIRs) for the virtual listener locations from the at least one sound source;

wherein the D-RIRs for the virtual listener locations from the at least one sound source comprise responses to sound arriving directly at the virtual listener locations from the at least one sound source, and responses to sound reflected within a given space and arriving at the virtual listener locations, based on information about the given space in a six degrees of freedom (6DoF) virtual reality environment.

11. An apparatus for spatial audio reproduction based on directional room impulse responses, the apparatus comprising:

memory configured to store at least one instruction; and

a processor configured to execute the at least one instruction;

wherein the processor is further configured to execute the at least one instruction to:

select measurement points around a listener based on the location of the listener;

calculate a directional room impulse response (D-RIR) for the location of the listener based on directional room impulse responses (D-RIRs) for the measurement points around the listener; and

reproduce spatial audio at the location of the listener based on the D-RIR for the location of the listener.

12. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to calculate the D-RIR for the location of the listener by interpolating the D-RIRs for the plurality of measurement points around the listener for the location of the listener.

13. The apparatus of claim 12 , wherein the processor is further configured to execute the at least one instruction to:

extract an attenuation level, a delay, and a direction of arrival of reflection from each of the multiple D-RIRs previously measured for the plurality of measurement points around the listener; and

calculate the D-RIR for the location of the listener by interpolating the D-RIR information, extracted from the multiple D-RIRs previously measured for the plurality of measurement points around the listener, for the location of the listener.

14. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to:

obtain D-RIRs arriving at the measurement points around the listener from at least one sound source;

interpolate the D-RIRs, arriving at the measurement points around the listener from the at least one sound source, for the location of the listener; and

obtain a D-RIR arriving at the location of the listener from the at least one sound source based on results of the interpolation.

15. The apparatus of claim 11 , wherein:

the D-RIRs for the plurality of measurement points around the listener are signals obtained using ambisonic microphones; and

the processor is further configured to execute the at least one instruction to obtain the D-RIRs arriving at the measurement points by detecting intervals of reflection components based on modeling using ambisonic microphones and calculating directions of arrival of reflections.

16. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to:

obtain D-RIRs arriving at the measurement points around the listener from the at least one first sound source;

perform first interpolation on the D-RIRs, arriving at the measurement points around the listener from the at least one first sound source, for a location of a new second sound source;

perform second interpolation on D-RIRs, arriving at the measurement points around the listener from the second sound source obtained as a result of the first interpolation, for the location of the listener; and

obtain a D-RIR, arriving at the location of the listener from the second sound source, based on results of the second interpolation.

17. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to select two or more measurement points around the listener from among a plurality of virtual listener measurement points each having a D-RIR arriving from at least one sound source based on relative locations the plurality of virtual listener measurement points and the location of the listener.

18. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to select the measurement points around the listener from among a listener measurement points that are plurality of virtual distributed in a given space in a virtual reality environment having six degrees of freedom (6DoF) and each of the plurality of virtual listener measurement points have a D-RIR.

19. The apparatus of claim 11 , wherein the processor is further configured to execute the at least one instruction to calculate the D-RIR for the location of the listener using the D-RIRs previously obtained for the measurement points around the listener in a given space in a virtual reality environment having 6DoF.

20. The apparatus of claim 11 , wherein:

the processor is further configured to execute the at least one instruction to:

select virtual listener locations as measurement points based on spatial information and at least one sound source location; and

obtain D-RIRs for the virtual listener locations from the at least one sound source; and

wherein the D-RIRs for the virtual listener locations from the at least one sound source comprise responses to sound arriving directly at the virtual listener locations from the at least one sound source, and responses to sound reflected within a given space and arriving at the virtual listener locations, based on information about the given space in a six degrees of freedom (6DoF) virtual reality environment.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 10, 2023
From: JANG, DAE YOUNG; ZHAO, JIAHONG; ZHENG, XIGUANG; RITZ, CHRISTIAN
To: ELECTRONICS AND TELECOMMUNICATIONS RESEARCH INSTITUTE
Reel/Frame 062663/0210 →
Priority Claims (2)
KR 10-2022-0017540 · Feb 10, 2022 · national
KR 10-2023-0017051 · Feb 8, 2023 · national
Continuity (1)
Related Publication 20230362572A1 · Nov 9, 2023
References Cited (7)
US 10602298B2 · Raghuvanshi · 2020 [cited by examiner]
US 11877143B2 · Raghuvanshi · 2024 [cited by examiner]
US 20190356999A1 · Raghuvanshi et al. · 2019 [cited by applicant]
US 20210029487A1 · Muynke et al. · 2021 [cited by applicant]
WO 2021237265A1 · 2021 [cited by applicant]
Kaspar Müller, et al., Auralization based on multi-perspective ambisonic room impulse responses, Published by EDP Sciences, 2020, Acta Acustica, https://doi.org/10.1051/aacus/2020024. [cited by applicant]
Antonello, Niccolò et al., Room impulse response interpolation using a sparse spatio-temporal representation of the sound field, IEEE Press, pp. 1-13, vol. 25, No. 10, Oct. 2017. [cited by applicant]