IP Library Granted Patent US 12,302,087
Granted Patent B2
US 12,302,087 · App. 17/986,877 · Granted May 13, 2025

Systems and methods for modifying room characteristics for spatial audio rendering over headphones

Inventors: Teck Chee Lee (Singapore, SG); Christopher Hummersone (Surrey, GB); Mark Anthony Davies (Middlesex, GB); Toh Onn Desmond Hii (Singapore, SG)
Assignee: Creative Technology Ltd.
H04S7/304H04S3/008H04S7/306H04S2400/01H04S2420/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,302,087
App. No.
17/986,877
Granted
May 13, 2025
Kind
B2
Abstract

An audio rendering system includes a processor that combines audio input signals with personalized spatial audio transfer functions having room responses. The personalized spatial audio transfer functions are selected from a database having a plurality of candidate transfer functions derived from in-ear microphone measurements for a plurality of individuals. Alternatively, the personalized transfer functions are derived from actual in-ear measurements of the listener. A room modification module allows the user to modify the personalized spatial audio transfer functions to substitute a different room or to modify the characteristics of the selected room without requiring additional in ear measurements. The module segments the selected transfer function into regions including one or more of direct; head and torso influenced; early reflection, and late reverberation regions. Extraction and modification operations are performed on one or more of the regions to alter the perceived sound.

Claims (28)

1. A method for generating modified Binaural Room Impulse Reponses (BRIRs) comprising:

generating a first Binaural Room Impulse Response (BRIR) based on a matching operation applied to extracted biometric properties;

segmenting the first BRIR into at least two regions by determining at least two regions of four regions that include a direct region, an early reflections region, a head and torso influenced region, and a late reverberation region for the first BRIR;

applying deconvolution to the direct region of the first BRIR to remove first loudspeaker effects from the direct region and create a deconvolved direct region of the first BRIR, wherein the first loudspeaker effects comprise a measured loudspeaker impulse response of a first loudspeaker;

convolving an impulse response for a target loudspeaker with the deconvolved direct region of the first BRIR to at least partially generate at least one modified region, wherein the at least one modified region includes the deconvolved direct region of the first BRIR corresponding to different loudspeaker acoustic properties from the first loudspeaker, wherein the at least one modified region is further at least partially generated by a modification operation comprising adapting different perceived loudspeaker acoustic properties from the first BRIR; and

combining the at least one modified region and any unmodified regions of the first BRIR to form a modified BRIR, wherein the at least one modified region corresponds to a removal of head-related transfer function (HRTF) effects on a listener relating to a second BRIR to obtain a result of the removal of the HRTF effects on the listener and substituting the result of the removal of the HRTF effects on the listener into the first BRIR.

2. The method as recited in claim 1 , wherein the generating the at least one modified region includes at least one of truncation, altering the slope of the decay rate, windowing, smoothing, ramping, and full room swapping.

3. The method as recited in claim 1 , wherein the modified BRIR is intended to mimic an audio processing performed by the target loudspeaker different from the first loudspeaker used for the second BRIR and the at least one modified region is generated from a corresponding region culled from an impulse response for the target loudspeaker; and convolving the impulse response for the second loudspeaker with the deconvolved direct region of the first BRIR.

4. The method recited in claim 1 , wherein the modified BRIR is intended to mimic an audio processing performed by a target room different from a first room used for the first BRIR and the at least one modified region is generated from a corresponding region derived from an impulse response for the target room, and wherein the segmenting includes determining at least one of the early reflections region and the late reverberation region of the first BRIR, and further comprising applying changes to at least one of the early reflections region and the late reverberations region to reflect sound characteristics of the target room.

5. The method as recited in claim 4 , wherein the first loudspeaker effects are deconvolved from the first BRIR and further comprising convolving the impulse response for the target loudspeaker with the deconvolved direct region of the first BRIR for the first loudspeaker.

6. The method as recited in claim 1 , wherein the first BRIR is a BRIR for an individual generated by accessing a database comprising a candidate pool of BRIRs for a population of individuals, each of the BRIRs in the candidate pool indexed according to the extracted biometric properties.

7. A system for modifying room or speaker characteristics for spatial audio rendering over headphones including a processor performing the steps of:

generating a first BRIR based on a matching operation applied to extracted biometric properties;

receiving the first BRIR corresponding to a first loudspeaker in a first room;

segmenting the first BRIR into at least two regions by determining at least two regions of four regions that include a direct region, an early reflections region, a head and torso influenced region, and a late reverberation region for the first BRIR;

applying deconvolution to the direct region of the first BRIR to remove first loudspeaker effects from the direct region and create a deconvolved direct region of the first BRIR, wherein the first loudspeaker effects comprise a measured loudspeaker impulse response of a first loudspeaker;

convolving a target loudspeaker response with the deconvolved direct region of the first BRIR to at least partially generate at least one modified region, wherein the at least one modified region includes the deconvolved direct region of the first BRIR corresponding to different loudspeaker acoustic properties from the first loudspeaker, wherein the at least one modified region is further at least partially generated by a modification operation comprising adapting different perceived loudspeaker acoustic properties from the first BRIR; and

combining the at least one modified region and the unmodified regions of the first BRIR to form a modified BRIR, wherein the at least one modified region corresponds to changed sound attributes for a loudspeaker-room-listener interrelationship.

8. The system as recited in claim 7 , wherein the modified BRIR is intended to mimic changes in the sound attributes for a loudspeaker-room-listener interrelationship that are derived from at least one of changes in loudspeaker composition, loudspeaker distance to room walls, loudspeaker distance to listener, room size, room and of dimensions, room construction, or room furnishings.

9. The system as recited in claim 7 , wherein the modified BRIR is synthesized to simulate non-room environments and further comprising:

using the processor to segment the first BRIR into regions that include the direct region, the early reflections region, the head and torso influenced region, and the late reverberation region; and

using ray tracing to synthesize the new reverberation to remove the late reverberation region and the early reflections region to create the non-room environment with no early reflection and late reverberation.

10. A method for generating modified spatial audio transfer functions comprising:

generating a first spatial audio transfer function customized for an individual by accessing a spatial audio transfer function indexed according to extracted image based properties, wherein the generating is based on a matching operation applied to the extracted image based properties;

segmenting the first spatial audio transfer function into at least two regions by determining at least two regions of four regions that include a direct region, an early reflections region, a head and torso influenced region, and a late reverberation region for the first spatial audio transfer function; and

performing a modification operation on at least one of the at least two regions to at least partially generate at least one modified region, wherein the modification operation comprises adapting different perceived loudspeaker acoustic properties or room acoustic properties from the first spatial audio transfer function and performing deconvolution to the direct region of the first spatial audio transfer function to remove first loudspeaker effects from the direct region and create a deconvolved direct region of the first spatial audio transfer function, wherein the first loudspeaker effects comprise a measured loudspeaker impulse response of a first loudspeaker;

convolving a target loudspeaker response with the deconvolved direct region of the first spatial audio transfer function, wherein the at least one modified region is further at least partially generated by a modification operation comprising adapting different perceived loudspeaker acoustic properties from the first loudspeaker from the first spatial audio transfer function; and

combining the at least one modified region and any unmodified regions of the first spatial audio transfer function of the at least two regions to form a modified spatial audio transfer function that provides virtualization of the target loudspeaker having different acoustic properties than the loudspeaker used as a basis for the first spatial audio transfer function.

Assignments (3)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 12, 2026
From: CREATIVE TECHNOLOGY LTD
To: ZEICA LABS PTE. LTD.
Reel/Frame 074063/0943 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 5, 2023
From: CREATIVE TECH (UK) LIMITED
To: CREATIVE TECHNOLOGY LTD.
Reel/Frame 062284/0774 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 5, 2023
From: LEE, TECK CHEE; HUMMERSONE, CHRISTOPHER; DAVIES, MARK ANTHONY; HII, TOH ONN DESMOND
To: CREATIVE TECHNOLOGY LTD.
Reel/Frame 062284/0966 →
Continuity (3)
Continuation 16653130 · Oct 15, 2019
Provisional Application 62750719 · Oct 25, 2018
Related Publication 20230072391A1 · Mar 9, 2023
References Cited (67)
US 5748758A · Menasco, Jr. · 1998 [cited by applicant]
US 7840019B2 · Slaney et al. · 2010 [cited by applicant]
US 7936887B2 · Smyth · 2011 [cited by applicant]
US 9030545B2 · Pedersen · 2015 [cited by applicant]
US 9544706B1 · Hirst · 2017 [cited by applicant]
US 9584946B1 · Lyren · 2017 [cited by examiner]
US 9591427B1 · Lyren · 2017 [cited by examiner]
US 10225682B1 · Lee et al. · 2019 [cited by applicant]
US 20030007648A1 · Currell · 2003 [cited by applicant]
US 20060045294A1 · Smyth · 2006 [cited by applicant]
US 20070270988A1 · Goldstein · 2007 [cited by applicant]
US 20080273708A1 · Sandgren · 2008 [cited by examiner]
US 20110268281A1 · Florencio et al. · 2011 [cited by applicant]
US 20120183161A1 · Agevik et al. · 2012 [cited by applicant]
US 20150073262A1 · Roth et al. · 2015 [cited by applicant]
US 20150223002A1 · Mehta · 2015 [cited by applicant]
US 20150312694A1 · Bilinski · 2015 [cited by examiner]
US 20150373477A1 · Norris et al. · 2015 [cited by applicant]
US 20160337779A1 · Davidson et al. · 2016 [cited by applicant]
US 20160379041A1 · Rhee et al. · 2016 [cited by applicant]
US 20170223478A1 · Jot et al. · 2017 [cited by applicant]
US 20170272890A1 · Oh et al. · 2017 [cited by applicant]
US 20180077514A1 · Lee et al. · 2018 [cited by applicant]
US 20180091920A1 · Family · 2018 [cited by applicant]
US 20180206059A1 · Fueg et al. · 2018 [cited by applicant]
US 20180218507A1 · Hyllus et al. · 2018 [cited by applicant]
US 20180249279A1 · Karapetyan et al. · 2018 [cited by applicant]
US 20200322727A1 · Smyth · 2020 [cited by examiner]
CN 105792090 · 2016 [cited by applicant]
CN 107835483 · 2018 [cited by applicant]
FR 3051951 · 2018 [cited by applicant]
JP 2008512015 · 2008 [cited by applicant]
WO 2006024850 · 2006 [cited by applicant]
WO 2015102920 · 2015 [cited by applicant]
WO 2017041922 · 2017 [cited by applicant]
WO 2017116308 · 2017 [cited by applicant]
WO 2017202634 · 2017 [cited by applicant]
WO 2017203011 · 2017 [cited by applicant]
WO WO2017203011A1 · 2017 [cited by examiner]
Japanese Patent Office, Japanese Office Action dated Apr. 12, 2021 in Application No. 2019-194536. [cited by applicant]
Japanese Patent Office, Japanese Search Report dated Dec. 23, 2020 in Application No. 2019-194536. [cited by applicant]
CNIPA, First Chinese Office Action dated Jan. 20, 2023 in Application No. 201911024774.7. [cited by applicant]
Meshram et al., “P-HRTF: Efficient Personalized HRTF Computation for High-Fidelity Spatial Sound,” 2014 IEEE ntemational Symposium on Mixed and Augmented Reality (ISMAR), 2014, pp. 53-61, Munich, Germany. [cited by applicant]
Dalena, Marco. “Selection of Head-Related Transfer Function through Ear Contour Matching for Personalized Binaural Rendering,” Politecnico Di Milano Master thesis for Master of Science in Computer Engineering, 2013, Mil… [cited by applicant]
Cootes et al., “Active Shape Models—Their Training and Application,” Computer Vision and Image Understanding, Jan. 1995, pp. 38-59, vol. 61, No. 1, Manchester, England. [cited by applicant]
John C. Middlebrooks, “Virtual localization improved by scaling nonindividualized external-ear transfer functions in frequency,” Journal of the Acoustical Society of America, Sep. 1999, pp. 1493-1510, vol. 106, No. 3, P… [cited by applicant]
Yukio Iwaya, “Individualization of head-related transfer functions with tournament-style listening test: Listening with other's ears,” Acoustical Science and Technology, 2006, vol. 27, Issue 6, Japan. [cited by applicant]
Slim Ghorbal, Theo Auclair, Catherine Soladie, & Renaud Segui ER, “Pinna Morphological Parameters influencing HRTF Sets,” Proceedings of the 20th International Conference on Digital Audio Effects (DAFx-17), Sep. 5-9, 20… [cited by applicant]
Slim Ghorbal, Renaud Seguier, & Xavier Bonjour, “Process of HRTF individualization by 30 statistical ear model,” Audio Engineering Society's 141st Convention e-Brief 283, Sep. 29, 2016-Oct. 2, 2016, Los Angeles, CA. [cited by applicant]
Robert P. Tame, Daniele Barchiesi, & Anssi Klapuri, “Headphone Virtualisation: Improved Localisation and Extemalisation of Non-individualised HRTFs by Cluster Analysis,” Audio Engineering Society's 133rd Convention Pape… [cited by applicant]
Zotkin, Dmitry et al., HRTF Personalization Using Anthropometric Measurements, 2003 IEEE Workshop on Applications of Signal Processing to Audio and Acouistics, Oct. 19-22, 2003, p. 157-160, New Paltz, NY. [cited by applicant]
AES Society, “Elevation Control in Binaural Rendering”, 140th Convention, Jun. 4-7, 2016, Paris, France, p. 1-4 (Year: 2016). [cited by applicant]
USPTO, Non-Final Office Action dated Nov. 27, 2019 in U.S. Appl. No. 16/653,130. [cited by applicant]
USPTO, Final Office Action dated Mar. 10, 2020 in U.S. Appl. No. 16/653,130. [cited by applicant]
USPTO, Non-Final Office Action dated Sep. 29, 2020 in U.S. Appl. No. 16/653,130. [cited by applicant]
USPTO, Final Office Action dated May 25, 2021 in U.S. Appl. No. 16/653,130. [cited by applicant]
USPTO, Non-Final Office Action dated Jan. 5, 2022 in U.S. Appl. No. 16/653,130. [cited by applicant]
USPTO, Notice of Allowance dated Jul. 19, 2022 in U.S. Appl. No. 16/653,130. [cited by applicant]
Korean Intellectual Propoerty Office, Korean Written Opinion dated Nov. 3, 2022 in Application No. 10-2019-01133368. [cited by applicant]
Japanese Intellectual Propoerty Office, Japanese Written Opinion dated Jul. 9, 2021 in Application No. 2019-194536. [cited by applicant]
Japanese Intellectual Propoerty Office, Japanese Written Opinion dated Jan. 25, 2022 in Application No. 2019-194536. [cited by applicant]
TIPO, Taiwan Office Action and Search Report dated Jul. 25, 2023 in Application No. 108137662. [cited by applicant]
EPO; European Summons to Attend Oral Proceedings dated Nov. 21, 2023 in European Application No. 19204434. [cited by applicant]
TIPO, Taiwan Office Action dated Sep. 6, 2023 in Application No. 108137662. [cited by applicant]
EPO; Summons to Attend Oral Proceedings dated Apr. 11, 2024 in European Application No. 19204434. [cited by applicant]
EPO; Extended European Search Report dated Feb. 27, 2020 in Application No. 192044345. [cited by applicant]
CNIPA; Chinese Search Report dated Jan. 20, 2023 in Application No. 201911024774. [cited by applicant]