IP Library Granted Patent US 12,470,883
Granted Patent B2
US 12,470,883 · App. 18/204,159 · Granted Nov 11, 2025

Apparatus, system and/or method for noise time-frequency masking based direction of arrival estimation for loudspeaker audio calibration

Inventors: Abdullah Kucuk (Novi, MI); Kadagattur Srinidhi (Novi, MI)
Assignee: HARMAN INTERNATIONAL INDUSTRIES, INCORPORATED
H04R29/001G10L21/0232H04R1/406H04R3/005G10L2021/02166H04R2430/21
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,470,883
App. No.
18/204,159
Granted
Nov 11, 2025
Kind
B2
Abstract

An audio system is provided that includes a first loudspeaker and a second loudspeaker. The first loudspeaker transmits an audio signal including audio data and a signature tone into a listening environment. The second loudspeaker includes at least one controller programmed to receive the audio signal including the audio data and the signature tone and to determine a direction of arrival (DOA) of the audio signal as received at the second loudspeaker based at least on the signature tone. The at least one controller is further programmed to perform a time frequency masking operation on the received audio signal prior to determining the DOA of the audio signal.

Claims (36)

1 . An audio system comprising:

a first loudspeaker to transmit an audio signal including audio data and a signature tone into a listening environment;

a second loudspeaker including:

at least one controller being programmed to:

receive the audio signal including the audio data and the signature tone;

determine a direction of arrival (DOA) of the audio signal as received at the second loudspeaker based at least on the signature tone;

perform a time frequency masking operation on the received audio signal prior to determining the DOA of the audio signal; and

perform cross correlation between the audio data and the signature tone to identify a frame that bounds the signature tone.

2 . The audio system of claim 1 , wherein the signature tone is transmitted at a predetermined length and at a predetermined frequency.

3 . The audio system of claim 1 , wherein the signature tone corresponds to at least an exponential sine sweep (ESS) signal having energy that is within a predetermined frequency range.

4 . The audio system of claim 3 , wherein the ESS signal is an inverse ESS signal.

5 . The audio system of claim 1 , wherein the second loudspeaker includes at least two microphones to receive the audio signal and a memory programmed to store the audio signal.

6 . The audio system of claim 5 , wherein the at least one controller is further programmed to perform the time frequency masking operation on the stored audio signal.

7 . The audio system of claim 6 , wherein the time frequency masking operation is one of an ideal binary mask (IBM), an ideal ratio mask (IRM), a complex ideal ratio mask (cIRM), and an optimal ratio mask (ORM).

8 . The audio system of claim 6 , wherein the at least one controller is further programmed to obtain a sample delay associated with the receipt of the signature tone on the audio signal between the at least two microphones.

9 . The audio system of claim 8 , wherein the sample delay is based at least on a sampling frequency, speed of sound, and a distance between the at least two microphones.

10 . The audio system of claim 8 , wherein the controller is further programmed to determine the DOA of the audio signal as received at the second loudspeaker further based at least on the sample delay.

11 . An audio system comprising:

a first loudspeaker including:

memory; and

at least one controller programmed to receive an audio signal including audio data and a signature tone from a second loudspeaker:

store the audio signal including the signature tone in the memory;

determine a direction of arrival (DOA) of the audio signal based at least on the signature tone;

perform a time frequency masking operation on the received audio signal prior to determining the DOA of the audio signal; and

perform cross correlation between the audio data and the signature tone to identify a frame that bounds the signature tone.

12 . The audio system of claim 11 , wherein the signature tone is transmitted from the second loudspeaker at a predetermined length and at a predetermined frequency.

13 . The audio system of claim 11 , wherein the signature tone corresponds to at least on exponential sine sweep (ESS) signal having energy that is within a predetermined frequency range.

14 . The audio system of claim 11 , wherein the first loudspeaker includes at least two microphones to receive the audio signal.

15 . The audio system of claim 14 , wherein the at least one controller is further programmed to perform a time frequency masking operation to reduce noise on the stored audio signal in the memory.

16 . The audio system of claim 15 , wherein the at least one controller is further programmed to obtain a sample delay associated with the receipt of the signature tone on the audio signal between the at least two microphones.

17 . The audio system of claim 16 , wherein the at least one controller is further programmed to determine the DOA of the audio signal further based at least on the sample delay.

18 . A computer-program product embodied in a non-transitory computer readable medium that is stored in memory and that is programmed and executable by at least one controller in an audio system, the computer-program product comprising instructions to:

receive, at a first loudspeaker, an audio signal from a second loudspeaker, the audio signal including audio data and a signature tone;

determine a direction of arrival (DOA) of the audio signal as received at the first loudspeaker based at least on the signature tone;

perform a time frequency masking operation on the received audio signal prior to determining the DOA of the audio signal; and

perform cross correlation between the audio data and the signature tone to identify a frame that bounds the signature tone.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 10, 2025
From: KUCUK, ABDULLAH; SRINIDHI, KADAGATTUR
To: HARMAN INTERNATIONAL INDUSTRIES, INCORPORATED
Reel/Frame 071659/0709 →
Continuity (1)
Related Publication 20240406646A1 · Dec 5, 2024
References Cited (27)
US 9270807B2 · Shivappa · 2016 [cited by examiner]
US 9360546B2 · Kim · 2016 [cited by examiner]
US 9794720B1 · Kadri · 2017 [cited by examiner]
US 10299060B2 · Satheesh · 2019 [cited by examiner]
US 10598543B1 · Mansour et al. · 2020 [cited by applicant]
US 11545172B1 · Chu · 2023 [cited by applicant]
US 11567162B2 · Chen · 2023 [cited by examiner]
US 12081949B2 · Vetter · 2024 [cited by examiner]
US 20050254662A1 · Blank et al. · 2005 [cited by applicant]
US 20130156198A1 · Kim et al. · 2013 [cited by applicant]
US 20170094437A1 · Kadri et al. · 2017 [cited by applicant]
US 20190228790A1 · Park et al. · 2019 [cited by applicant]
US 20190253801A1 · Arteaga et al. · 2019 [cited by applicant]
US 20210116555A1 · Zaccá · 2021 [cited by applicant]
US 20220291328A1 · Ozturk et al. · 2022 [cited by applicant]
US 20230040846A1 · Thomas et al. · 2023 [cited by applicant]
US 20230162750A1 · Murgai · 2023 [cited by examiner]
US 20240381046A1 · Thomas · 2024 [cited by examiner]
EP 2429214A2 · 2012 [cited by applicant]
WO 2022118072A1 · 2022 [cited by applicant]
Extended European Search Report dated Oct. 17, 2024 for European Patent Application No. 24175735.0, 8 pages. [cited by applicant]
Extended European Search Report dated Oct. 15, 2024 for European Patent Application No. 24175717.8, 9 pages. [cited by applicant]
Extended European Search Report dated Nov. 4, 2024 for European Patent Application No. 24175728.5, 9 pages. [cited by applicant]
Liang, S. et al., “The Optimal Ratio Time-Frequency Mask for Speech Separation in Terms of the Signal-to-Noise Ratio”, The Journal of the Acoustical Society of America 134, Oct. 16, 2013, 8 pgs. [cited by applicant]
Xia, S. et al., “Using Optimal Ratio Mask as Training Target for Supervised Speech Separation”, In 2017 Asia-Pacific Signal and Information Processing Association Annual Summit and Conference (APSIPA ASC), Sep. 11, 2017… [cited by applicant]
Knapp, C. et al., “The generalized correlation method for estimation of time delay”, IEEE transactions on acoustics, speech, and signal processing, Aug. 1976, 8 pgs. [cited by applicant]
Non-Final Office Action; related U.S. Appl. No. 18/204,150, filed May 31, 2023; date of mailing Mar. 24, 2025, 17 pgs. [cited by applicant]