IP Library › Granted Patent US 12,342,136
Granted Patent B2
US 12,342,136 · App. 17/571,361 · Granted Jun 24, 2025

Signal processing methods and system for beam forming with improved signal to noise ratio

Inventor: Dietmar Ruwisch (Berlin, DE)
Assignee: Analog Devices International Unlimited Company
H04R3/005H04R1/406G10L2021/02166G10L21/0232H04R2410/01H04R2410/07
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,342,136
App. No.
17/571,361
Granted
Jun 24, 2025
Kind
B2
Abstract

A method and apparatus are provided for generating a directional output signal from sound received by at least two microphones arranged as microphone array. The method includes transforming the sound received by each of the microphones and represented by analog-to-digital converted time-domain signals into corresponding complex-valued frequency-domain microphone signals each having a frequency component value for each of a plurality of frequency components, calculating from the complex-valued frequency-domain microphone signals for a Beam Focus Direction a Beam Focus Spectrum by means of a Characteristic Function with values between zero and one, said Beam Focus Spectrum comprises, for each of the plurality of frequency components, a time-dependent, real-valued attenuation factor, multiplying, for each of the plurality of frequency components, the attenuation factor with the frequency component value of the complex-valued frequency-domain microphone signal to obtain a directional frequency component value, and forming a frequency-domain directional output signal.

Claims (56)

1. A method of generating a directional output signal from sound received by at least two microphones, said method comprising:

transforming analog-to-digital converted time-domain signals provided by respective ones of said at least two microphones into corresponding complex-valued frequency-domain microphone signals, with the at least two microphones closely spaced apart from one another and arranged as a microphone array within a vehicle, the analog-to-digital converted time-domain signals representing respective sounds received by the at least two microphones, wherein each one of the corresponding complex-valued frequency-domain microphone signals has a frequency component value for each frequency component of a plurality of frequency components;

calculating from the corresponding complex-valued frequency-domain microphone signals, for a Beam Focus Direction, a Beam Focus Spectrum by means of a Characteristic Function that controls a shape of the Beam Focus Direction and depends on frequency by means of a frequency-dependent exponent, said Beam Focus Spectrum comprises, for each frequency component of the plurality of frequency components, a time-dependent, real-valued attenuation factor, with the Beam Focus Direction pointing to a position of an occupant of the vehicle;

multiplying, for each frequency component of the plurality of frequency components, the time-dependent, real-valued attenuation factor with the frequency component value of a complex-valued frequency-domain microphone signal of one of said at least two microphones to obtain a directional frequency component value;

forming a frequency-domain directional output signal from directional frequency component values for each of the plurality of frequency components;

synthesizing, from the frequency-domain direction output signal, by means of inverse transformation, a time-domain directional output signal, with time-domain directional output signal having the Beam Focus Direction; and

causing output, via a loudspeaker, of the time-domain directional output signal.

2. The method of claim 1 , further comprising:

calculating a linear combination of respective microphone signals of said at least two microphones; and

multiplying, for each frequency component of the plurality of frequency components, the time-dependent, real-valued attenuation factor by a frequency component value of a complex-valued frequency-domain microphone signal resulting from the linear combination of the respective microphone signals.

3. The method of claim 2 , wherein, when the Beam Focus Spectrum for the Beam Focus Direction is provided, for each of the plurality of frequency components, Characteristic function values of different Beam Spectra are multiplied.

4. The method of claim 1 , wherein calculating the Beam Focus Spectrum further comprises:

calculating, for each of the plurality of frequency components, a real-valued Beam Spectra value from the corresponding complex-valued frequency-domain microphone signals for the Beam Focus Direction by means of predefined, microphone-specific, time-constant, complex-valued Transfer Functions; and

wherein, for each of the plurality of frequency components, said real-valued Beam Spectra value is an argument of said Characteristic Function, providing the Beam Focus Spectrum for said Beam Focus Direction.

5. The method of claim 4 , wherein said predefined, microphone-specific, time-constant, complex-valued Transfer Functions are calculated by means of an analytic formula incorporating a spatial distance of the at least two microphones, and a speed of sound.

6. The method of claim 1 , further comprising:

calculating, for each of the plurality of frequency components of the complex-valued frequency-domain microphone signal of at least one particular microphone of said at least two microphones, a respective tolerance compensated frequency component value by multiplying the frequency component value of the complex-valued frequency-domain microphone signal of said at least one particular microphone with a real-valued correction factor;

wherein, for each of the plurality of frequency components, said real-valued correction factor is calculated as temporal average of frequency component values of a plurality of real-valued Deviation Spectra;

wherein, for each of the plurality of frequency components, each frequency component value of a Deviation Spectrum of said plurality of real-valued Deviation Spectra is calculated by dividing a frequency component magnitude of a frequency-domain reference signal by the frequency component magnitude of the complex-valued frequency-domain microphone signal of said at least one particular microphone; and

wherein the Beam Focus Spectrum for a Beam Focus Direction is calculated from respective tolerance compensated frequency component values for said at least one particular microphone.

7. The method of claim 6 , for generating a wind-reduced directional output signal, further comprising:

calculating, for each of the plurality of frequency components, real-valued Wind Reduction Factors as minima of reciprocal frequency components of said plurality of real-valued Deviation Spectra; and

wherein, for each of the plurality of frequency components, said real-valued Wind Reduction Factors are multiplied with the frequency component values of said frequency-domain directional output signal, forming a frequency-domain wind-reduced directional output signal.

8. The method of claim 7 , wherein a time-domain wind-reduced directional output signal is synthesized from the frequency-domain wind-reduced directional output signal by means of inverse transformation.

9. The method of claim 6 , wherein said temporal average of the frequency component values is executed in response to said frequency component value of said Deviation Spectrum being greater than a predefined threshold value.

10. An apparatus comprising at least one processor configured to carry out the method of claim 1 .

11. An apparatus for generating a directional output signal from sound received by at least two microphones, said apparatus comprising at least one processor configured to:

transform analog-to-digital converted time-domain signals provided by respective ones of said at least two microphones into corresponding complex-valued frequency-domain microphone signals, with the at least two microphones closely spaced apart from one another and arranged as a microphone array within a vehicle, the analog-to-digital converted time-domain signals representing respective sounds received by the at least two microphones, wherein each one of the corresponding complex-valued frequency-domain microphone signals has a frequency component value for each frequency component of a plurality of frequency components;

calculate, from the corresponding complex-valued frequency-domain microphone signals, for a Beam Focus Direction, a Beam Focus Spectrum by means of a Characteristic Function that controls a shape of the Beam Focus Direction and depends on frequency by means of a frequency-dependent exponent, said Beam Focus Spectrum comprises, for each frequency component of the plurality of frequency components, a time-dependent, real-valued attenuation factor, with the Beam Focus Direction pointing to a position of an occupant of the vehicle;

multiply, for each frequency component of the plurality of frequency components, the time-dependent, real-valued attenuation factor with the frequency component value of a complex-valued frequency-domain microphone signal of one of said at least two microphones to obtain a directional frequency component value;

form a frequency-domain directional output signal from directional frequency component values for each of the plurality of frequency components; and

synthesizing, from the frequency-domain direction output signal, by means of inverse transformation, a time-domain directional output signal, with time-domain directional output signal having the Beam Focus Direction; and

causing output, via a loudspeaker, of the time-domain directional output signal.

12. The apparatus of claim 11 , further comprising said at least two microphones.

13. One or more non-transitory computer-readable media having instructions stored thereon, the instructions for generating a directional output signal from sound received by at least two microphones arranged as microphone array, wherein the instructions, in response to being executed, cause one or more processors to perform operations comprising:

transforming analog-to-digital converted time-domain signals provided by respective ones of said at least two microphones into corresponding complex-valued frequency-domain microphone signals, with the at least two microphones closely spaced apart from one another and arranged as a microphone array within a vehicle, the analog-to-digital converted time-domain signals representing respective sounds received by the at least two microphones, wherein each one of the corresponding complex-valued frequency-domain microphone signals has a frequency component value for each frequency component of a plurality of frequency components;

calculating, from the corresponding complex-valued frequency-domain microphone signals, for a Beam Focus Direction, a Beam Focus Spectrum by means of a Characteristic Function that controls a shape of the Beam Focus Direction and depends on frequency by means of a frequency-dependent exponent, said Beam Focus Spectrum comprises, for each frequency component of the plurality of frequency components, a time-dependent, real-valued attenuation factor, with the Beam Focus Direction pointing to a position of an occupant of the vehicle;

multiplying, for each frequency component of the plurality of frequency components, the time-dependent, real-valued attenuation factor with the frequency component value of a complex-valued frequency-domain microphone signal of one of said at least two microphones to obtain a directional frequency component value;

forming a frequency-domain directional output signal from directional frequency component values for each of the plurality of frequency components; and

synthesizing, from the frequency-domain direction output signal, by means of inverse transformation, a time-domain directional output signal, with time-domain directional output signal having the Beam Focus Direction; and

causing output, via a loudspeaker, of the time-domain directional output signal.

14. The one or more non-transitory computer-readable media of claim 13 , wherein the operations further comprise:

calculating a linear combination of respective microphone signals of said at least two microphones, and

multiplying, for each frequency component of the plurality of frequency components, the time-dependent, real-valued attenuation factor by the frequency component value of a complex-valued frequency-domain microphone signal resulting from the linear combination of the respective microphone signals.

15. The one or more non-transitory computer-readable media of claim 13 , wherein calculating the Beam Focus Spectrum further comprises:

calculating, for each of the plurality of frequency components, a real-valued Beam Spectra value from the corresponding complex-valued frequency-domain microphone signals for the Beam Focus Direction by means of predefined, microphone-specific, time-constant, complex-valued Transfer Functions; and

wherein, for each of the plurality of frequency components, said real-valued Beam Spectra value is an argument of said Characteristic Function, providing the Beam Focus Spectrum for said Beam Focus Direction.

16. The one or more non-transitory computer-readable media of claim 15 , wherein said predefined, microphone-specific, time-constant, complex-valued Transfer Functions are calculated by means of an analytic formula incorporating a spatial distance of the at least two microphones, and a speed of sound.

17. The one or more non-transitory computer-readable media of claim 13 , further comprising:

calculating, for each of the plurality of frequency components of the complex-valued frequency-domain microphone signal of at least one particular microphone of said at least two microphones, a respective tolerance compensated frequency component value by multiplying the frequency component value of the complex-valued frequency-domain microphone signal of said at least one particular microphone with a real-valued correction factor;

wherein, for each of the plurality of frequency components, said real-valued correction factor is calculated as temporal average of frequency component values of a plurality of real-valued Deviation Spectra;

wherein, for each of the plurality of frequency components, each frequency component value of a Deviation Spectrum of said plurality of real-valued Deviation Spectra is calculated by dividing a frequency component magnitude of a frequency-domain reference signal by the frequency component magnitude of the complex-valued frequency-domain microphone signal of said at least one particular microphone; and

wherein the Beam Focus Spectrum for a Beam Focus Direction is calculated from respective tolerance compensated frequency component values for said at least one particular microphone.

18. The one or more non-transitory computer-readable media of claim 17 , for generating a wind-reduced directional output signal, the operations further comprising:

calculating, for each of the plurality of frequency components, real-valued Wind Reduction Factors as minima of reciprocal frequency components of said plurality of real-valued Deviation Spectra; and

wherein, for each of the plurality of frequency components, said real-valued Wind Reduction Factors are multiplied with the frequency component values of said frequency-domain directional output signal, forming a frequency-domain wind-reduced directional output signal.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 24, 2022
From: RUWISCH, DIETMAR
To: ANALOG DEVICES INTERNATIONAL UNLIMITED COMPANY
Reel/Frame 059094/0594 →
Priority Claims (1)
EP 19185502 · Jul 10, 2019 · regional
Continuity (2)
Continuation PCTEP2020069599 · Jul 10, 2020
Related Publication 20220132241A1 · Apr 28, 2022
References Cited (50)
US 6683961B2 · Ruwisch · 2004 [cited by applicant]
US 6820053B1 · Ruwisch · 2004 [cited by applicant]
US 7327852B2 · Ruwisch · 2008 [cited by applicant]
US 7522737B2 · Solderits · 2009 [cited by applicant]
US 7885420B2 · Hetherington et al. · 2011 [cited by applicant]
US 8477964B2 · Ruwisch · 2013 [cited by applicant]
US 9330677B2 · Ruwisch · 2016 [cited by examiner]
US 9813833B1 · Vesa · 2017 [cited by applicant]
US 10506356B2 · Walser et al. · 2019 [cited by applicant]
US 20030179888A1 · Burnett et al. · 2003 [cited by applicant]
US 20050195988A1 · Tashev · 2005 [cited by examiner]
US 20070050161A1 · Taenzer et al. · 2007 [cited by applicant]
US 20070263847A1 · Konchitsky · 2007 [cited by applicant]
US 20080232607A1 · Tashev et al. · 2008 [cited by applicant]
US 20090097670A1 · Jeong · 2009 [cited by examiner]
US 20090136057A1 · Taenzer · 2009 [cited by examiner]
US 20110015931A1 · Kawahara et al. · 2011 [cited by applicant]
US 20110038489A1 · Visser et al. · 2011 [cited by applicant]
US 20110257967A1 · Every et al. · 2011 [cited by applicant]
US 20120121100A1 · Zhang et al. · 2012 [cited by applicant]
US 20130117016A1 · Ruwisch · 2013 [cited by applicant]
US 20140193000A1 · Ruwisch · 2014 [cited by examiner]
US 20150016629A1 · Kanamori et al. · 2015 [cited by applicant]
US 20160050488A1 · Matheja et al. · 2016 [cited by applicant]
US 20170337932A1 · Iyengar et al. · 2017 [cited by applicant]
US 20170347206A1 · Pedersen et al. · 2017 [cited by applicant]
US 20190364492A1 · Azizi et al. · 2019 [cited by applicant]
CN 1851806 · 2006 [cited by applicant]
DE 10043064 · 2002 [cited by applicant]
DE 19948308 · 2002 [cited by applicant]
DE 102004005998 · 2005 [cited by applicant]
DE 102010001935 · 2012 [cited by applicant]
EP 1571875A2 · 2005 [cited by applicant]
EP 2752848A1 · 2014 [cited by applicant]
JP 2007336232A · 2007 [cited by applicant]
WO 2003043374 · 2003 [cited by applicant]
WO 2006041735 · 2006 [cited by applicant]
International Search Report and Written Opinion mailed Dec. 2, 2020 in PCT/EP2020/069599, 13 pages. [cited by applicant]
Grimm et al., [cited by applicant]
Abstract in English for CN1851806, 1 page. [cited by applicant]
Extended European Search Report in EP19185498.3, mailed Jan. 20, 2020, 8 pages. [cited by applicant]
Extended European Search Report in EP19185502.2, mailed Jan. 8, 2020, 7 pages. [cited by applicant]
Extended European Search Report in EP19185507.1, mailed Jan. 28, 2020, 10 pages. [cited by applicant]
Extended European Search Report in EP19185513.9, mailed Dec. 2, 2019, 9 pages. [cited by applicant]
Extended European Search Report in EP19185514.7, mailed Jan. 24, 2020, 7 pages. [cited by applicant]
International Search Report and Written Opinion in PCT/EP2020/069592, mailed Sep. 23, 2020, 13 pages. [cited by applicant]
International Search Report and Written Opinion in PCT/EP2020/069607, mailed Nov. 12, 2020, 18 pages. [cited by applicant]
International Search Report and Written Opinion in PCT/EP2020/069617, mailed Aug. 17, 2020, 17 pages. [cited by applicant]
International Search Report and Written Opinion in PCT/EP2020/069621, mailed Sep. 18, 2020, 12 pages. [cited by applicant]
Takahashi et al., “Structure Selection Algorithm for Less Musical-Noise Generation in Integration Systems of Beamforming and Spectral Subtraction,” IEEE/SP 15th Workshop on Statistical Signal Processing, Aug. 31, 2009, … [cited by applicant]