IP Library Granted Patent US 12,260,873
Granted Patent B2
US 12,260,873 · App. 17/850,936 · Granted Mar 25, 2025

Method and apparatus of noise reduction, electronic device and readable storage medium

Inventor: Li Kang (Chongqing, CN)
Assignee: UNISOC (CHONGQING) TECHNOLOGIES CO., LTD.
G10L21/0232G10L2021/02165G10L2021/02166
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,260,873
App. No.
17/850,936
Granted
Mar 25, 2025
Kind
B2
Abstract

A method of noise reduction, which is applied to an electronic device. The electronic device includes a first sound collector and a second sound collector, installation positions of the first sound collector and the second sound collectors are different; the method includes: determining a desired sound signal and an interference sound signal based on a first sound signal collected by the first sound collector and a second sound signal collected by the second sound collector (S 102 ); obtaining a third sound signal by performing coherent noise elimination processing on the desired sound signal based on the interfering sound signal (S 103 ); and then obtaining a target sound signal by performing incoherent noise suppression processing on the third sound signal based on a probability of existence of a speech in the third sound signal (S 104 ).

Claims (90)

1. A method of noise reduction, wherein the method is applied to an electronic device, the electronic device comprises a first sound collector and a second sound collector, and installation positions of the first sound collector and the second sound collector are different, the method comprises:

acquiring a first sound signal collected by the first sound collector and a second sound signal collected by the second sound collector;

determining a desired sound signal and an interference sound signal based on the first sound signal and the second sound signal;

obtain a third sound signal by performing coherent noise elimination processing on the desired sound signal based on the interference sound signal; and

obtain a target sound signal by performing incoherent noise suppression processing on the third sound signal based on a probability of existence of a speech in the third sound signal;

wherein obtaining the third sound signal by performing the coherent noise elimination processing on the desired sound signal based on the interference sound signal comprises:

obtaining the third sound signal by calculating a difference between the desired sound signal and a product of the interfering sound signal and an adaptive filter coefficient;

wherein the (n+1)-th adaptive filter coefficient is obtained based on the n-th adaptive filter coefficient, an update step size, a variable update step size, a preset parameter and a conjugate correlation between the interference sound signal and the third sound signal, and the variable update step size changes with a change of a power ratio of the desired sound signal and the interference sound signal.

2. The method according to claim 1 , wherein determining the desired sound signal and the interference sound signal based on the first sound signal and the second sound signal comprises:

determining a first frequency domain signal of the first sound signal in a frequency domain, and a second frequency domain signal of the second sound signal in the frequency domain; and

obtain the desired sound signal and the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal.

3. The method according to claim 2 , wherein obtaining the desired sound signal and the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal comprises:

determining a delay duration between a collection moment of the first sound signal and a collection moment of the second sound signal; and

obtaining the desired sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a fixed beamforming filter, and obtaining the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a blocking matrix filter based on the delay duration.

4. The method according to claim 3 , wherein obtaining the desired sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using the fixed beamforming filter, and obtaining the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using the blocking matrix filter based on the delay duration, comprises:

calculating the desired sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and an exponential function related to the delay duration;

calculating the interfering sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and the exponential function related to the delay duration;

or,

calculating the desired sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and an exponential function related to the delay duration;

calculating the interfering sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and the exponential function related to the delay duration.

5. The method according to claim 1 , wherein determining the desired sound signal and the interference sound signal based on the first sound signal and the second sound signal comprises:

determining a first frequency domain signal of the first sound signal in a frequency domain and a second frequency domain signal of the second sound signal in the frequency domain; and

determining the first frequency domain signal as the desired sound signal, and determining the second frequency domain signal as the interference sound signal; or, determining the second frequency domain signal as the desired sound signal, and determining the first frequency domain signal as the interference sound signal.

6. The method according to claim 1 , wherein obtaining the target sound signal by performing the incoherent noise suppression processing on the third sound signal based on the probability of existence of the speech in the third sound signal comprises:

determining a smoothed power spectrum corresponding to the third sound signal;

determining a probability of absence of a priori speech corresponding to the third sound signal based on the smoothed power spectrum;

determining a probability of existence of a posteriori speech corresponding to the third sound signal based on the probability of absence of the priori speech;

determining an incoherent noise signal existing in the third sound signal by using the probability of existence of the posteriori speech, and determining an effective gain function corresponding to the third sound signal based on the incoherent noise signal; and

performing the incoherent noise suppression processing on the third sound signal by using the effective gain function.

7. An apparatus of noise reduction, wherein the apparatus is applied to an electronic device, the electronic device comprises a first sound collector and a second sound collector, installation positions of the first sound collector and the first sound collector are different; the apparatus comprises:

at least one processor; and

a memory communicatively connected with the at least one processor;

the at least one processor executes computer-executable instructions stored in the memory to cause the at least one processor to:

acquire a first sound signal collected by the first sound collector and a second sound signal collected by the second sound collector;

determine a desired sound signal and an interference sound signal based on the first sound signal and the second sound signal;

obtain a third sound signal by performing coherent noise elimination processing on the desired sound signal based on the interfering sound signal; and

obtain a target sound signal by performing incoherent noise suppression processing on the third sound signal based on a probability of existence of a speech in the third sound signal;

wherein the at least one processor is configured to:

obtain the third sound signal by calculating a difference between the desired sound signal and a product of the interfering sound signal and an adaptive filter coefficient;

wherein the (n+1)-th adaptive filter coefficient is obtained based on the n-th adaptive filter coefficient, an update step size, a variable update step size, a preset parameter and a conjugate correlation between the interference sound signal and the third sound signal, and the variable update step size changes with a change of a power ratio of the desired sound signal and the interference sound signal.

8. The apparatus according to claim 7 , wherein the at least one processor is further configured to:

determine a first frequency domain signal of the first sound signal in a frequency domain, and a second frequency domain signal of the second sound signal in the frequency domain; and

obtain the desired sound signal and the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal.

9. The apparatus according to claim 8 , wherein the at least one processor is further configured to:

determine a delay duration between a collection moment of the first sound signal and a collection moment of the second sound signal; and

obtain the desired sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a fixed beamforming filter, and obtain the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a blocking matrix filter based on the delay duration.

10. The apparatus according to claim 9 , wherein the at least one processor is further configured to:

calculate the desired sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and an exponential function related to the delay duration;

calculate the interfering sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and the exponential function related to the delay duration;

or,

calculate the desired sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and an exponential function related to the delay duration;

calculate the interfering sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and the exponential function related to the delay duration.

11. The apparatus according to claim 7 , wherein the at least one processor is further configured to:

determine a first frequency domain signal of the first sound signal in a frequency domain and a second frequency domain signal of the second sound signal in the frequency domain; and

determine the first frequency domain signal as the desired sound signal, and determine the second frequency domain signal as the interference sound signal; or, determine the second frequency domain signal as the desired sound signal, and determine the first frequency domain signal as the interference sound signal.

12. The apparatus according to claim 7 , wherein the at least one processor is further configured to:

determine a smoothed power spectrum corresponding to the third sound signal;

determine a probability of absence of a priori speech corresponding to the third sound signal based on the smoothed power spectrum;

determine a probability of existence of a posteriori speech corresponding to the third sound signal based on the probability of absence of the priori speech;

determine an incoherent noise signal existing in the third sound signal by using the probability of existence of the posteriori speech, and determine an effective gain function corresponding to the third sound signal based on the incoherent noise signal; and

perform the incoherent noise suppression processing on the third sound signal by using the effective gain function.

13. A non-transitory computer-readable storage medium, wherein computer-executed instructions are stored in the computer-readable storage medium, and when a processor executes the computer-executed instructions, the processor is enabled to:

acquire a first sound signal collected by a first sound collector and a second sound signal collected by a second sound collector;

determine a desired sound signal and an interference sound signal based on the first sound signal and the second sound signal;

obtain a third sound signal by performing coherent noise elimination processing on the desired sound signal based on the interference sound signal; and

obtain a target sound signal by performing incoherent noise suppression processing on the third sound signal based on a probability of existence of a speech in the third sound signal;

wherein when the processor executes the computer-executed instructions, the processor is enabled to:

obtain the third sound signal by calculating a difference between the desired sound signal and a product of the interfering sound signal and an adaptive filter coefficient;

wherein the (n+1)-th adaptive filter coefficient is obtained based on the n-th adaptive filter coefficient, an update step size, a variable update step size, a preset parameter and a conjugate correlation between the interference sound signal and the third sound signal, and the variable update step size changes with a change of a power ratio of the desired sound signal and the interference sound signal.

14. The non-transitory computer-readable storage medium according to claim 13 , wherein when the processor executes the computer-executed instructions, the processor is further enabled to:

determine a first frequency domain signal of the first sound signal in a frequency domain, and a second frequency domain signal of the second sound signal in the frequency domain; and

obtain the desired sound signal and the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal.

15. The non-transitory computer-readable storage medium according to claim 14 , wherein when the processor executes the computer-executed instructions, the processor is further enabled to:

determine a delay duration between a collection moment of the first sound signal and a collection moment of the second sound signal; and

obtain the desired sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a fixed beamforming filter, and obtain the interference sound signal by performing spatial filtering on the first frequency domain signal and the second frequency domain signal by using a blocking matrix filter based on the delay duration.

16. The non-transitory computer-readable storage medium according to claim 15 , wherein when the processor executes the computer-executed instructions, the processor is further enabled to:

calculate the desired sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and an exponential function related to the delay duration;

calculate the interfering sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and the exponential function related to the delay duration;

or,

calculate the desired sound signal based on a difference between the second frequency domain signal and a product of the first frequency domain signal and an exponential function related to the delay duration;

calculate the interfering sound signal based on a difference between the first frequency domain signal and a product of the second frequency domain signal and the exponential function related to the delay duration.

17. The non-transitory computer-readable storage medium according to claim 13 , wherein when the processor executes the computer-executed instructions, the processor is further enabled to:

determine a first frequency domain signal of the first sound signal in a frequency domain and a second frequency domain signal of the second sound signal in the frequency domain; and

determine the first frequency domain signal as the desired sound signal, and determine the second frequency domain signal as the interference sound signal; or, determine the second frequency domain signal as the desired sound signal, and determine the first frequency domain signal as the interference sound signal.

18. The non-transitory computer-readable storage medium according to claim 13 , wherein when the processor executes the computer-executed instructions, the processor is further enabled to:

determine a smoothed power spectrum corresponding to the third sound signal;

determine a probability of absence of a priori speech corresponding to the third sound signal based on the smoothed power spectrum;

determine a probability of existence of a posteriori speech corresponding to the third sound signal based on the probability of absence of the priori speech;

determine an incoherent noise signal existing in the third sound signal by using the probability of existence of the posteriori speech, and determine an effective gain function corresponding to the third sound signal based on the incoherent noise signal; and

perform the incoherent noise suppression processing on the third sound signal by using the effective gain function.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 18, 2025
From: KANG, LI
To: UNISOC (CHONGQING) TECHNOLOGIES CO., LTD.
Reel/Frame 070240/0742 →
Priority Claims (1)
CN 201911368908.7 · Dec 26, 2019 · national
Continuity (2)
Continuation PCTCN2020086639 · Apr 24, 2020
Related Publication 20220328058A1 · Oct 13, 2022
References Cited (22)
US 8175291B2 · Chan · 2012 [cited by examiner]
US 20100246851A1 · Buck · 2010 [cited by examiner]
US 20160192068A1 · Ng · 2016 [cited by examiner]
US 20200336833A1 · Cho · 2020 [cited by examiner]
CN 105590630A · 2016 [cited by applicant]
CN 106653043A · 2017 [cited by applicant]
CN 107993670A · 2018 [cited by applicant]
CN 109308904A · 2019 [cited by applicant]
CN 109473118A · 2019 [cited by applicant]
CN 109994120A · 2019 [cited by applicant]
WO 2009043066A1 · 2009 [cited by applicant]
WO WO2017002525A1 · 2017 [cited by examiner]
WO 2017132958A1 · 2017 [cited by applicant]
Cohen, Israel. “Two-Channel Signal Detection and Speech Enhancement Based on the Transient Beam-to-Reference Ratio” (Year: 2003). [cited by examiner]
Israel Cohen et al: “Two-channel signal detection and speech enhancement based on the transient beam-to-reference ratio”, 2003 IEEE International CONFE, vol. 5, Apr. 6, 2003 (Apr. 6, 2003), pp. V 233-V 236. [cited by applicant]
Junfeng Li et al: “Theoretical Analysis of Microphone Arrays With Postfiltering for Coherent and Incoherent Noise Suppression in Noisy Environments”, Hoboken, NJ : Wiley-Interscience, Sep. 12, 2005 (Sep. 12, 2005), pp. … [cited by applicant]
Extended European Search Report received in the corresponding European Application 20905296.8, mailed Dec. 14, 2022. [cited by applicant]
International Search Report and Written Opinion mailed in International Application PCT/CN2020/086639 on Sep. 30, 2020. [cited by applicant]
The first Office Action received in CN Application 201911368908.7 on Dec. 8, 2020. [cited by applicant]
The second Office Action received in CN Application 201911368908.7 on Jul. 5, 2021. [cited by applicant]
Robert J. Mcaulay et al., “Speech Enhancement Using a Soft-Decision Noise Suppression Filter”, issued on IEEE Transactions on Acoustics, Speech, and Signal Processing, vol. ASSP-28, No. 2, Apr. 1980. [cited by applicant]
Ni Zhong, “The Research of Speech Enhancement Method Based on Microphone Array”, a thesis submitted in partial satisfaction of the requirements for the degree of Master of Engineering in Electronic and Communication Eng… [cited by applicant]