IP Library Granted Patent US 12,335,698
Granted Patent B2
US 12,335,698 · App. 18/135,101 · Granted Jun 17, 2025

Audio denoising method and system

Inventors: Jinbo Zheng (Shenzhen, CN); Meilin Zhou (Shenzhen, CN); Fengyun Liao (Shenzhen, CN); Xin Qi (Shenzhen, CN)
Assignee: SHENZHEN SHOKZ CO., LTD.
H04R3/04
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,335,698
App. No.
18/135,101
Granted
Jun 17, 2025
Kind
B2
Abstract

In an audio denoising method and system provided in the present disclosure, a gain coefficient corresponding to each frequency unit may be generated based on a parameter related to a frequency by using a frequency of an audio signal as a unit, and gain processing is performed on each frequency unit separately by using the gain coefficient. The gain coefficient corresponding to a frequency unit including more valid audio signals may be larger, and a gain coefficient corresponding to a frequency unit including fewer valid audio signals may be smaller, so that more audio signals corresponding to frequency parts including more valid audio signals are preserved, while less audio signals corresponding to frequency parts including fewer valid audio signals are preserved. In this way, fidelity and intelligibility of an audio signal are improved while quality of the audio signal is improved and noise is reduced.

Claims (76)

1. An audio denoising system, comprising:

at least one storage medium storing at least one set of instruction for audio denoising; and

at least one processor in communication with the at least one storage medium, wherein during operation, the at least one processor executes the set of instructions to:

obtain a to-be-processed audio signal,

obtain at least one modulation parameter related to a frequency of the to-be-processed audio signal, wherein the at least one modulation parameter at least includes a plurality of frequency units of the to-be-processed audio signal,

based on the at least one modulation parameter and a preset gain function, obtain at least one gain coefficient corresponding to the at least one modulation parameter, wherein the at least one gain coefficient is in a negative correlation with the plurality of frequency units,

perform gain processing on the to-be-processed audio signal based on the at least one gain coefficient corresponding to the at least one modulation parameter to obtain a target audio signal, and

output the target audio signal.

2. The audio denoising system according to claim 1 , wherein the at least one modulation parameter further includes:

a plurality of signal-to-noise ratios corresponding to the plurality of frequency units.

3. The audio denoising system according to claim 1 , wherein the to-be-processed audio signal includes an audio signal obtained after an original audio signal is processed by using a first audio denoising algorithm.

4. The audio denoising system according to claim 3 , wherein the original audio signal includes at least one of:

a first audio signal output by a first-type microphone, a second audio signal output by a second-type microphone, or an audio signal obtained after fusion of the first audio signal and the second audio signal.

5. The audio denoising system according to claim 1 , wherein to perform the gain processing on the to-be-processed audio signal based on the at least one gain coefficient corresponding to the at least one modulation parameter to obtain the target audio signal, the at least one processor further executes the set of instructions to:

generate, based on the at least one modulation parameter and the preset gain function, the at least one gain coefficient corresponding to the at least one modulation parameter,

wherein the preset gain function includes a correlation between the at least one gain coefficient and the at least one modulation parameter; and

perform the gain processing on the to-be-processed audio signal based on the at least one gain coefficient to obtain the target audio signal.

6. The audio denoising system according to claim 5 , wherein the preset gain function is a monotonic function;

wherein the at least one gain coefficient is in a positive correlation with the plurality of signal-to-noise ratios.

7. The audio denoising system according to claim 6 , wherein

the at least one modulation parameter is the plurality of frequency units;

the preset gain function is a first gain function including a correlation between at least one first gain coefficient and the frequency;

the at least one gain coefficient is the at least one first gain coefficient; and

to generate, based on the at least one modulation parameter and the preset gain function, the at least one gain coefficient corresponding to the at least one modulation parameter, the at least one processor further executes the set of instructions to:

generate, based on the plurality of frequency units and the first gain function, a plurality of first gain coefficients corresponding to the plurality of frequency units.

8. The audio denoising system according to claim 6 , wherein

the at least one modulation parameter is the plurality of signal-to-noise ratios corresponding to the plurality of frequency units;

the preset gain function is a second gain function including a correlation between at least one second gain coefficient and the plurality of signal-to-noise ratios;

the at least one gain coefficient is the at least one second gain coefficient; and

to generate, based on the at least one modulation parameter and the preset gain function, the at least one gain coefficient corresponding to the at least one modulation parameter, the at least one processor further executes the set of instructions to:

generate, based on the plurality of signal-to-noise ratios and the second gain function, a plurality of second gain coefficients corresponding to the plurality of frequency units.

9. The audio denoising system according to claim 6 , wherein

the at least one modulation parameter is the plurality of frequency units and the plurality of signal-to-noise ratios corresponding to the plurality of frequency units;

the preset gain function is a third gain function including a correlation between at least one third gain coefficient and the frequency and the plurality of signal-to-noise ratios; the at least one gain coefficient is the at least one third gain coefficient; and

to generate, based on the at least one modulation parameter and the preset gain function, the at least one gain coefficient corresponding to the at least one modulation parameter, the at least one processor further executes the set of instructions to:

generate, based on the plurality of signal-to-noise ratios, the plurality of frequency units, and the third gain function, a plurality of third gain coefficients corresponding to the plurality of frequency units.

10. The audio denoising system according to claim 6 , wherein the preset gain function is a function based on a sigmoid function.

11. The audio denoising system according to claim 5 , wherein to perform the gain processing on the to-be-processed audio signal based on the at least one gain coefficient to obtain the target audio signal, the at least one processor further executes the set of instructions to:

perform the gain processing on each of the plurality of frequency units based on the at least one gain coefficient, to obtain the target audio signal.

12. The audio denoising system according to claim 1 , wherein to obtain the at least one modulation parameter related to the frequency of the to-be-processed audio signal, the at least one processor further executes the set of instructions to:

obtain at least one initial modulation parameter corresponding to the frequency of the to-be-processed audio signal; and

perform smoothing processing on a value of the at least one initial modulation parameter by using the frequency as a variable to obtain the at least one modulation parameter.

13. The audio denoising system according to claim 12 , wherein to perform the smoothing processing on the value of the at least one initial modulation parameter with the frequency as the variable, the at least one processor further executes the set of instructions to:

perform feature fusion processing on an initial signal-to-noise ratio corresponding to each of the plurality of frequency units and an initial signal-to-noise ratio corresponding to at least one frequency unit near a current frequency unit, to obtain a signal-to-noise ratio corresponding to the current frequency unit.

14. An audio denoising method, comprising:

obtaining a to-be-processed audio signal;

obtaining at least one modulation parameter related to a frequency of the to-be-processed audio signal;

based on the at least one modulation parameter and a preset gain function, obtaining at least one gain coefficient corresponding to the at least one modulation parameter, wherein the at least one gain coefficient is in a negative correlation with the plurality of frequency units;

performing gain processing on the to-be-processed audio signal based on at least one gain coefficient corresponding to the at least one modulation parameter to obtain a target audio signal; and

outputting the target audio signal,

wherein the at least one modulation parameter at least includes plurality of frequency units of the to-be-processed audio signal.

15. The audio denoising method according to claim 14 , wherein the performing of the gain processing on the to-be-processed audio signal based on the at least one gain coefficient corresponding to the at least one modulation parameter to obtain the target audio signal includes:

generating, based on the at least one modulation parameter and the preset gain function, the at least one gain coefficient corresponding to the at least one modulation parameter, wherein the preset gain function includes a correlation between the at least one gain coefficient and the at least one modulation parameter; and

performing the gain processing on the to-be-processed audio signal based on the at least one gain coefficient to obtain the target audio signal.

16. The audio denoising method according to claim 14 , wherein the preset gain function is a monotonic function;

wherein the at least one gain coefficient is in a positive correlation with the plurality of signal-to-noise ratios.

17. The audio denoising method according to claim 16 , wherein

the at least one modulation parameter is the plurality of frequency units;

the preset gain function is a first gain function including a correlation between at least one first gain coefficient and the frequency;

the at least one gain coefficient is the at least one first gain coefficient; and

the generating, based on the at least one modulation parameter and the preset gain function, of the at least one gain coefficient corresponding to the at least one modulation parameter includes:

generating, based on the plurality of frequency units and the first gain function, a plurality of first gain coefficients corresponding to the plurality of frequency units.

18. The audio denoising method according to claim 16 , the at least one modulation parameter further includes a plurality of signal-to-noise ratios corresponding to the plurality of frequency units, wherein

the at least one modulation parameter is the plurality of signal-to-noise ratios corresponding to the plurality of frequency units;

the preset gain function is a second gain function including a correlation between at least one second gain coefficient and the plurality of signal-to-noise ratios;

the at least one gain coefficient is the at least one second gain coefficient; and

the generating, based on the at least one modulation parameter and the preset gain function, of the at least one gain coefficient corresponding to the at least one modulation parameter includes:

generating, based on the plurality of signal-to-noise ratios and the second gain function, a plurality of second gain coefficients corresponding to the plurality of frequency units.

19. The audio denoising method according to claim 16 , the at least one modulation parameter further includes a plurality of signal-to-noise ratios corresponding to the plurality of frequency units, wherein

the at least one modulation parameter is the plurality of frequency units and the plurality of signal-to-noise ratios corresponding to the plurality of frequency units;

the preset gain function is a third gain function including a correlation between at least one third gain coefficient and the frequency and the plurality of signal-to-noise ratios;

the at least one gain coefficient is the at least one third gain coefficient; and

the generating, based on the at least one modulation parameter and the preset gain function, of the at least one gain coefficient corresponding to the at least one modulation parameter includes:

generating, based on the plurality of signal-to-noise ratios, the plurality of frequency units, and the third gain function, a plurality of third gain coefficients corresponding to the plurality of frequency units.

20. The audio denoising method according to claim 14 , wherein the obtaining of the at least one modulation parameter related to the frequency of the to-be-processed audio signal includes: obtaining at least one initial modulation parameter corresponding to the frequency of the to-be-processed audio signal, and performing smoothing processing on a value of the at least one initial modulation parameter by using the frequency as a variable to obtain the at least one modulation parameter,

wherein the performing of the smoothing processing on the value of the at least one initial modulation parameter with the frequency as the variable includes: performing feature fusion processing on an initial signal-to-noise ratio corresponding to each of the plurality of frequency units and an initial signal-to-noise ratio corresponding to at least one frequency unit near a current frequency unit, to obtain a signal-to-noise ratio corresponding to the current frequency unit.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 8, 2023
From: ZHENG, JINBO; ZHOU, MEILIN; LIAO, FENGYUN; QI, XIN
To: SHENZHEN SHOKZ CO., LTD.
Reel/Frame 063559/0118 →
Continuity (2)
Continuation PCTCN2020140214 · Dec 28, 2020
Related Publication 20230262390A1 · Aug 17, 2023
References Cited (32)
US 6766292B1 · Chandran et al. · 2004 [cited by applicant]
US 10636434B1 · Ramprashad · 2020 [cited by examiner]
US 20090119111A1 · Goto et al. · 2009 [cited by applicant]
US 20110081026A1 · Ramakrishnan et al. · 2011 [cited by applicant]
US 20110191101A1 · Uhle · 2011 [cited by examiner]
US 20160042746A1 · Fujieda · 2016 [cited by applicant]
US 20170345439A1 · Jensen et al. · 2017 [cited by applicant]
CN 104103278A · 2014 [cited by applicant]
CN 105810203A · 2016 [cited by applicant]
CN 107910011A · 2018 [cited by applicant]
CN 108630221A · 2018 [cited by applicant]
CN 109686347A · 2019 [cited by applicant]
CN 110634497A · 2019 [cited by applicant]
CN 111131947A · 2020 [cited by applicant]
CN 111554321A · 2020 [cited by applicant]
CN 111627455A · 2020 [cited by applicant]
EP 2149985A · 2010 [cited by applicant]
JP 2000047697A · 2000 [cited by applicant]
JP 2011259319A · 2011 [cited by applicant]
JP 2016038551A · 2016 [cited by applicant]
TW 1609366B · 2017 [cited by applicant]
WO 2007052612A1 · 2007 [cited by applicant]
WO 2019112468A · 2019 [cited by applicant]
WO 2019112468A1 · 2019 [cited by applicant]
WO 2019210605A1 · 2019 [cited by applicant]
International Search Report of PCT/CN2020/140214 (Sep. 28, 2021). [cited by applicant]
Jurgen Tchorz et al., ‘SNR Estimation Based on Amplitude Modulation Analysis With Applications to Noise Suppression’, IEEE Transactions on Speech and Audio Processing, vol. 11, No. 3, May 2003. [cited by applicant]
Matthias Dorbecker et al., ‘Combination of TwoChannel Spectral Subtraction and daptive Wener PostFiltering for Noise Reduction and reverberation’, EUSIPCO 1996, Sep. 1996. [cited by applicant]
Furuta et al. “A Study of Noise Suppression Method Based on Mutual Control of Spectral Subtraction and Spectral Amplitude Suppression,” IEICE Transactions on Information and Systems, Pt. 2, vol. JB7-D-2, Feb. 1, 2004, p… [cited by applicant]
Matthias Dorbecker et al., ‘Combination of TwoChannel Spectral Subtraction and daptive Wiener PostFiltering for Noise Reduction and reverberation’, EUSIPCO 1996, Sep. 1996. [cited by applicant]
Yumie Ono, “Back to Basics: Digital Signal Processing”, Biomedical Engineering, vol. 57, Nos. 2-3. Jun. 2019. [cited by applicant]
Szu-Chen Jou et al., “Adaptation for Soft Whisper Recognition Using a Throat Microphone”, Interspeech 2004—ICSLP. Oct. 2004. [cited by applicant]