IP Library Granted Patent US 9,489,963
Granted Patent B2
US 9,489,963 · App. 14/658,873 · Granted Nov 8, 2016

Correlation-based two microphone algorithm for noise reduction in reverberation

Inventors: Nima Yousefian Jazi (Rochester Hills, MI); Rogerio Guedes Alves (Macomb Township, MI)
Assignee: QUALCOMM TECHNOLOGIES INTERNATIONAL, LTD.
G10L21/0208G10L21/0316G10L2021/02165
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,489,963
App. No.
14/658,873
Granted
Nov 8, 2016
Kind
B2
Abstract

Embodiments are directed towards providing speech enhancement of audio signals from a target source and noise reduction of audio signals from a noise source. A coherence between a first audio signal from a first microphone and a second audio signal from a second microphone may be determined. A first gain function may be determined based on real components of a coherence function, wherein the real components include coefficients based on the previously determined coherence. A second gain function may be determined based on imaginary components of the coherence function. And a third gain function may be determined based on a relationship between a real component of the coherence function and a threshold range. An enhanced audio signal may be generated by applying a combination of the first gain function, the second gain function, and the third gain function to the first audio signal.

Claims (40)

1. A method to provide speech enhancement of audio signals from a target source and noise reduction of audio signals from a noise source, comprising:

determining a coherence function between a first audio signal from a first microphone and a second audio signal from a second microphone;

determining a first gain function based on real components of the coherence function;

determining a second gain function based on imaginary components of the coherence function;

determining a third gain function based on a relationship between the real components of the coherence function and a threshold range;

determining a final gain function based on the first gain function, the second gain function, and the third gain function; and

generating an enhanced audio signal by applying final gain function to the first audio signal.

2. The method of claim 1 , wherein the third gain function is a small constant value when the real component of the coherence function is outside of the threshold range and one when the real component of the coherence function is inside of the threshold range.

3. The method of claim 1 , wherein the first gain function, the second gain function, and the third gain function are determined independent of each other.

4. The method of claim 1 , wherein the final gain function is a product of the first gain function, the second gain function, and the third gain function raised to a power.

5. The method of claim 1 , wherein the first gain function and the second gain function are based on differences between values of the coherence function and values of the coherence function expected for a high signal-to-noise ratio.

6. The method of claim 5 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined using a direct-to-reverberant energy ratio.

7. The method of claim 6 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined further utilizing an angle of incidence of the target source.

8. A network computer to provide speech enhancement of audio signals from a target source and noise reduction of audio signals from a noise source, comprising:

a memory for storing at least instructions; and

a processor that executes the instructions to perform actions, including:

determining a coherence function between a first audio signal from a first microphone and a second audio signal from a second microphone;

determining a first gain function based on real components of the coherence function;

determining a second gain function based on imaginary components of the coherence function;

determining a third gain function based on a relationship between the real component of the coherence function and a threshold range;

determining a final gain function based on the first gain function, the second gain function, and the third gain function; and

generating an enhanced audio signal by applying the final gain function to the first audio signal.

9. The network computer of claim 8 , wherein the third gain function is a small constant value when the real component of the coherence function is outside of the threshold range and one when the real component of the coherence function is inside of the threshold range.

10. The network computer of claim 8 , wherein the first gain function, the second gain function, and the third gain function are determined independent of each other.

11. The network computer of claim 8 , wherein the final gain function is a product of the first gain function, the second gain function, and the third gain function raised to a power.

12. The network computer of claim 8 , wherein the first gain function and the second gain function are based on differences between values of the coherence function and values of the coherence function expected for a high signal-to-noise ratio.

13. The network computer of claim 12 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined using a direct-to-reverberant energy ratio.

14. The network computer of claim 13 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined further utilizing an angle of incidence of the target source.

15. A processor readable non-transitory storage media that includes instructions to provide speech enhancement of audio signals from a target source and noise reduction of audio signals from a noise source, wherein execution of the instructions by a processor performs actions, comprising:

determining a coherence function between a first audio signal from a first microphone and a second audio signal from a second microphone;

determining a first gain function based on real components of the coherence function;

determining a second gain function based on imaginary components of the coherence function;

determining a third gain function based on a relationship between the real component of the coherence function and a threshold range;

determining a final gain function based on the first gain function, the second gain function, and the third gain function; and

generating an enhanced audio signal by applying the final gain function to the first audio signal.

16. The media of claim 15 , wherein the third gain function is a small constant value when the real component of the coherence function is outside of the threshold range and one when the real component of the coherence function is inside of the threshold range.

17. The media of claim 15 , wherein the final gain function is a product of the first gain function, the second gain function, and the third gain function raised to a power.

18. The media of claim 15 , wherein the first gain function and the second gain function are based on differences between values of the coherence function and values of the coherence function expected for a high signal-to-noise ratio.

19. The media of claim 18 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined using a direct-to-reverberant energy ratio.

20. The media of claim 19 , wherein the values of the coherence function expected for a high signal-to-noise ratio are determined further utilizing an angle of incidence of the target source.

Assignments (4)
CHANGE OF NAME Recorded Feb 29, 2016
From: CAMBRIDGE SILICON RADIO LIMITED
To: QUALCOMM TECHNOLOGIES INTERNATIONAL, LTD.
Reel/Frame 037853/0185 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 13, 2016
From: JAZI, NIMA YOUSEFIAN; ALVES, ROGERIO GUEDES
To: CAMBRIDGE SILICON RADIO LIMITED
Reel/Frame 037482/0649 →
CHANGE OF NAME Recorded Jan 13, 2016
From: CAMBRIDGE SILICON RADIO LIMITED
To: QUALCOMM TECHNOLOGIES INTERNATIONAL, LTD.
Reel/Frame 037482/0667 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 17, 2015
From: JAZI, NIMA YOUSEFIAN; ALVES, ROGERIO GUEDES
To: CAMBRIDGE SILICON RADIO LIMITED
Reel/Frame 035178/0855 →
Continuity (1)
Related Publication 20160275966A1 · Sep 22, 2016