IP Library Granted Patent US 9,437,181
Granted Patent B2
US 9,437,181 · App. 14/452,938 · Granted Sep 6, 2016

Off-axis audio suppression in an automobile cabin

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,437,181
App. No.
14/452,938
Granted
Sep 6, 2016
Kind
B2
Abstract

The suppression of off-axis audio in an audio environment is provided. Off-axis audio may be considered audio that does not originate from a region of interest. The off-axis audio is suppressed by comparing a phase difference between signals from two microphones to a target slope of the phase difference between signals originating from the region of interest. The target slope can be adapted to allow the region of interest to move with the location of a human speaker such as a driver.

Claims (42)

1. A method of off-axis audio suppression in an audio environment comprising:

receiving first and second audio signals from first and second microphones positioned within the audio environment;

calculating a phase difference between the first and second audio signals;

calculating a direction error between the calculated phase difference and a target slope, the target slope defining a desired phase difference between signals from the first and second microphones corresponding to audio originating from a region of interest; and

processing the first and second audio signals based on the calculated direction error to determine if the first and second audio signals correspond to audio originating from the region of interest and to suppress off-axis audio relative to the positions of the first and second microphones and the region of interest and further comprising adjusting the target slope based on the calculated phase difference between the first and second audio signals to adapt the region of interest based on a location of a human speaker within the audio environment.

2. The method of claim 1 , wherein the first and second audio signals are frequency domain representations of a frame of audio received at the corresponding microphone over a time interval, and wherein the method is repeated for subsequent frames of audio.

3. The method of claim 2 , wherein processing the first and second audio signals comprises:

determining if the direction error is less than an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to voice audio originating from the region of interest; and

combining the first and second audio signals to enhance the frame of audio when the direction error is less than the on-axis threshold.

4. The method of claim 2 , wherein processing the first and second audio signals comprises:

determining if the direction error is greater than an off-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to noise audio or to voice audio originating from outside the region of interest; and

combining the first and second audio signals to suppress the frame of audio when the direction error is greater than the off-axis threshold.

5. The method of claim 2 , wherein processing the first and second audio signals comprises:

determining if the direction error is between an off-axis threshold and an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to a combination of voice audio originating from the region of interest and noise audio or to voice audio originating from outside the region of interest;

calculating a mixing mask as a function of frequency; and

combining the first and second audio signals using the mixing mask when the direction error is between than the off-axis threshold and the on-axis threshold.

6. An apparatus performing off-axis audio suppression in an audio environment comprising:

a processor and memory configuring the apparatus to provide:

a target slope stored in memory defining a desired phase difference between signals from first and second microphones corresponding to audio originating from a region of interest;

a target adaptation component adjusting the target slope based on a calculated phase difference between the first and second audio signals to adapt the region of interest based on a location of a human speaker within the audio environment

a source-locating component calculating a direction error between the target slope and a phase difference between first and second audio signals received from the first and second microphones; and

an audio mixer processing the first and second audio signals based on the calculated direction error to suppress off-axis audio relative to the positions of the first and second microphones and the region of interest.

7. The apparatus of claim 6 , further comprising a signal processing component to convert the first and second audio signals to frequency domain representations of a frame of audio received at the corresponding microphone over a time interval.

8. The apparatus of claim 7 , wherein the audio mixer determines if the direction error is less than an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to voice audio originating from the region of interest and combines the first and second audio signals to enhance the frame of audio when the direction error is less than the on-axis threshold.

9. The apparatus of claim 7 , wherein the audio mixer determines if the direction error is greater than an off-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to noise audio or to voice audio originating from outside the region of interest and combines the first and second audio signals to suppress the frame of audio when the direction error is greater than the off-axis threshold.

10. The apparatus of claim 7 , wherein the audio mixer determines if the direction error is between an off-axis threshold and an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to a combination of voice audio originating from the region of interest and noise audio or to voice audio originating from outside the region of interest and combines the first and second audio signals using a mixing mask calculated as a function of frequency when the direction error is between the off-axis threshold and the on-axis threshold.

11. A computer readable non-transitory memory containing instructions which when executed by a processor perform a method of off-axis audio suppression in an audio environment comprising:

receiving first and second audio signals from first and second microphones positioned within the audio environment;

calculating a phase difference between the first and second audio signals;

calculating a direction error between the calculated phase difference and a target slope, the target slope defining a desired phase difference between signals from the first and second microphones corresponding to audio originating from a region of interest; and

processing the first and second audio signals based on the calculated direction error to determine if the first and second audio signals correspond to audio originating from the region of interest and to suppress off-axis audio relative to the positions of the first and second microphones and the region of interest and further comprising adjusting the target slope based on the calculated phase difference between the first and second audio signals to adapt the region of interest based on a location of a human speaker within the audio environment.

12. The computer readable non-transitory memory of claim 11 , wherein the first and second audio signals are frequency domain representations of a frame of audio received at the corresponding microphone over a time interval, and wherein the method is repeated for subsequent frames of audio.

13. The computer readable non-transitory memory of claim 11 , wherein processing the first and second audio signals comprises:

determining if the direction error is less than an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to voice audio originating from the region of interest; and

combining the first and second audio signals to enhance the frame of audio when the direction error is less than the on-axis threshold.

14. The computer readable non-transitory memory of claim 11 , wherein processing the first and second audio signals comprises:

determining if the direction error is greater than an off-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to noise audio or to voice audio originating from outside the region of interest; and

combining the first and second audio signals to suppress the frame of audio when the direction error is greater than the off-axis threshold.

15. The computer readable non-transitory memory of claim 11 , wherein processing the first and second audio signals comprises:

determining if the direction error is between an off-axis threshold and an on-axis threshold, indicating that the frame of audio represented by the first and second audio signals corresponds to a combination of voice audio originating from the region of interest and noise audio or to voice audio originating from outside the region of interest;

calculating a mixing mask as a function of frequency; and

combining the first and second audio signals using the mixing mask when the direction error is between than the off-axis threshold and the on-axis threshold.

Assignments (8)
NUNC PRO TUNC ASSIGNMENT Recorded Jun 19, 2023
From: BLACKBERRY LIMITED
To: MALIKIE INNOVATIONS LIMITED
Reel/Frame 064271/0199 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 16, 2023
From: BLACKBERRY LIMITED
To: MALIKIE INNOVATIONS LIMITED
Reel/Frame 064104/0103 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 22, 2020
From: 2236008 ONTARIO INC.
To: BLACKBERRY LIMITED
Reel/Frame 053313/0315 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE CORPORATE IDENTIFIER INADVERTENTLY LISTED ON THE ASSIGNMENT AND COVERSHEET AS "LIMITED" PREVIOUSLY RECORDED ON REEL 035700 FRAME 0845. ASSIGNOR(S) HEREBY CONFIRMS THE IDENTIFIER SHOULD HAVE STATED "INC.". Recorded May 27, 2015
From: QNX SOFTWARE SYSTEMS LIMITED
To: 2236008 ONTARIO INC.
Reel/Frame 035785/0156 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 22, 2015
From: QNX SOFTWARE SYSTEMS LIMITED
To: 2236008 ONTARIO LIMITED
Reel/Frame 035700/0845 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 4, 2015
From: 8758271 CANADA INC.
To: 2236008 ONTARIO INC.
Reel/Frame 034902/0538 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 4, 2015
From: QNX SOFTWARE SYSTEMS LIMITED
To: 8758271 CANADA INC.
Reel/Frame 034895/0490 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 11, 2014
From: FALLAT, MARK RYAN; HETHERINGTON, PHILLIP ALAN; PERCY, MICHAEL ANDREW
To: QNX SOFTWARE SYSTEMS LIMITED
Reel/Frame 033505/0802 →