IP Library Granted Patent US 9,224,404
Granted Patent B2
US 9,224,404 · App. 13/751,724 · Granted Dec 29, 2015

Dynamic audio processing parameters with automatic speech recognition

Inventor: Anthony Andrew Poliak (Lake Stevens, WA)
Assignee: 2236008 Ontario Inc.
G10L21/0208G10L15/30G10L15/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,224,404
App. No.
13/751,724
Granted
Dec 29, 2015
Kind
B2
Abstract

A communication system includes a front-end audio gateway or bridge and a hands-free device. An automatic speech recognition platform accessible to the hands-free device provides or makes available one or more preprocessing schemes and/or acoustic models to the front-end audio gateway or bridge. The preprocessing schemes or acoustic models can be identified by or provided before a connection is established between the front-end audio gateway and the automatic speech recognition platform, when a connection occurs between the front-end audio gateway and the automatic speech recognition platform, and/or during a speech recognition session.

Claims (19)

1. A communication system comprising:

an audio-gateway that converts input audio signals into a compatible form used by a receiving network;

a noise reduction module resident to the audio gateway configured to reduce in-vehicle echo and noise; and

a speech recognition engine remote from the audio gateway generates and transmits commands through a wireless network that cause the audio gateway to modify the audio gateway's noise reduction processing state in response to a recognized request at the audio gateway for an automated speech recognition;

where the noise reduction module applies knowledge of a reverberation time, knowledge of a plurality of microphone placement, and knowledge of driver's location stored in a data store to reduce the in-vehicle echo and noise.

2. A communication process comprising:

transferring a plurality of preprocessing schemes or acoustic models that can be implemented by an audio gateway through a short-range network;

comparing a spoken utterances to a grammar-based vocabulary to generate a recognition result and a confidence score at the audio gateway; and

selecting one of the plurality of preprocessing schemes or acoustic models by a automatic speech recognition process remote from the audio gateway based on the recognition result and the confidence score when the recognition result and the confidence score indicates a request for an automated speech recognition.

3. The communication process of claim 2 where the transfer of the plurality of preprocessing schemes or acoustic models occurs during a speech recognition session.

4. A communication system comprising:

an audio-gateway that converts input audio signals into a compatible form used by a receiving network;

a noise reduction module resident to the audio gateway configured to reduce in-vehicle echo and noise; and

a speech recognition engine remote from the audio gateway generates and transmits commands through a wireless network that cause the audio gateway to modify the audio gateway's noise reduction processing state in response to a request for an automated speech recognition

where the noise reduction module applies knowledge of a reverberation time, knowledge of a plurality of microphone placement, and knowledge of driver's location stored in a data store to reduce the in-vehicle echo and noise.

5. A communication process comprising:

transferring a plurality of noise or echo preprocessing schemes that are implemented by an audio gateway in response to commands received through a short-range network and the detection of a speech event at the audio gateway;

comparing a spoken utterances to a grammar-based vocabulary to generate a recognition result and a confidence score at the audio gateway; and

selecting one of the plurality of the nose or echo preprocessing schemes in response to an automatic speech recognition process remote from the audio gateway based on the recognition result and the confidence score when the recognition result and the confidence score indicates a request for automated speech recognition.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 22, 2020
From: 2236008 ONTARIO INC.
To: BLACKBERRY LIMITED
Reel/Frame 053313/0315 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 4, 2014
From: 8758271 CANADA INC.
To: 2236008 ONTARIO INC.
Reel/Frame 032607/0674 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 4, 2014
From: QNX SOFTWARE SYSTEMS LIMITED
To: 8758271 CANADA INC.
Reel/Frame 032607/0943 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 24, 2013
From: QNX SOFTWARE SYSTEMS, INC.
To: QNX SOFTWARE SYSTEMS LIMITED
Reel/Frame 030276/0885 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 2, 2013
From: POLIAK, ANTHONY ANDREW
To: QNX SOFTWARE SYSTEMS, INC.
Reel/Frame 030136/0858 →
Continuity (1)
Related Publication 20140214414A1 · Jul 31, 2014