IP Library Granted Patent US 7,983,917
Granted Patent B2
US 7,983,917 · App. 12/608,544 · Granted Jul 19, 2011

Dynamic speech sharpening

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,983,917
App. No.
12/608,544
Granted
Jul 19, 2011
Kind
B2
Abstract

An enhanced system for speech interpretation is provided. The system may include receiving a user verbalization and generating one or more preliminary interpretations of the verbalization by identifying one or more phonemes in the verbalization. An acoustic grammar may be used to map the phonemes to syllables or words, and the acoustic grammar may include one or more linking elements to reduce a search space associated with the grammar. The preliminary interpretations may be subject to various post-processing techniques to sharpen accuracy of the preliminary interpretation. A heuristic model may assign weights to various parameters based on a context, a user profile, or other domain knowledge. A probable interpretation may be identified based on a confidence score for each of a set of candidate interpretations generated by the heuristic model. The model may be augmented or updated based on various information associated with the interpretation of the verbalization.

Claims (16)

1. A method for reducing a search space for a recognition grammar used when interpreting natural language speech utterances, the method comprising:

creating a phonemic representation of acoustic elements associated with an acoustic speech model, the acoustic elements including at least an unstressed central vowel and plurality of phonemic elements;

representing syllables associated with the acoustic speech model using the phonemic representation, each of the represented syllables including a series of acoustic elements;

constructing, via an electronic device, an acoustic grammar that contains transitions between the acoustic elements of the represented syllables, the transitions constrained according to phonotactic rules of the acoustic speech model; and

using the unstressed central vowel as a linking element between sequential phonemic elements contained in the constructed acoustic grammar.

2. The method of claim 1 , the series of acoustic elements defining an onset, a nucleus, and a coda for a given one of the represented syllables.

3. The method of claim 1 , the unstressed central vowel including a schwa acoustic element.

4. A system for reducing a search space for a recognition grammar used when interpreting natural language speech utterances, the system comprising:

at least one input device that receives an utterance from a user and generates an electronic signal corresponding to the utterance; and

a speech interpretation engine that receives the electronic signal corresponding to the utterance, the speech interpretation engine operable to:

create a phonemic representation of acoustic elements associated with an acoustic speech model, the acoustic elements including at least an unstressed central vowel and plurality of phonemic elements;

represent syllables associated with the acoustic speech model using the phonemic representation, each of the represented syllables including a series of acoustic elements;

construct an acoustic grammar that contains transitions between the acoustic elements of the represented syllables, the transitions constrained according to phonotactic rules of the acoustic speech model; and

use the unstressed central vowel as a linking element between sequential phonemic elements contained in the constructed acoustic grammar.

5. The system of claim 4 , the series of acoustic elements defining an onset, a nucleus, and a coda for a given one of the represented syllables.

6. The system of claim 4 , the unstressed central vowel including a schwa acoustic element.

Assignments (5)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 16, 2022
From: VOICE INVENTIONS, LLC
To: DIALECT, LLC
Reel/Frame 060820/0857 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 10, 2021
From: NUANCE COMMUNICATIONS, INC.
To: VOICE INVENTIONS, LLC
Reel/Frame 056185/0742 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 14, 2015
From: VOICEBOX TECHNOLOGIES CORPORATION
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 034703/0429 →
MERGER Recorded Apr 7, 2014
From: VOICEBOX TECHNOLOGIES, INC.
To: VOICEBOX TECHNOLOGIES CORPORATION
Reel/Frame 032620/0956 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 29, 2009
From: KENNEWICK, ROBERT A.; KE, MIN; TJALVE, MICHAEL; DI CRISTO, PHILIPPE
To: VOICEBOX TECHNOLOGIES, INC.
Reel/Frame 023444/0105 →