IP Library Granted Patent US 7,143,035
Granted Patent B2
US 7,143,035 · App. 10/107,723 · Granted Nov 28, 2006

Methods and apparatus for generating dialog state conditioned language models

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,143,035
App. No.
10/107,723
Granted
Nov 28, 2006
Kind
B2
Abstract

Techniques are provided for generating improved language modeling. Such improved modeling is achieved by conditioning a language model on a state of a dialog for which the language model is employed. For example, the techniques of the invention may improve modeling of language for use in a speech recognizer of an automatic natural language based dialog system. Improved usability of the dialog system arises from better recognition of a user's utterances by a speech recognizer, associated with the dialog system, using the dialog state-conditioned language models. By way of example, the state of the dialog may be quantified as: (i) the internal state of the natural language understanding part of the dialog system; or (ii) words in the prompt that the dialog system played to the user.

Claims (22)

1. A method for use in accordance with a dialog system, the method comprising the steps of:

generating at least one language model, the at least one language model being conditioned on a state of dialog associated with the dialog system; and

storing the at least one language model for subsequent use in accordance with a speech recognizer associated with the dialog system;

wherein the step of generating the at least one language model conditioned on a state of dialog associated with the dialog system further comprises the steps of:

caching words in a prompt presented to a user by the dialog system;

building a unigram model on the cached words; and

interpolating the unigram model with a baseline model.

2. The method of claim 1 , wherein the baseline model is a trigram language model trained from available training data for a domain of the dialog system.

3. Apparatus for use in accordance with a dialog system, the apparatus comprising:

at least one processor operative to generate at least one language model, the at least one language model being conditioned on a state of dialog associated with the dialog system; and

memory, coupled to the at least one processor, for storing the at least one language model for subsequent use in accordance with a speech recognizer associated with the dialog system;

wherein the operation of generating the at least one language model conditioned on a state of dialog associated with the dialog system further comprises: (i) caching words in a prompt presented to a user by the dialog system; (ii) building a unigram model on the cached words; and (iii) interpolating the unigram model with a baseline model.

4. An article of manufacture for use in accordance with a dialog system, comprising a machine readable medium containing one or more programs which when executed implement the steps of:

generating at least one language model, the at least one language model being conditioned on a state of dialog associated with the dialog system; and

storing the at least one language model for subsequent use in accordance with a speech recognizer associated with the dialog system;

wherein the step of generating the at least one language model conditioned on a state of dialog associated with the dialog system further comprises: (i) caching words in a prompt presented to a user by the dialog system; (ii) building a unigram model on the cached words; and (iii) interpolating the unigram model with a baseline model.

5. A data structure for use in accordance with a dialog system, the data structure comprising:

at least one language model, the at least one language model being conditioned on a state of dialog associated with the dialog system, and the at least one language model being for subsequent use in accordance with a speech recognizer associated with the dialog system;

wherein the at least one language model is generatable by: (i) caching words in a prompt presented to a user by the dialog system; (ii) building a unigram model on the cached words; and (iii) interpolating the unigram model with a baseline model.

6. A dialog system, comprising:

at least one language model, the at least one language model being conditioned on a state of dialog associated with the dialog system, wherein the at least one language model is generatable by: (i) caching words in a prompt presented to a user by the dialog system; (ii) building a unigram model on the cached words; and (iii) interpolating the unigram model with a baseline model; and

a speech recognizer for recognizing speech utilizing the at least one language model.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 13, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065552/0934 →