IP Library Granted Patent US 7,853,451
Granted Patent B1
US 7,853,451 · App. 10/739,380 · Granted Dec 14, 2010

System and method of exploiting human-human data for spoken language understanding systems

Assignee: AT&T Intellectual Property II, L.P.
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,853,451
App. No.
10/739,380
Granted
Dec 14, 2010
Kind
B1
Abstract

A method is disclosed for generating labeled utterances from human-human utterances for use in training a semantic classification model for a spoken dialog system. The method comprises augmenting received human-human utterances with data that relates to call-type gaps in the human-human utterances, augmenting the received human-human utterances by placing at least one word in the human-human utterances that improves the training ability of the utterances according to the conversation patterns of the spoken dialog system, clausifying the human-human utterances, labeling the clausified and augmented human-human utterances and building the semantic classification model for the spoken dialog system using the labeled utterances.

Claims (19)

1. A method for generating labeled utterances from human-human utterances for use in training a semantic classification model for a spoken dialog system, the method comprising:

identifying via a processor call-type gaps in human-human utterances to yield identified call-type gaps, wherein the identified call-type gaps include one of a missing call-type and an infrequent call-type;

augmenting the human-human utterances with data that relates to the identified call-type gaps in the human-human utterances;

augmenting the human-human utterances, to yield augmented human-human utterances by placing at least one word within the text of the human-human utterances that improves a training ability of the human-human utterances according to conversation patterns of the spoken dialog system;

clausifying the augmented human-human utterances to yield clausified, augmented human-human utterances;

labeling the clausified, augmented human-human utterances to yield labeled utterances; and

building a semantic classification model for the spoken dialog system using the labeled utterances.

2. The method of claim 1 , wherein clausifying the received human-human utterances occurs before the augmenting steps.

3. The method of claim 1 , wherein clausifying the received human-human utterances further comprises:

detecting sentence boundaries within a speech utterance text;

editing the speech utterance text to remove unneeded words; and

detecting conjunctions within the speech utterance text, wherein the clausifer outputs annotated text having identifiable clauses according to the sentence boundaries, edited text, and conjunctions within the speech utterance text.

4. The method of claim 1 , wherein the data that relates to call-type gaps used to augment the received human-human utterances is borrowed from other spoken dialog system applications.

5. The method of claim 4 , wherein the other spoken dialog system applications have a related function to the spoken dialog system.

6. The method of claim 1 , wherein only a portion of the clausified human-human utterances are labeled and used to build the semantic classification model.

7. The method of claim 6 , wherein the portion of the clausified human-human utterances that are labeled and used to build the semantic classification model relate to identifying the intent of the speaker.

8. The method of claim 1 , wherein the at least one word is either no or yes.

9. The method of claim 1 , wherein the at least one word relates to a phrase related to a computer-human interaction.

10. The method of claim 1 , wherein missing or infrequent call-types include call types which are missing or infrequent in the human-human utterances, but would be more common in human-machine dialogs.

Assignments (6)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 14, 2023
From: NUANCE COMMUNICATIONS, INC.
To: MICROSOFT TECHNOLOGY LICENSING, LLC
Reel/Frame 065566/0013 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041512/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 23, 2017
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 041045/0468 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 23, 2017
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 041045/0512 →
CORRECTIVE ASSIGNMENT TO CORRECT THE ASSIGNEE NAME PREVIOUSLY RECORDED AT REEL: 014843 FRAME: 0058. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Jan 23, 2017
From: GUPTA, NARENDRA K.; TUR, GOKHAN
To: AT&T CORP.
Reel/Frame 041069/0990 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 18, 2003
From: GUPTA, NARENDRA K.; TUR, GOKHAN
To: AT&T CORPORATION
Reel/Frame 014843/0058 →