IP Library Granted Patent US 8,386,262
Granted Patent B2
US 8,386,262 · App. 13/481,031 · Granted Feb 26, 2013

System and method of spoken language understanding in human computer dialogs

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,386,262
App. No.
13/481,031
Granted
Feb 26, 2013
Kind
B2
Abstract

A system and method are disclosed that improve automatic speech recognition in a spoken dialog system. The method comprises partitioning speech recognizer output into self-contained clauses, identifying a dialog act in each of the self-contained clauses, qualifying dialog acts by identifying a current domain object and/or a current domain action, and determining whether further qualification is possible for the current domain object and/or current domain action. If further qualification is possible, then the method comprises identifying another domain action and/or another domain object associated with the current domain object and/or current domain action, reassigning the another domain action and/or another domain object as the current domain action and/or current domain object and then recursively qualifying the new current domain action and/or current object. This process continues until nothing is left to qualify.

Claims (39)

1. A method comprising:

partitioning, via a processor, a speech recognizer output into independent clauses;

identifying, independent of domain, a dialog act for each of the independent clauses;

identifying, dependent on domain, an object within each of the independent clauses;

generating a semantic representation using the dialog act and the object; and

recursively extending the semantic representation by qualifying the object in each of the independent clauses.

2. The method of claim 1 , wherein the semantic representation is used by a dialog manager in a spoken dialog system to determine a response to a user input.

3. The method of claim 1 , further comprising:

identifying, dependent on domain, an action within each of the independent clauses, wherein generating the semantic representation further comprises using the action.

4. The method of claim 1 , wherein qualifying the object in each of the independent clauses comprises extracting additional objects from each of the independent clauses.

5. The method of claim 1 , wherein identifying the object comprises using a domain specific classifier.

6. The method of claim 1 , wherein generating the semantic representation further comprises identifying relationships between the dialog act and the object.

7. The method of claim 1 , wherein generating the semantic representation further comprises filling in a predefined data structure associated with the dialog act.

8. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed on the processor, perform a method comprising:

partitioning a speech recognizer output into independent clauses;

identifying, independent of domain, a dialog act for each of the independent clauses;

identifying, dependent on domain, an object within each of the independent clauses;

generating a semantic representation using the dialog act and the object; and

recursively extending the semantic representation by qualifying the object in each of the independent clauses.

9. The system of claim 8 , wherein the semantic representation is used by a dialog manager in a spoken dialog system to determine a response to a user input.

10. The system of claim 8 , the computer-readable storage medium having additional instructions stored which, when executed on the processor, perform a method comprising:

identifying, dependent on domain, an action within each of the independent clauses, wherein generating the semantic representation further comprises using the action.

11. The system of claim 8 , wherein qualifying the object in each of the independent clauses comprises extracting additional objects from each of the independent clauses.

12. The system of claim 8 , wherein identifying the object comprises using a domain specific classifier.

13. The system of claim 8 , wherein generating the semantic representation further comprises identifying relationships between the dialog act and the object.

14. The system of claim 8 , wherein generating the semantic representation further comprises filling in a predefined data structure associated with the dialog act.

15. A computer-readable storage device having instructions stored which, when executed on a computing device, perform a method comprising:

partitioning a speech recognizer output into independent clauses;

identifying, independent of domain, a dialog act for each of the independent clauses;

identifying, dependent on domain, an object within each of the independent clauses;

generating a semantic representation using the dialog act and the object; and

recursively extending the semantic representation by qualifying the object in each of the independent clauses.

16. The computer-readable storage device of claim 15 , wherein the semantic representation is used by a dialog manager in a spoken dialog system to determine a response to a user input.

17. The computer-readable storage device of claim 15 , the computer-readable storage device having additional instructions stored which, when executed on the computing device, perform a method comprising identifying, dependent on domain, an action within each of the independent clauses, wherein generating the semantic representation further comprises using the action.

18. The computer-readable storage device of claim 15 , wherein qualifying the object in each of the independent clauses comprises extracting additional objects from each of the independent clauses.

19. The computer-readable storage device of claim 15 , wherein identifying the object comprises using a domain specific classifier.

20. The computer-readable storage device of claim 15 , wherein generating the semantic representation further comprises identifying relationships between the dialog act and the object.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041512/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2016
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 038275/0041 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 28, 2016
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 038275/0130 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 29, 2012
From: BANGALORE, SRINIVAS; GUPTA, NARENDRA K.; RAHIM, MAZIN G.
To: AT&T CORP.
Reel/Frame 028278/0004 →