IP Library Granted Patent US 10,199,039
Granted Patent B2
US 10,199,039 · App. 14/963,408 · Granted Feb 5, 2019

Library of existing spoken dialog data for use in generating new natural language spoken dialog systems

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,199,039
App. No.
14/963,408
Granted
Feb 5, 2019
Kind
B2
Abstract

A machine-readable medium may include a group of reusable components for building a spoken dialog system. The reusable components may include a group of previously collected audible utterances. A machine-implemented method to build a library of reusable components for use in building a natural language spoken dialog system may include storing a dataset in a database. The dataset may include a group of reusable components for building a spoken dialog system. The reusable components may further include a group of previously collected audible utterances. A second method may include storing at least one set of data. Each one of the at least one set of data may include ones of the reusable components associated with audible data collected during a different collection phase.

Claims (42)

1. A method comprising:

collecting, via a plurality of industry specific spoken dialog systems and during a plurality of collection phases comprising respective defined periods of time and in which respective conversations between a user and a respective industry specific spoken dialog system of the plurality of industry specific spoken dialog systems occurs, a plurality of audible utterances for each respective conversation;

organizing, via a processor, the plurality of audible utterances into a plurality of datasets having call-type labels, wherein each dataset in the plurality of datasets pertains to a unique industrial sector in a plurality of industrial sectors;

identifying a positive example utterance for each of the call-type labels;

generating a natural language spoken dialog system using the plurality of datasets and the positive example utterance for each of the call-type labels;

receiving audible speech at the natural language spoken dialog system; and

converting the audible speech into text via the natural language spoken dialog system.

2. The method of claim 1 , further comprising, prior to the generating of the natural language spoken dialog system, comparing, for each of the call-type labels, each utterance in the plurality of audible utterances to a negative example utterance for an associated call-type.

3. The method of claim 1 , wherein organizing of the plurality of audible utterances further comprises using a corresponding industry specific spoken dialog system of the plurality of industry specific spoken dialog systems.

4. The method of claim 3 , wherein the corresponding industry specific spoken dialog system and the unique industrial sector share a common task domain.

5. The method of claim 1 , wherein the plurality of datasets are stored in an extensible markup language database.

6. The method of claim 1 , wherein the plurality of datasets are stored in a relational database.

7. The method of claim 1 , wherein utterances in the plurality of audible utterances are each associated with a respective utterance-type category.

8. A system comprising:

a processor; and

a computer-readable storage medium having instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

collecting, via a plurality of industry specific spoken dialog systems and during a plurality of collection phases comprising respective defined periods of time and in which respective conversations between a user and a respective industry specific spoken dialog system of the plurality of industry specific spoken dialog systems occurs, a plurality of audible utterances for each respective conversation;

organizing the plurality of audible utterances into a plurality of datasets having call-type labels, wherein each dataset in the plurality of datasets pertains to a unique industrial sector in a plurality of industrial sectors;

identifying a positive example utterance for each of the call-type labels;

generating a natural language spoken dialog system using the plurality of datasets and the positive example utterance for each of the call-type labels;

receiving audible speech at the natural language spoken dialog system; and

converting the audible speech into text via the natural language spoken dialog system.

9. The system of claim 8 , the computer-readable storage medium having additional instructions stored which, when executed by the processor, cause the processor to perform operations comprising:

prior to the generating of the natural language spoken dialog system, comparing, for each of the call-type labels, each utterance in the plurality of audible utterances to a negative example utterance for an associated call-type.

10. The system of claim 8 , wherein organizing of the plurality of audible utterances further comprises using a corresponding industry specific spoken dialog system of the plurality of industry specific spoken dialog systems.

11. The system of claim 10 , wherein the corresponding industry specific spoken dialog system and the unique industrial sector share a common task domain.

12. The system of claim 8 , wherein the plurality of datasets are stored in an extensible markup language database.

13. The system of claim 8 , wherein the plurality of datasets are stored in a relational database.

14. The system of claim 8 , wherein utterances the plurality of audible utterances are each associated with a respective utterance-type category.

15. A computer-readable storage device having instructions stored which, when executed by a computing device, cause the computing device to perform operations comprising:

collecting, via a plurality of industry specific spoken dialog systems and during a plurality of collection phases comprising respective defined periods of time and in which respective conversations between a user and a respective industry specific spoken dialog system of the plurality of industry specific spoken dialog systems occurs, a plurality of audible utterances for each respective conversation;

organizing the plurality of audible utterances into a plurality of datasets having call-type labels, wherein each dataset in the plurality of datasets pertains to a unique industrial sector in a plurality of industrial sectors;

identifying a positive example utterance for each of the call-type labels;

generating a natural language spoken dialog system using the plurality of datasets and the positive example utterance for each of the call-type labels;

receiving audible speech at the natural language spoken dialog system; and

converting the audible speech into text via the natural language spoken dialog system.

16. The computer-readable storage device of claim 15 , having additional instructions stored which, when executed by the computing device, cause the computing device to perform operations comprising:

prior to the generating of the natural language spoken dialog system, comparing, for each of the call-type labels, each utterance in the plurality of audible utterances to a negative example utterance for an associated call-type.

17. The computer-readable storage device of claim 15 , wherein organizing of the plurality of audible utterances further comprises using a corresponding industry specific spoken dialog system of the plurality of industry specific spoken dialog systems.

18. The computer-readable storage device of claim 17 , wherein the corresponding industry specific spoken dialog system and the unique industrial sector share a common task domain.

19. The computer-readable storage device of claim 15 , wherein the plurality of datasets are stored in an extensible markup language database.

20. The computer-readable storage device of claim 15 , wherein the plurality of datasets are stored in a relational database.

Assignments (4)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 26, 2017
From: AT&T INTELLECTUAL PROPERTY II, L.P.
To: NUANCE COMMUNICATIONS, INC.
Reel/Frame 041512/0608 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 26, 2016
From: AT&T CORP.
To: AT&T PROPERTIES, LLC
Reel/Frame 038529/0164 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 26, 2016
From: AT&T PROPERTIES, LLC
To: AT&T INTELLECTUAL PROPERTY II, L.P.
Reel/Frame 038529/0240 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 15, 2016
From: BEGEJA, LEE; DI FABBRIZIO, GIUSEPPE; GIBBON, DAVID CRAWFORD; HAKKANI-TUR, DILEK Z.; LIU, ZHU; RENGER, BERNARD S.; SHAHRARAY, BEHZAD; TUR, GOKHAN
To: AT&T CORP.
Reel/Frame 038294/0391 →