IP Library Granted Patent US 10,102,847
Granted Patent B2
US 10,102,847 · App. 15/235,961 · Granted Oct 16, 2018

Automated learning for speech-based applications

Inventor: Charles C Wooters (Sunnyvale, CA)
Assignee: VERINT AMERICAS INC.
G10L15/065G10L15/08G10L15/22G10L15/26G10L15/18G10L2015/0638G10L2015/223
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,102,847
App. No.
15/235,961
Granted
Oct 16, 2018
Kind
B2
Abstract

Systems and methods for modifying a computer-based speech recognition system. A speech utterance is processed with the computer-based speech recognition system using a set of internal representations, which may comprise parameters for recognizing speech in a speech utterance, such as parameters of an acoustic model and/or a language model. The computer-based speech recognition system may perform a first task in response to the processed speech utterance. The utterance may also be provided to a human who performs a second task based on the utterance. Data indicative of the first task, performed by the computer system, is compared to data indicative of a second task, performed by the human in response to the speech utterance. Based on the comparison, the set of internal representations may be updated or modified to improve the speech recognition performance and capabilities of the speech recognition system.

Claims (54)

1. A method comprising:

receiving, by a computer-based speech recognition system that is communicatively coupled to a communication network, a speech input including one or more words or phrases, the computer-based speech recognition system comprising at least one processor and at least one memory device, wherein the speech input is from a call to a call center via the communication network;

determining, by the computer-based speech recognition system, to provide the speech input to a human;

receiving a response from the human that identifies a first task for the speech input, the first task including a first transcription of the speech input;

performing the first task that is identified by the human for the speech input;

processing the speech input to identify a second task for the speech input, the processing using a set of internal representations of the computer-based speech recognition system, the set of internal representations comprising one or more machine-readable parameters for recognizing speech in a speech utterance, the second task including a second transcription of the speech input;

comparing the first transcription of the speech input included in the first task identified by the human with the second transcription of the speech input included in the second task identified by the computer-based recognition system to determine one or more differences between the first transcription and the second transcription;

modifying the set of internal representations of the computer-based speech recognition system based at least in part on the one or more differences between the first transcription and the second transcription to create a modified set of internal representations, wherein the modifying includes adjusting at least a portion of the set of internal representations;

checking the performance of the modified set of internal representations to prevent the modification from degrading the set internal representations, wherein the checking comprises determining that a performance difference between the set of internal representations before and after modification is within a margin of error;

receiving, by the computer-based speech recognition system, another input; and

processing, by the at least one processor and based at least in part on the modified set of internal representations, the other input to identify a third task for the other input.

2. The method of claim 1 , wherein:

the input comprises a speech utterance.

3. The method of claim 1 , further comprising, in response to determining to provide the speech input to the human, providing the speech input to the human in real-time as the speech input is delivered to the computer-based speech recognition system.

4. The method of claim 1 , wherein comparing the first transcription identified by the human with the second transcription identified by the computer-based recognition system to determine one or more differences between the first transcription and the second transcription comprises determining that the second transcription is not within an acceptable margin of error to the first transcription.

5. The method of claim 1 , further comprising:

receiving, by the computer-based speech recognition system, another speech input; and

processing, by the at least one processor and based at least partly on the modified set of internal representations, the other speech input to identify a particular transcription for the other speech input.

6. A computer-based speech recognition system, comprising:

one or more processors communicatively coupled to a communication network; and

memory storing instructions that, when executed by the one or more processors, cause the computer-based speech recognition system to perform acts comprising:

obtaining a speech input including one or more words or phrases, wherein the speech input is from a call to a call center via the communication network;

determining to provide the speech input to a human;

receiving a response from the human that identifies a first task for the input, the first task including a first transcription of the speech input;

processing the speech input to identify a second task for the speech input, the processing using a set of internal representations of the computer-based speech recognition system, the set of internal representations comprising one or more machine-readable parameters for recognizing speech in a speech utterance, the second task including a second transcription of the speech input;

comparing the speech input and the first transcription of the speech input included in the first task with the second transcription of the speech input included in the second task, the set of internal representations comprising one or more machine-readable parameters for recognizing speech in a speech utterance;

modifying the set of internal representations of the computer-based speech recognition system based at least in part on the comparing the first transcription with the second transcription to create a modified set of internal representations, wherein the modifying includes adjusting at least a portion of the set of internal representations;

checking the performance of the modified set of internal representations to prevent the modification from degrading the set internal representations, wherein the checking comprises determining that a performance difference between the set of internal representations before and after modification is within a margin of error;

receiving, by the computer-based speech recognition system, another input; and

processing, by the one or more processors and based at least in part on the modified set of internal representations, the other input to identify a third task for the other input.

7. The computer-based speech recognition system of claim 6 , the acts further comprising processing the speech input to identify a second task for the input, the processing using the set of internal representations of the computer-based speech recognition system.

8. The computer-based speech recognition system of claim 7 , wherein comparing the speech input and the first task with the set of internal representations of the computer-based speech recognition system comprises comparing the first transcription of the speech input included in the first task and the second transcription of the speech input included in the second task to determine one or more differences between the transcription and the second transcription.

9. The computer-based speech recognition system of claim 8 , wherein the modifying the set of internal representations is based at least in part one the one or more differences between the first transcription and the second transcription.

10. The computer-based speech recognition system of claim 6 , wherein:

the speech input comprises a speech utterance.

11. The computer-based speech recognition system of claim 6 , the acts further comprising, in response to determining to provide the speech input to the human, providing the speech input to the human in real-time as the speech input is delivered to the computer-based speech recognition system.

12. The computer-based speech recognition system of claim 6 , the acts further comprising performing the first task that is identified by the human for the speech input, wherein at least a portion of the first task that is identified by the human is performed by the human.

13. One or more non-transitory computer-readable storage media storing instructions that, when executed by one or more processors, communicatively coupled to a communication network, configure the one or more processors to perform acts comprising:

receiving, by a computer-based speech recognition system, a speech input including one or more words or phrases, wherein the speech input is from a call to a call center via the communication network;

determining, by the computer-based speech recognition system, to provide the speech input to a human;

receiving a response from the human that identifies a first task for the speech input, the first task including a first transcription of the speech input;

processing the speech input to identify a second task for the speech input, the processing using a set of internal representations of the computer-based speech recognition system, the set of internal representations comprising one or more machine-readable parameters for recognizing speech in a speech utterance, the second task including a second transcription of the speech input;

comparing the speech input and the first transcription of the speech input included in the first task with the second transcription of the speech input included in the second task, the set of internal representations comprising one or more machine-readable parameters for recognizing speech in a speech utterance;

modifying the set of internal representations of the computer-based speech recognition system based at least in part on the comparing the first transcription with the second transcription to create a modified set of internal representations, wherein the modifying includes adjusting at least a portion of the set of internal representations;

checking the performance of the modified set of internal representations to prevent the modification from degrading the set internal representations, wherein the checking comprises determining that a performance difference between the set of internal representations before and after modification is within a margin of error;

receiving, by the computer-based speech recognition system, another input; and

processing, by the one or more processors and based at least in part on the modified set of internal representations, the other input to identify a third task for the other input.

14. The one or more non-transitory computer-readable storage media of claim 13 , the acts further comprising processing the speech input to identify a second task for the speech input, the processing using the set of internal representations of the computer-based speech recognition system.

15. The one or more non-transitory computer-readable storage media of claim 14 , wherein comparing the speech input and the first transcription of the speech input included in the first task with the set of internal representations of the computer-based speech recognition system comprises comparing the first transcription and the second transcription of the speech input included in the second task to determine one or more differences between the first task and the second task.

16. The computer-based speech recognition system of claim 15 , wherein the modifying the set of internal representations is based at least in part one the one or more differences between the first transcription and the second transcription.

17. The computer-based speech recognition system of claim 13 , wherein:

the speech input comprises a speech utterance.

18. The computer-based speech recognition system of claim 13 , the acts further comprising, in response to determining to provide the input to the human, providing the speech input to the human in real-time as the speech input is delivered to the computer-based speech recognition system.

19. The computer-based speech recognition system of claim 13 , the acts further comprising performing the first task that is identified by the human for the speech input, wherein at least a portion of the first task that is identified by the human is performed by the human.

Assignments (5)
SECURITY INTEREST Recorded Dec 23, 2025
From: VERINT AMERICAS INC.
To: ALTER DOMUS (US) LLC, AS COLLATERAL AGENT
Reel/Frame 074034/0292 →
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (050612/0972) Recorded Nov 26, 2025
From: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
To: VERINT AMERICAS INC.
Reel/Frame 073796/0675 →
PATENT SECURITY AGREEMENT Recorded Oct 3, 2019
From: VERINT AMERICAS INC.
To: JPMORGAN CHASE BANK, N.A.
Reel/Frame 050612/0972 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 19, 2018
From: NEXT IT CORPORATION
To: VERINT AMERICAS INC.
Reel/Frame 044963/0046 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 12, 2016
From: WOOTERS, CHARLES C
To: NEXT IT CORPORATION
Reel/Frame 039423/0330 →
Continuity (4)
Continuation 14610891 · Jan 30, 2015
Continuation 12584770 · Sep 11, 2009
Provisional Application 61096095 · Sep 11, 2008
Related Publication 20160351186A1 · Dec 1, 2016