IP Library Granted Patent US 9,412,362
Granted Patent B2
US 9,412,362 · App. 14/319,847 · Granted Aug 9, 2016

System and method for determining the compliance of agent scripts

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,412,362
App. No.
14/319,847
Granted
Aug 9, 2016
Kind
B2
Abstract

Systems and methods of script identification in audio data obtained from audio data. The audio data is segmented into a plurality of utterances. A script model representative of a script text is obtained. The plurality of utterances are decoded with the script model. A determination is made if the script text occurred in the audio data.

Claims (41)

1. A method of script identification in audio data, the method comprising: obtaining audio data segmenting the audio data into a plurality of utterances;

obtaining a script model representative of a script text:

decoding the plurality of utterances by applying the script model to the plurality of utterances;

determining if the script text occurred in the audio data from the decoded plurality of utterances;

applying speech analytics to the plurality of utterances to identify at least one event;

upon identification of the at least one event selecting, the script model based upon the identified at least one event; and

decoding a subset of the plurality of utterances occurring, after the identified at least one event with the selected script model;

filtering the plurality of utterances to include only utterances attributed to a customer service agent;

extracting acoustic features from the filtered plurality of utterances;

using the extracted acoustic features in decoding the plurality of utterances;

determine if the audio data is non-compliant; and

if so, initiate at least one remedial action that is selected from a graphical display to present on screen guidance to a customer service agent.

2. The method of claim 1 , wherein decoding the plurality of utterances comprises applying a plurality of script models to the plurality of utterances and producing an indication of script models identified in the plurality of data.

3. The method of claim 1 , wherein if the script text is determined to have occurred in the audio data, further comprising: transcribing the utterance containing the script to produce an utterance transcription; comparing the script text to the utterance transcription; determining a script accuracy.

4. The method of claim 3 , further comprising evaluating a compliance of the audio data with a script requirement threshold by comparing the determined script accuracy to the script requirement threshold.

5. The method of claim 4 , wherein the script requirement threshold is a minimum word error rate.

6. The method of claim 4 , further comprising:

presenting additional information to a customer; and

producing an alert of a non-compliant script.

7. The method of claim 3 , wherein if the script text is determined to have not occurred, initiating, at least one remedial action.

8. The method of claim 1 , further comprising: compiling the script model from at least the script text; and modifying the script model to include acceptable variations in the script text.

9. The method of claim 8 , wherein the script model is further compiled from at least one speaker acoustic model.

10. The method of claim 1 , wherein the audio data is streaming audio data.

11. A non-transient computer readable medium programmed with computer readable code that upon execution by a computer processor causes the computer processor to:

obtain audio data; segment the audio data into a plurality of utterances;

obtain a script model representative of a script text;

decode the plurality of utterances by applying the script model to the plurality of utterances;

determine if the script text occurred in the audio data from the decoded plurality of utterances;

filter the plurality of utterances to include only utterances attributed to a customer service agent;

extract acoustic features from the filtered plurality of utterances;

use the extracted acoustic features in decoding the plurality of utterances;

determine if the audio data is non-compliant; and

if so, initiate at least one remedial action that is selected from a graphical display to present on screen guidance to a customer service agent.

12. The non-transient computer readable medium of claim 11 , wherein execution of the computer readable code further causes the processor to: apply speech analytics to the plurality of utterances to identify at least one event; upon identification of the at least one event select the script model based upon the identified at least one event; and decode a subset of the plurality of utterances occurring after the identified least one event with the selected script model.

13. The non-transient computer readable medium of claim 12 , wherein execution of the computer readable code further causes the processor to: transcribe the utterance containing the script to produce an utterance transcription; compare the script text to the utterance transcription; and determine a script accuracy.

14. The non-transient computer readable medium of claim 13 , wherein execution of the computer readable code further causes the processor to: evaluate a compliance of the audio data with a script requirement threshold by comparing the determined script accuracy to the script requirement threshold; and initiate a remedial action based upon the evaluation.

15. A system for identification of a script in audio data, the system comprising:

an audio data source; a script model database comprising a plurality of script models each script model of the plurality representative of at least one script text;

a processing system communicatively connected to the script model database and the audio data source, the processing system obtains audio data from the audio data source, segments the audio data into a plurality of utterances, obtains at least one script model from the script model database, decodes the plurality of utterances by applying the script model to the plurality of utterances, determines if the script text occurred in the audio data from the decoded plurality of utterances, and based upon the determination initiates a remedial action if the script text did not occur in the audio data; and

a graphical display communicatively connected to the processing system and operable by the processing system, wherein the processing system operates the graphical display to present visual guidance to a customer service agent as the remedial action,

wherein the processor further filters the plurality of utterances to include only utterances attributed to the customer service agent, extracts acoustic features from the filtered plurality of utterances, applies speech analytics to the filtered plurality of utterances to identify at least one event, upon identification of the at least one event selects the script model based upon the identified at least one event and decodes a subset of the filtered plurality of utterances occurring after the identified at least one event with the selected script model.

Assignments (3)
SECURITY INTEREST Recorded Dec 23, 2025
From: VERINT SYSTEMS INC.
To: ALTER DOMUS (US) LLC, AS COLLATERAL AGENT
Reel/Frame 074034/0919 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 22, 2021
From: VERINT SYSTEMS LTD.
To: VERINT SYSTEMS INC.
Reel/Frame 057568/0183 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 30, 2014
From: IANNONE, JEFFERY MICHAEL; WEIN, RON; ZIV, OMER
To: VERINT SYSTEMS LTD.
Reel/Frame 033423/0579 →