IP Library Granted Patent US 8,326,631
Granted Patent B1
US 8,326,631 · App. 12/415,688 · Granted Dec 4, 2012

Systems and methods for speech indexing

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,326,631
App. No.
12/415,688
Granted
Dec 4, 2012
Kind
B1
Abstract

A speech index for a recording or other representation of an audio signal containing speech is generated using a phonetic automatic voice recognition engine. A second speech index is also generated using a more accurate, but slower, automatic voice recognition engine such as a large vocabulary speech recognition (LVSR) engine. These two speech indexes are compared. The results of the comparison are then used to adjust certain parameters used by the phonetic engine while generating a speech index. The results may also be used to correct all or parts of the speech index generated by the phonetic automatic speech recognition engine.

Claims (46)

1. A method of indexing speech, comprising:

associating a first phonetic sequence with a first position in an audio signal using a phonetic recognizer;

associating said first phonetic sequence to a first linguistic element based on a first parameter;

associating a second linguistic element with a second position in said audio signal using a large vocabulary speech recognizer (LVSR);

comparing said first position and said second position to determine a phrase window;

comparing said first linguistic element to said second linguistic element if said phrase window meets a first criteria; and

adjusting said first parameter based upon a result of said step of comparing said first linguistic element

wherein said step of associating said second linguistic element is performed on a lesser portion of said audio signal than said step of associating said first phonetic sequence with said first position;

wherein said step of associating said first phonetic sequence to said first linguistic element also associates said first linguistic element with a confidence value and said lesser portion of said audio signal is selected to correspond to said first linguistic element based upon said confidence value.

2. The method of claim 1 , further comprising:

associating said first position with said second linguistic element.

3. The method of claim 1 , further comprising:

associating said first position with said second linguistic element.

4. The method of claim 1 wherein said step of adjusting said first parameter comprises increasing a probability that said second linguistic element will be associated with said first phonetic sequence by said step of associating said first phonetic sequence to a first linguistic element based on said first parameter.

5. The method of claim 1 , wherein said step of comparing further comprises:

correlating a second phonetic sequence associated with said second linguistic element with said first phonetic sequence.

6. A system for indexing speech, comprising:

a phonetic decoder that associates audio features of an audio signal with a first phonetic sequence at a first position in said audio signal;

a lexical interpreter that associates said first phonetic sequence with a first linguistic element based on a first parameter;

large vocabulary speech recognizer that associates a second linguistic element with a second position in said audio signal;

a speech index comparator that compares said first position and said second position to determine a phrase window;

and, said speech index comparator also compares said first linguistic element to said second linguistic element if said phrase window meets a first criteria; and

a parameter adjuster that adjusts said first parameter based upon a result of said speech index comparator

wherein said large vocabulary speech recognizer performs said association on a lesser portion of said audio signal than said phonetic decoder;

wherein said lexical interpreter also associates said first linguistic element with a confidence value and said lesser portion of said audio signal is selected to correspond to said first linguistic element based upon said confidence value.

7. The system of claim 6 , further comprising:

an index updater that associates said first position with said second linguistic element.

8. The system of claim 6 , further comprising:

an index updater that associates said first position with said second linguistic element.

9. The system of claim 6 , wherein adjusting said parameter adjuster increases a probability that said second linguistic element will be associated with said first phonetic sequence by said lexical interpreter.

10. The system of claim 6 , further comprising:

a phonetic sequence correlator that correlates a second phonetic sequence associated with said second linguistic element with said first phonetic sequence.

11. A program storage device readable by a machine, tangibly embodying a program of instructions executable by the machine to perform method steps for indexing speech, comprising:

associating a first phonetic sequence with a first position in an audio signal using a phonetic recognizer;

associating said first phonetic sequence to a first linguistic element based on a first parameter;

associating a second linguistic element with a second position in said audio signal using a large vocabulary speech recognizer a (LVSR);

comparing said first position and said second position to determine a phrase window;

comparing said first linguistic element to said second linguistic element if said phrase window meets a first criteria; and,

adjusting said first parameter based upon a result of said step of comparing said first linguistic element

wherein said step of associating said second linguistic element is performed on a lesser portion of said audio signal than said step of associating said first phonetic sequence with said first position;

wherein said step of associating said first phonetic sequence to said first linguistic element also associates said first linguistic element with a confidence value and said lesser portion of said audio signal is selected to correspond to said first linguistic element based upon said confidence value.

12. The program storage device of claim 11 , wherein the method further comprises:

associating said first position with said second linguistic element.

13. The program storage device of claim 11 , wherein the method further comprises:

associating said first position with said second linguistic element.

14. The program storage device of claim 11 wherein said step of adjusting said first parameter comprises increasing a probability that said second linguistic element will be associated with said first phonetic sequence by said step of associating said first phonetic sequence to a first linguistic element based on said first parameter.

Assignments (6)
SECURITY INTEREST Recorded Dec 23, 2025
From: VERINT SYSTEMS INC.
To: ALTER DOMUS (US) LLC, AS COLLATERAL AGENT
Reel/Frame 074034/0919 →
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (043293/0567) Recorded Nov 26, 2025
From: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
To: VERINT AMERICAS INC.
Reel/Frame 073796/0639 →
GRANT OF SECURITY INTEREST IN PATENT RIGHTS Recorded Jul 21, 2017
From: VERINT AMERICAS INC.
To: JPMORGAN CHASE BANK, N.A., AS COLLATERAL AGENT
Reel/Frame 043293/0567 →
RELEASE OF SECURITY INTEREST Recorded Jun 30, 2017
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
To: VERINT SYSTEMS INC.
Reel/Frame 043066/0318 →
GRANT OF SECURITY INTEREST IN PATENT RIGHTS Recorded Oct 21, 2013
From: VERINT SYSTEMS INC.
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
Reel/Frame 031465/0314 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 31, 2009
From: WATSON, JOESPH ALVA
To: VERINT SYSTEMS INC.
Reel/Frame 022478/0974 →