IP Library Granted Patent US 10,388,272
Granted Patent B1
US 10,388,272 · App. 16/209,640 · Granted Aug 20, 2019

Training speech recognition systems using word sequences

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,388,272
App. No.
16/209,640
Granted
Aug 20, 2019
Kind
B1
Abstract

A method may include obtaining first audio data of a communication session between a first device and a second device, obtaining a text string that is a transcription of the first audio data, and selecting a contiguous sequence of words from the text string as a first word sequence. The method may further include comparing the first word sequence to multiple word sequences obtained before the communication session and in response to the first word sequence corresponding to one of the multiple word sequences, incrementing a counter of multiple counters associated with the one of the multiple word sequences. The method may also include deleting the text string and the first word sequence and training and after deleting the text string and the first word sequence, training a language model of an automatic transcription system using the multiple word sequences and the multiple counters.

Claims (54)

1. A method comprising:

obtaining first audio data of a communication session between a first device of a first user and a second device of a second user, the communication session configured for verbal communication;

obtaining, during the communication session, a text string that is a transcription of the first audio data from an automatic transcription system;

selecting, during the communication session, a contiguous sequence of words from the text string as a first word sequence;

comparing, during the communication session, the first word sequence to a plurality of word sequences obtained before the communication session, each of the plurality of word sequences associated with a corresponding one of a plurality of counters;

in response to the first word sequence corresponding to one of the plurality of word sequences based on the comparison, incrementing, during the communication session, a counter of the plurality of counters associated with the one of the plurality of word sequences;

after incrementing the counter of the plurality of counters, deleting the text string and the first word sequence, wherein the first word sequence is deleted during the communication session; and

training, after deleting the text string and the first word sequence, a language model of the automatic transcription system using the plurality of word sequences and the plurality of counters.

2. The method of claim 1 , wherein each one of the plurality of counters indicates a number of occurrences that a corresponding one of the plurality of words sequences is included in a plurality of transcriptions of a plurality of communication sessions that occur between a plurality of devices, the plurality of devices not including the first device and the second device.

3. The method of claim 1 , wherein the text string is deleted during the communication session.

4. The method of claim 1 , wherein the text string is denormalized before selecting the contiguous sequence of words as the first word sequence.

5. The method of claim 1 , further comprising:

selecting a second contiguous sequence of words from the text string as a second word sequence;

comparing the second word sequence to the plurality of word sequences; and

in response to the second word sequence not corresponding to any of the plurality of word sequences based on the comparison, adding a third word sequence based on the second word sequence to the plurality of word sequences and adding a second counter with a count of one to the plurality of counters that is associated with the third word sequence of the plurality of word sequences,

wherein the training the language model of the automatic transcription system using the plurality of word sequences and the plurality of counters occurs after adding the second word sequence to the plurality of word sequences.

6. The method of claim 5 , wherein

the third word sequence is the same as the second word sequence,

the third word sequence includes fewer words than the second word sequence, or

the third word sequence includes a replacement word that is a generic word of one of the words in the second word sequence, the replacement word used in place of the one of the words in the second word sequence such that the third word sequence and the second word sequence include a same number of words.

7. The method of claim 6 , wherein the one of the words in the second word sequence are replaced based on the one of the words meeting a sensitive criteria, and

wherein removal words removed from the second word sequence to generate the third word sequence are removed based on the removal words meeting the sensitive criteria, wherein the third word sequence includes fewer words than the second word sequence.

8. The method of claim 6 , further comprising:

adding the one of the words in the second word sequence that is replaced by the replacement word to the plurality of word sequences; and

adding a third counter with a count of one to the plurality of counters that is associated with the one of the words in the second word sequence.

9. At least one non-transitory computer readable media configured to store one or more instructions that in response to be executed by at least one computing system cause performance of the method of claim 1 .

10. A method comprising:

obtaining first audio data of a communication session between a first device of a first user and a second device of a second user, the communication session configured for verbal communication;

obtaining, during the communication session, a text string that is a transcription of the first audio data;

selecting a contiguous sequence of words from the text string as a first word sequence;

comparing the first word sequence to a plurality of word sequences obtained before the communication session, each of the plurality of word sequences associated with a corresponding one of a plurality of counters;

in response to the first word sequence corresponding to one of the plurality of word sequences based on the comparison, incrementing a counter of the plurality of counters associated with the one of the plurality of word sequences;

after incrementing the counter of the plurality of counters, deleting the text string and the first word sequence; and

training, after deleting the text string and the first word sequence, a language model of an automatic transcription system using the plurality of word sequences and the plurality of counters.

11. The method of claim 10 , wherein the text string is obtained from the automatic transcription system.

12. The method of claim 10 , wherein each one of the plurality of counters indicates a number of occurrences that a corresponding one of the plurality of words sequences is included in a plurality of transcriptions of a plurality of communication sessions that occur between a plurality of devices, the plurality of devices not including the first device and the second device.

13. The method of claim 10 , wherein the steps of selecting the contiguous sequence of words, comparing the first word sequence, and incrementing the counter of the plurality of counters, each occur during the communication session.

14. The method of claim 10 , wherein the text string and the first word sequence are deleted during the communication session.

15. The method of claim 10 , wherein the text string is denormalized before selecting the contiguous sequence of words as the first word sequence.

16. The method of claim 10 , further comprising:

selecting a second contiguous sequence of words from the text string as a second word sequence;

comparing the second word sequence to the plurality of word sequences; and

in response to the second word sequence not corresponding to any of the plurality of word sequences based on the comparison, adding a third word sequence based on the second word sequence to the plurality of word sequences and adding a second counter with a count of one to the plurality of counters that is associated with the third word sequence of the plurality of word sequences,

wherein the training the language model of the automatic transcription system using the plurality of word sequences and the plurality of counters occurs after adding the second word sequence to the plurality of word sequences.

17. The method of claim 16 , wherein

the third word sequence is the same as the second word sequence,

the third word sequence includes fewer words than the second word sequence, or

the third word sequence includes a replacement word that is a generic word of one of the words in the second word sequence, the replacement word used in place of the one of the words in the second word sequence such that the third word sequence and the second word sequence include a same number of words.

18. The method of claim 17 , wherein the one of the words in the second word sequence are replaced based on the one of the words meeting a sensitive criteria, and

wherein removal words removed from the second word sequence to generate the third word sequence that includes fewer words than the second word sequence are removed based on the removal words meeting the sensitive criteria.

19. The method of claim 17 , further comprising:

adding the one of the words in the second word sequence that is replaced by the replacement word to the plurality of word sequences; and

adding a third counter with a count of one to the plurality of counters that is associated with the one of the words in the second word sequence.

20. At least one non-transitory computer readable media configured to store one or more instructions that in response to be executed by at least one computing system cause performance of the method of claim 10 .

Assignments (11)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
RELEASE OF SECURITY INTEREST Recorded Dec 16, 2021
From: CORTLAND CAPITAL MARKET SERVICES LLC
To: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 058533/0467 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
LIEN Recorded Feb 11, 2020
From: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CORTLAND CAPITAL MARKET SERVICES LLC
Reel/Frame 051894/0665 →
RELEASE OF SECURITY INTEREST Recorded May 8, 2019
From: U.S. BANK NATIONAL ASSOCIATION
To: SORENSON COMMUNICATIONS, LLC; SORENSON IP HOLDINGS, LLC; CAPTIONCALL, LLC; INTERACTIVECARE, LLC
Reel/Frame 049115/0468 →
RELEASE OF SECURITY INTEREST Recorded May 7, 2019
From: JPMORGAN CHASE BANK, N.A.
To: SORENSON COMMUNICATIONS, LLC; SORENSON IP HOLDINGS, LLC; CAPTIONCALL, LLC; INTERACTIVECARE, LLC
Reel/Frame 049109/0752 →
PATENT SECURITY AGREEMENT Recorded Apr 29, 2019
From: SORENSEN COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
Reel/Frame 050084/0793 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 4, 2019
From: HOLM, MICHAEL; BLACK, DAVID; BAROCIO, JESSE; THOMSON, DAVID; BOEKWEG, SCOTT; ROYLANCE, SHANE; CLEMENTS, KIERSTEN; BOEHME, KENNETH; ADAMS, JADIE; SKAGGS, JONATHAN; ORZECHOWSKI, GRZEGORZ; MCCLELLAN, JOSHUA
To: CAPTIONCALL, LLC
Reel/Frame 047905/0805 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 4, 2019
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 047905/0819 →
Cited By (25)
US 12,505,282 US 12,505,832 US 12,513,479 US 12,518,748 US 12,518,756 US 12,530,111 US 12,530,112 US 12,555,581 US 12,561,536 US 12,562,152 US 12,563,048 US 12,567,416 US 12,567,418 US 12,579,978 US 12,579,982 US 12,639,330 US 12,645,874 US 12,645,875 US 12,646,507 US 12,664,695 US 12,670,911 US 12,675,484 US 12,694,865 US 12,699,543 US 12,711,962