IP Library Granted Patent US 11,600,279
Granted Patent B2
US 11,600,279 · App. 17/279,512 · Granted Mar 7, 2023

Transcription of communications

Inventors: Brian Chevrier (Highland, UT); Shane Roylance (Farmington, UT); Kenneth Boehme (South Jordan, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G10L15/06G10L15/08G10L15/22G10L2015/088G10L2015/221
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,600,279
App. No.
17/279,512
Granted
Mar 7, 2023
Kind
B2
Abstract

A method to transcribe communications may include obtaining audio data originating at a first device during a communication session between the first device and a second device and providing the audio data to an automated speech recognition system configured to transcribe the audio data. The method may further include obtaining multiple hypothesis transcriptions generated by the automated speech recognition system. Each of the multiple hypothesis transcriptions may include one or more words determined by the automated speech recognition system to be a transcription of a portion of the audio data. The method may further include determining one or more consistent words that are included in two or more of the multiple hypothesis transcriptions and in response to determining the one or more consistent words, providing the one or more consistent words to the second device for presentation of the one or more consistent words by the second device.

Claims (35)

1. A method to transcribe communications, the method comprising:

obtaining a plurality of hypothesis transcriptions of audio data;

determining a plurality of consistent words that are included in two or more of the plurality of hypothesis transcriptions;

in response to determining the plurality of consistent words, directing the plurality of consistent words to a device for presentation of the plurality of consistent words; and

presenting the plurality of consistent words in a rolling fashion, a pace of the presentation of the plurality of consistent words in the rolling fashion being variable such that when more words are to be presented the words are presented quicker than when fewer words are to be presented.

2. The method of claim 1 , further comprising:

determining an update word in a final transcription of the audio data that is different from any of the plurality of consistent words; and

in response to determining the update word, directing an indication of the update word to the device, the update word replacing one or more of the plurality of consistent words in the presentation of the plurality of consistent words.

3. The method of claim 1 , wherein the presentation of the plurality of consistent words by the device is configured to occur before a final transcription of the audio data is provided to the device.

4. The method of claim 1 , wherein the plurality of hypothesis transcriptions are obtained from a speech recognition system that includes a plurality of speech engines configured to recognize speech.

5. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 1 .

6. A method to transcribe communications, the method comprising:

obtaining a plurality of hypothesis transcriptions of audio data;

determining one or more consistent words that are included in two or more of the plurality of hypothesis transcriptions;

in response to determining the one or more consistent words, directing the one or more consistent words to a device for presentation of the one or more consistent words;

determining an update word in a final transcription of the audio data that is different from any of the one or more consistent words; and

in response to determining the update word and after directing the one or more consistent words to the device, directing an indication of the update word to the device, the update word replacing one or more of the one or more consistent words in the presentation of the one or more consistent words on the device.

7. The method of claim 6 , further comprising obtaining the audio data during a communication session between a second device and the device, the audio data originating at the second device.

8. The method of claim 6 , wherein the presentation of the one or more consistent words by the device is configured to occur before the final transcription of the audio data is provided to the device.

9. The method of claim 6 , wherein the one or more consistent words are presented in a rolling fashion, a pace of the presentation of the one or more consistent words in the rolling fashion being variable.

10. The method of claim 6 , wherein the plurality of hypothesis transcriptions are obtained sequentially over time and determining the one or more consistent words includes comparing a first hypothesis transcription of the plurality of hypothesis transcriptions with a second hypothesis transcription of the plurality of hypothesis transcriptions, the second hypothesis transcription directly following the first hypothesis transcription among the plurality of hypothesis transcriptions.

11. The method of claim 6 , wherein the plurality of hypothesis transcriptions are obtained from a speech recognition system that includes a plurality of speech engines configured to recognize speech.

12. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 6 .

13. A method to transcribe communications, the method comprising: obtaining a plurality of hypothesis transcriptions of audio data, each of the plurality of hypothesis transcriptions including one or more words determined to be a transcription of portions of the audio data;

determining one or more consistent words that are included in two or more of the plurality of hypothesis transcriptions, each of the two or more of the plurality of hypothesis transcriptions including words from a first portion of the audio data;

in response to determining the one or more consistent words, providing the one or more consistent words to a device for presentation of the one or more consistent words by the device, the presentation of the one or more consistent words configured to occur before a final transcription of the audio data is provided to the device; and

presenting the one or more consistent words in a rolling fashion, a pace of the presentation of the one or more consistent words in the rolling fashion being variable such that when more words are to be presented the words are presented quicker than when fewer words are to be presented.

14. The method of claim 13 , wherein the plurality of hypothesis transcriptions are obtained from a speech recognition system that includes a plurality of speech engines configured to recognize speech.

15. The method of claim 13 , wherein the plurality of hypothesis transcriptions are obtained from a speech recognition system that includes a single speech engine configured to recognize speech.

16. The method of claim 13 , wherein a first portion of the audio data associated with a first one of the plurality of hypothesis transcriptions includes all of the audio data associated with at least one of the plurality of hypothesis transcriptions obtained previous to obtaining the first one of the plurality of hypothesis transcriptions.

17. The method of claim 13 , wherein the plurality of hypothesis transcriptions are obtained sequentially over time and the determining the one or more consistent words includes comparing a first hypothesis transcription of the plurality of hypothesis transcriptions with a second hypothesis transcription of the plurality of hypothesis transcriptions, the second hypothesis transcription directly following the first hypothesis transcription among the plurality of hypothesis transcriptions.

18. The method of claim 13 , further comprising:

determining an update word in the final transcription that is different from any of the one or more consistent words; and

in response to determining the update word and after directing the one or more consistent words to the device, directing an indication of the update word to the device, the update word replacing one or more of the one or more consistent words in the presentation of the one or more consistent words on the device.

19. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 13 .

Assignments (6)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2021
From: CHEVRIER, BRIAN; ROYLANCE, SHANE; BOEHME, KENNETH
To: CAPTIONCALL, LLC
Reel/Frame 057653/0860 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2021
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 057653/0872 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
Continuity (2)
Continuation 16154553 · Oct 8, 2018
Related Publication 20210398538A1 · Dec 23, 2021