IP Library Granted Patent US 10,504,519
Granted Patent B1
US 10,504,519 · App. 16/370,214 · Granted Dec 10, 2019

Transcription of communications

Inventors: Brian Chevrier (Highland, UT); Shane Roylance (Farmington, UT); Kenneth Boehme (South Jordan, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G10L15/06G10L15/08G10L2015/088
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,504,519
App. No.
16/370,214
Granted
Dec 10, 2019
Kind
B1
Abstract

A method to transcribe communications may include obtaining audio data originating at a first device during a communication session between the first device and a second device and providing the audio data to an automated speech recognition system configured to transcribe the audio data. The method may further include obtaining multiple hypothesis transcriptions generated by the automated speech recognition system. Each of the multiple hypothesis transcriptions may include one or more words determined by the automated speech recognition system to be a transcription of a portion of the audio data. The method may further include determining one or more consistent words that are included in two or more of the multiple hypothesis transcriptions and in response to determining the one or more consistent words, providing the one or more consistent words to the second device for presentation of the one or more consistent words by the second device.

Claims (38)

1. A method to transcribe communications, the method comprising:

obtaining a plurality of hypothesis transcriptions of audio data generated by a speech recognition system;

determining a plurality of consistent words that are included in two or more of the plurality of hypothesis transcriptions;

in response to determining the plurality of consistent words, directing the plurality of consistent words to a device for presentation of the plurality of consistent words; and

presenting the plurality of consistent words in a rolling fashion, a pace of the presentation of the plurality of consistent words in the rolling fashion being variable.

2. The method of claim 1 , further comprising:

determining an update word in a final transcription of the audio data that is different from any of the plurality of consistent words; and

in response to determining the update word, directing an indication of the update word to the device, the update word replacing one or more of the plurality of consistent words in the presentation of the plurality of consistent words.

3. The method of claim 1 , wherein the presentation of the plurality of consistent words by the device is configured to occur before a final transcription of the audio data is provided to the device.

4. The method of claim 1 , wherein the variable pace is determined to allow for all of the plurality of consistent words to be presented in less than a threshold period of time.

5. The method of claim 1 , wherein the variable pace is determined based on one or more of: a display screen size of the device and an average talking speed of a user associated with the audio data.

6. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 1 .

7. A system comprising:

at least one processor; and

at least one non-transitory computer-readable media communicatively coupled to the at least one processor and configured to store one or more instructions that when executed by the at least one processor cause or direct the system to perform operations comprising:

obtain a plurality of hypothesis transcriptions of audio data generated by a speech recognition system;

determine one or more consistent words that are included in two or more of the plurality of hypothesis transcriptions;

in response to determining the one or more consistent words, direct the one or more consistent words to a device for presentation of the one or more consistent words;

determine an update word in a final transcription of the audio data that is different from any of the one or more consistent words; and

in response to determining the update word, direct an indication of the update word to the device, the update word replacing the one or more consistent words in the presentation of the one or more consistent words.

8. The system of claim 7 , wherein the operations further comprise obtain the audio data during a communication session between a second device and the device, the audio data originating at the second device.

9. The system of claim 7 , wherein the presentation of the one or more consistent words by the device is configured to occur before the final transcription of the audio data is provided to the device.

10. The system of claim 7 , wherein the speech recognition system is included in the system.

11. The system of claim 7 , wherein the final transcription includes a plurality of words that the speech recognition system outputs together as the final transcription of the audio data that includes one or more of the one or more consistent words.

12. The system of claim 7 , wherein the one or more consistent words are presented in a rolling fashion, a pace of the presentation of the one or more consistent words in the rolling fashion being variable.

13. The system of claim 7 , wherein the plurality of hypothesis transcriptions are obtained sequentially over time and the operation to determine the one or more consistent words includes compare a first hypothesis transcription of the plurality of hypothesis transcriptions with a second hypothesis transcription of the plurality of hypothesis transcriptions, the second hypothesis transcription directly following the first hypothesis transcription among the plurality of hypothesis transcriptions.

14. A method to transcribe communications, the method comprising:

obtaining a plurality of hypothesis transcriptions of audio data generated by a speech recognition system;

determining one or more consistent words that are included in two or more of the plurality of hypothesis transcriptions;

in response to determining the one or more consistent words, directing the one or more consistent words to a device for presentation of the one or more consistent words;

determining an update word in a final transcription of the audio data that is different from any of the one or more consistent words; and

in response to determining the update word, directing an indication of the update word to the device, the update word replacing the one or more consistent words in the presentation of the one or more consistent words.

15. The method of claim 14 , further comprising obtaining the audio data during a communication session between a second device and the device, the audio data originating at the second device.

16. The method of claim 14 , wherein the presentation of the one or more consistent words by the device is configured to occur before the final transcription of the audio data is provided to the device.

17. The method of claim 14 , wherein the final transcription includes a plurality of words that the speech recognition system outputs together as the final transcription of the audio data that includes one or more of the one or more consistent words.

18. The method of claim 14 , wherein the one or more consistent words are presented in a rolling fashion, a pace of the presentation of the one or more consistent words in the rolling fashion being variable.

19. The method of claim 14 , wherein the plurality of hypothesis transcriptions are obtained sequentially over time and determining the one or more consistent words includes comparing a first hypothesis transcription of the plurality of hypothesis transcriptions with a second hypothesis transcription of the plurality of hypothesis transcriptions, the second hypothesis transcription directly following the first hypothesis transcription among the plurality of hypothesis transcriptions.

20. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 14 .

Assignments (9)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
RELEASE OF SECURITY INTEREST Recorded Dec 16, 2021
From: CORTLAND CAPITAL MARKET SERVICES LLC
To: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 058533/0467 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
LIEN Recorded Feb 11, 2020
From: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CORTLAND CAPITAL MARKET SERVICES LLC
Reel/Frame 051894/0665 →
PATENT SECURITY AGREEMENT Recorded Apr 29, 2019
From: SORENSEN COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
Reel/Frame 050084/0793 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 16, 2019
From: CHEVRIER, BRIAN; ROYLANCE, SHANE; BOEHME, KENNETH
To: CAPTIONCALL, LLC
Reel/Frame 048896/0596 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 16, 2019
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 048896/0639 →