IP Library Granted Patent US 11,741,964
Granted Patent B2
US 11,741,964 · App. 16/885,039 · Granted Aug 29, 2023

Transcription generation technique selection

Inventor: David Thomson (Bountiful, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G06F18/2178G10L15/30G10L21/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,741,964
App. No.
16/885,039
Granted
Aug 29, 2023
Kind
B2
Abstract

A method to transcribe communications may include selecting a first transcription generation technique from among multiple transcription generation techniques for generating transcriptions of audio of one or more communication sessions that involve a user device and obtaining performances of the multiple transcription generation techniques with respect to generating the transcriptions of the audio. The method may also include monitoring comparisons between the performances of the multiple transcription generation techniques and obtaining input from the user with respect to the comparisons. The method may further include selecting a second transcription generation technique from among the multiple transcription generation techniques based on the input from the user.

Claims (36)

1. A method to transcribe communications, the method comprising:

obtaining a performance of a first transcription generation technique with respect to generating transcriptions of audio of a first communication session associated with a user;

obtaining a performance of a second transcription generation technique with respect to generating transcriptions of the audio of the first communication session;

determining a report based on the performance of the first transcription generation technique and the performance of the second transcription generation technique;

directing the report to a first device associated with the user;

in response to the report, obtaining an indication from the first device; and

directing a transcription of a second communication session to a second device for presentation to the user, the transcription being generated by the second transcription generation technique in response to the indication from the first device, wherein the second communication session is independent and distinct from the first communication session and an entirety of the second communication session occurs after the first communication session.

2. The method of claim 1 , wherein the performance of the second transcription generation technique is based on one or more of the following: transcription accuracy, transcription latency, and number of transcription corrections.

3. The method of claim 2 , wherein the performance of the second transcription generation technique is based on a plurality of communications sessions that include the first communication session.

4. The method of claim 1 , wherein the report includes a recommendation for the second transcription generation technique and the indication includes a selection of the second transcription generation technique.

5. The method of claim 1 , wherein the first device and the second device are the same device.

6. The method of claim 1 , further comprising before determining the report, directing a second transcription of the first communication session that involves the second device to the second device, the second transcription generated by the first transcription generation technique.

7. The method of claim 6 , wherein the steps of directing the report and of obtaining the indication occur during the first communication session.

8. The method of claim 1 , wherein one of the first transcription generation technique and the second transcription generation technique includes a revoicing of audio before transcription generation.

9. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 1 .

10. A method to transcribe communications, the method comprising:

selecting a first transcription generation technique from among a plurality of transcription generation techniques for generating transcriptions of audio of one or more communication sessions that involve a user device of a user;

obtaining performances of the plurality of transcription generation techniques with respect to generating the transcriptions of the audio;

monitoring comparisons between the performances of the plurality of transcription generation techniques;

obtaining input from the user with respect to the comparisons; and

selecting a second transcription generation technique from among the plurality of transcription generation techniques based on the input from the user, the selected second transcription generation technique being used for generating transcriptions of audio of a second communication session that involves the user device, the second communication session occurring after the one or more communication sessions.

11. The method of claim 10 , wherein the performances of the plurality of transcription generation techniques are based on one or more of the following: transcription accuracy and transcription latency.

12. The method of claim 10 , further comprising directing a report to the user based on the comparison, wherein the input is obtained in response to report.

13. The method of claim 10 , wherein the second transcription generation technique does not generate a transcription of the audio such that the performance of the second transcription generation technique is an estimated performance.

14. The method of claim 10 , wherein the selection of the first transcription generation technique is based on the performance of the first transcription generation technique.

15. The method of claim 10 , wherein the monitoring comparisons between the performances of the plurality of transcription generation techniques occur with respect to a first communication session that involves the user device.

16. The method of claim 15 , the second transcription generation technique is selected to generate transcriptions of audio of the first communication session during the first communication session.

17. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 10 .

18. A system comprising:

one or more processors; and

one or more non-transitory computer-readable mediums configured to store instructions that when executed by the processors cause or direct the system to perform operations, the operations comprising:

select a first transcription generation technique from among a plurality of transcription generation techniques for generating transcriptions of audio of one or more communication sessions that involve a user device;

obtain performances of the plurality of transcription generation techniques with respect to generating the transcriptions of the audio;

monitor comparisons between the performances of the plurality of transcription generation techniques;

obtain input from the user device with respect to the comparisons; and

select a second transcription generation technique from among the plurality of transcription generation techniques based on the input from the user device, the selected second transcription generation technique being used for generating transcriptions of audio of a second communication session that involves the user device, the second communication session occurring after the one or more communication sessions.

Assignments (7)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 19, 2020
From: BOEKWEG, SCOTT; CHEVRIER, BRIAN; MONTERO, ADAM; PETERSON, BRUCE; ROYLANCE, SHANE
To: CAPTIONCALL, LLC
Reel/Frame 052986/0491 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 19, 2020
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 052986/0456 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 28, 2020
From: THOMSON, DAVID
To: CAPTIONCALL, LLC
Reel/Frame 052770/0801 →