IP Library Granted Patent US 12,266,366
Granted Patent B2
US 12,266,366 · App. 18/457,184 · Granted Apr 1, 2025

Transcription generation technique selection

Inventor: David Thomson (Bountiful, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G06F18/2178G10L15/30G10L21/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,266,366
App. No.
18/457,184
Granted
Apr 1, 2025
Kind
B2
Abstract

A method to transcribe communications may include selecting a first transcription generation technique from among multiple transcription generation techniques for generating transcriptions of audio of one or more communication sessions that involve a user device and obtaining performances of the multiple transcription generation techniques with respect to generating the transcriptions of the audio. The method may also include monitoring comparisons between the performances of the multiple transcription generation techniques and obtaining input from the user with respect to the comparisons. The method may further include selecting a second transcription generation technique from among the multiple transcription generation techniques based on the input from the user.

Claims (49)

1. A method to transcribe communications, the method comprising:

obtaining a performance of at least one of a plurality of transcription generation techniques with respect to generating transcriptions of audio;

determining a report based on the performance of the at least one of the plurality of transcription generation techniques, wherein the performance of the at least one of the plurality of transcription generation techniques is based on one or more of the following: transcription accuracy and transcription latency;

directing the report to a device that obtains transcriptions of a first audio session involving the device using one of the plurality of transcription generation techniques, the performance of the at least one of the plurality of transcription generation techniques is based on the transcriptions of the first audio session;

after directing the report, obtaining an indication from the device; and

selecting, based on the indication from the device, another one of the plurality of transcription generation techniques to generate transcriptions of a future audio session involving the device that occurs after the first audio session.

2. The method of claim 1 , wherein the report includes a recommendation for the other one of the plurality of transcription generation techniques and the indication includes a selection of the other one of the plurality of transcription generation techniques.

3. The method of claim 2 , further comprising determining the recommendation based on a cost associated with the one of the plurality of transcription generation techniques and a cost associated with the other one of the plurality of transcription generation techniques.

4. The method of claim 1 , wherein the steps of directing the report and of obtaining the indication occur during the first audio session.

5. The method of claim 1 , further comprising:

obtaining a performance of the other one of the plurality of transcription generation techniques with respect to generating transcriptions of audio; and

comparing the performance of the other one of the plurality of transcription generation techniques and the performance of the at least one of the plurality of transcription generation techniques,

wherein the report is based on the comparison.

6. The method of claim 5 , wherein comparing the performance of the other one of the plurality of transcription generation techniques and the performance of the at least one of the plurality of transcription generation techniques includes comparing two or more aspects of the performance of the other one of the plurality of transcription generation techniques with two or more aspects of the performance of the at least one of the plurality of transcription generation techniques,

wherein the report includes aspects of the performance of the other one of the plurality of transcription generation techniques that are better than aspects of the performance of the at least one of the plurality of transcription generation techniques.

7. The method of claim 1 , wherein one of the plurality of transcription generation techniques includes a revoicing of audio before transcription generation.

8. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 1 .

9. A method to transcribe communications, the method comprising:

selecting a first transcription generation technique from among a plurality of transcription generation techniques for generating transcriptions of audio obtained by a device during a first audio session;

obtaining a performance of at least one of the plurality of transcription generation techniques with respect to generating transcriptions;

after presentation of the transcriptions of the audio by the device, obtaining input from a user of the device regarding the performance of at least one of the plurality of transcription generation techniques with respect to generating transcriptions of the first audio session; and

selecting a second transcription generation technique from among the plurality of transcription generation techniques in response to the input from the device for a second audio session that occurs after the first audio session, the second transcription generation technique being used for generating transcriptions of second audio obtained by the device during the second audio session.

10. The method of claim 9 , wherein the performance of the at least one of the plurality of transcription generation techniques is based on one or more of: transcription accuracy and transcription latency.

11. The method of claim 9 , wherein the performance of the at least one of the plurality of transcription generation techniques is measured without consideration of the transcriptions generated during the first audio session.

12. The method of claim 9 , wherein the selection of the first transcription generation technique occurs before the second audio session begins.

13. The method of claim 9 , wherein the performance of the at least one of the plurality of transcription generation techniques is the performance of the first transcription generation technique, the method further comprising:

obtaining performance of the second transcription generation technique with respect to generating transcriptions; and

before obtaining the input from the device, providing a report to the device based on a comparison between the performance of the first transcription generation technique and the performance of the second transcription generation technique.

14. The method of claim 13 , further comprising:

determining an aspect of the performance of the second transcription generation technique that is better than a same aspect of the performance of the first transcription generation technique; and

generating the report to include the aspect of the performance of the second transcription generation technique.

15. The method of claim 13 , wherein the second transcription generation technique does not generate a transcription of audio such that the performance of the second transcription generation technique is an estimated performance.

16. The method of claim 9 , wherein the performance of at least one of the plurality of transcription generation techniques is the performance of the second transcription generation technique.

17. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 9 .

18. A system comprising:

one or more processors; and

one or more non-transitory computer-readable mediums configured to store instructions that when executed by the processors cause or direct the system to perform operations, the operations comprising:

obtaining a performance of at least one of a plurality of transcription generation techniques with respect to generating transcriptions of audio;

determining a report based on the performance of the at least one of the plurality of transcription generation techniques, wherein the performance of the at least one of the plurality of transcription generation techniques is based on one or more of the following: transcription accuracy and transcription latency;

directing the report to a device that obtains transcriptions of a first audio session involving the device using one of the plurality of transcription generation techniques, the performance of the at least one of the plurality of transcription generation techniques is based on the transcriptions of the first audio session;

after directing the report, obtaining an indication from the device; and

selecting, based on the indication from the device, another one of the plurality of transcription generation techniques to generate transcriptions of a future audio session involving the device that occurs after the first audio session.

19. A method to transcribe communications, the method comprising:

obtaining a performance of at least one of a plurality of transcription generation techniques with respect to generating transcriptions of audio;

determining a report based on the performance of the at least one of the plurality of transcription generation techniques, wherein the performance of the at least one of the plurality of transcription generation techniques is based on one or more of the following: transcription accuracy and transcription latency;

directing the report to a device that obtains transcriptions of a first audio session involving the device using one of the plurality of transcription generation techniques, the performance of the at least one of the plurality of transcription generation techniques is based on the transcriptions of the first audio session;

after directing the report, obtaining an indication from the device; and

selecting, based on the indication from the device, another one of the plurality of transcription generation techniques to generate transcriptions of a future audio session involving the device that occurs after the first audio session.

20. The method of claim 19 , wherein the report includes a recommendation for the other one of the plurality of transcription generation techniques and the indication includes a selection of the other one of the plurality of transcription generation techniques.

Assignments (4)
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 30, 2023
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 064747/0960 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 30, 2023
From: BOEKWEG, SCOTT; CHEVRIER, BRIAN; MONTERO, ADAM; PETERSON, BRUCE; ROYLANCE, SHANE
To: CAPTIONCALL, LLC
Reel/Frame 064747/0962 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 30, 2023
From: THOMSON, DAVID
To: CAPTIONCALL, LLC
Reel/Frame 064747/0969 →
Continuity (2)
Continuation 16885039 · May 27, 2020
Related Publication 20230410815A1 · Dec 21, 2023
References Cited (50)
US 5875436A · Kikinis · 1999 [cited by applicant]
US 6185535B1 · Hedin · 2001 [cited by examiner]
US 7228275B1 · Endo et al. · 2007 [cited by applicant]
US 7908145B2 · Bennett et al. · 2011 [cited by applicant]
US 8682672B1 · Ha et al. · 2014 [cited by applicant]
US 9318110B2 · Roe · 2016 [cited by examiner]
US 9443518B1 · Gauci · 2016 [cited by examiner]
US 9497315B1 · Pakidko · 2016 [cited by examiner]
US 9704111B1 · Antunes · 2017 [cited by examiner]
US 9715876B2 · Hager · 2017 [cited by examiner]
US 9736309B1 · Bentitou · 2017 [cited by examiner]
US 9773501B1 · Brooksby · 2017 [cited by examiner]
US 9967380B2 · Engelke et al. · 2018 [cited by applicant]
US 10224057B1 · Chevrier · 2019 [cited by applicant]
US 10389876B2 · Engelke et al. · 2019 [cited by applicant]
US 10542141B2 · Engelke et al. · 2020 [cited by applicant]
US 10917519B2 · Engelke et al. · 2021 [cited by applicant]
US 10971157B2 · Willett · 2021 [cited by examiner]
US 11315569B1 · Talieh et al. · 2022 [cited by applicant]
US 11620566B1 · Shevchenko et al. · 2023 [cited by applicant]
US 20010005825A1 · Engelke et al. · 2001 [cited by applicant]
US 20060149558A1 · Kahn et al. · 2006 [cited by applicant]
US 20110054892A1 · Jung · 2011 [cited by examiner]
US 20120221321A1 · Nakamura · 2012 [cited by examiner]
US 20130066630A1 · Roe · 2013 [cited by examiner]
US 20150073790A1 · Steuble · 2015 [cited by examiner]
US 20150340036A1 · Weeks · 2015 [cited by applicant]
US 20170201613A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206808A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206888A1 · Engelke et al. · 2017 [cited by applicant]
US 20170206914A1 · Engelke et al. · 2017 [cited by applicant]
US 20170208172A1 · Engelke et al. · 2017 [cited by applicant]
US 20180034961A1 · Engelke et al. · 2018 [cited by applicant]
US 20180270350A1 · Engelke et al. · 2018 [cited by applicant]
US 20180285059A1 · Zurek et al. · 2018 [cited by applicant]
US 20190037072A1 · Engelke et al. · 2019 [cited by applicant]
US 20190312973A1 · Engelke et al. · 2019 [cited by applicant]
US 20190333517A1 · Nelson · 2019 [cited by applicant]
US 20200007679A1 · Engelke et al. · 2020 [cited by applicant]
US 20200075013A1 · Holm · 2020 [cited by applicant]
US 20200153957A1 · Engelke et al. · 2020 [cited by applicant]
US 20200153958A1 · Engelke et al. · 2020 [cited by applicant]
US 20200243094A1 · Thomson · 2020 [cited by examiner]
US 20200244800A1 · Engelke et al. · 2020 [cited by applicant]
US 20200252507A1 · Engelke et al. · 2020 [cited by applicant]
US 20200394258A1 · Chen et al. · 2020 [cited by applicant]
US 20210056950A1 · Niehaus et al. · 2021 [cited by applicant]
US 20210224695A1 · Stefanov et al. · 2021 [cited by applicant]
US 20210295826A1 · Morabia et al. · 2021 [cited by applicant]
InnoCaption, ASR Development Update, Feb. 27, 2020 and Mar. 2, 2020. [cited by applicant]