IP Library Granted Patent US 11,783,837
Granted Patent B2
US 11,783,837 · App. 16/950,653 · Granted Oct 10, 2023

Transcription generation technique selection

Inventor: Michael Holm (Bountiful, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G10L15/01G10L15/22G10L15/08G10L15/14G10L15/16G10L15/18G10L15/28G10L15/30H04L67/10
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,783,837
App. No.
16/950,653
Granted
Oct 10, 2023
Kind
B2
Abstract

According to one or more aspects of the present disclosure, operations related to selecting a transcription generation technique may be disclosed. In some embodiments, the operations may include obtaining multiple user ratings that each correspond to a different one of multiple transcriptions. Each transcription may be obtained using a first transcription generation technique and may correspond to a different one of multiple communication sessions. The operations may further include selecting, for a subsequent communication session that occurs after the multiple communication sessions, a second transcription generation technique based on the user ratings. In addition, the operations may include providing the subsequent transcription to a device during the subsequent communication session.

Claims (61)

1. A method comprising:

obtaining, at a system, a first rating of a first transcription of audio from a second user device during a first communication session involving a first user device and the second user device, the first rating indicating feedback from the first user device regarding the first transcription and the first transcription obtained using a first transcription generation technique, wherein the system is not part of the first communication session;

after termination of the first communication session, obtaining, at the system, a second rating of a second transcription of audio from a third user device during a second communication session involving the first user device and the third user device, where the third user device is different from the second user device, the second rating indicating feedback from the first user device regarding the second transcription and the second transcription obtained using the first transcription generation technique, wherein the system and the second user device are not part of the second communication session;

selecting, using the first and second ratings, a second transcription generation technique for a third communication session involving the first user device and a fourth user device that is different from the second user device and the third user device, the second transcription generation technique selected to generate a third transcription of audio from the fourth user device; and

obtaining, during the third communication session, the third transcription that is generated using the second transcription generation technique.

2. The method of claim 1 , wherein:

the first transcription generation technique includes using a fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a re-voicing speech recognition system to generate the third transcription.

3. The method of claim 1 , wherein:

the first transcription generation technique includes using a first fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a second fully machine based automatic speech recognition system to generate the third transcription.

4. The method of claim 1 , wherein obtaining the first rating includes:

directing presentation of a request for the first rating on the first user device after termination of the first communication session; and

obtaining the first rating from the first user device based on a response to the request.

5. The method of claim 1 , wherein selecting the second transcription generation technique includes:

determining a number of negative ratings of a plurality of ratings from the first user device, which include the first rating and the second rating, that meet a negative rating threshold; and

selecting the second transcription generation technique based on the number of negative ratings satisfying a number threshold.

6. The method of claim 1 , wherein selecting the second transcription generation technique includes:

determining a ratio of ratings that meet a negative rating threshold with respect to a plurality of ratings from the first user device that include the first rating and the second rating; and

selecting the second transcription generation technique based on the ratio of ratings satisfying a ratio threshold.

7. The method of claim 1 , wherein selecting the second transcription generation technique includes:

determining an average rating of a plurality of ratings from the first user device that include the first rating and the second rating; and

selecting the second transcription generation technique based on the average rating satisfying a negative rating threshold.

8. At least one non-transitory computer-readable media configured to store one or more instructions that in response to being executed by at least one computing system cause performance of the method of claim 1 .

9. A method comprising:

obtaining, at a system, a first rating of a first transcription of audio from a second user device during a first communication session involving a first user device and the second user device, the first rating indicating a quality of the first transcription, the first transcription obtained using a first transcription generation technique, wherein the system is not part of the first communication session;

obtaining, at the system, a second rating of a second transcription of audio from a third user device during a second communication session involving the first user device and the third user device, where the third user device is different from the second user device, the second rating indicating a quality of the second transcription, the second transcription obtained using the first transcription generation technique, wherein the system and the second user device are not part of the second communication session;

selecting, using the first and second ratings, a second transcription generation technique for a third communication session involving a fourth user device that is different from the second user device and the third user device, the second transcription generation technique selected to generate a third transcription of audio from the fourth user device; and

obtaining, during the third communication session, the third transcription that is generated using the second transcription generation technique.

10. The method of claim 9 , wherein:

the first transcription generation technique includes using a fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a re-voicing speech recognition system to generate the third transcription.

11. The method of claim 9 , wherein:

the first transcription generation technique includes using a first fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a second fully machine based automatic speech recognition system to generate the third transcription.

12. The method of claim 9 , wherein obtaining the first rating includes:

directing presentation of a request for the first rating on the first user device after termination of the first communication session; and

obtaining the first rating from the first user device based on a response to the request.

13. The method of claim 9 , wherein selecting the second transcription generation technique includes:

determining a number of negative ratings of a plurality of ratings from the first user device, which include the first rating and the second rating, that meet a negative rating threshold; and

selecting the second transcription generation technique based on the number of negative ratings satisfying a number threshold.

14. The method of claim 9 , wherein selecting the second transcription generation technique includes:

determining a ratio of ratings that meet a negative rating threshold with respect to a plurality of ratings from the first user device that include the first rating and the second rating; and

selecting the second transcription generation technique based on the ratio of ratings satisfying a ratio threshold.

15. The method of claim 9 , wherein selecting the second transcription generation technique includes:

determining an average rating of a plurality of ratings from the first user device that include the first rating and the second rating; and

selecting the second transcription generation technique based on the average rating satisfying a negative rating threshold.

16. At least one non-transitory computer-readable media configured to store one or more instructions that in response to being executed by at least one computing system cause performance of the method of claim 9 .

17. A system comprising:

one or more processors; and

one or more computer-readable media configured to store instructions that in response to being executed by the one or more processors cause the system to perform operations, the operations comprising:

obtain a first rating of a first transcription of audio from a second user device during a first communication session involving a first user device and the second user device, the first rating indicating a quality of the first transcription, the first transcription obtained using a first transcription generation technique, wherein the system is not part of the first communication session;

obtain a second rating of a second transcription of audio from a third user device during a second communication session involving the first user device and the third user device, where the third user device is different from the second user device, the second rating indicating a quality of the second transcription, the second transcription obtained using the first transcription generation technique, wherein the system and the second user device are not part of the second communication session;

select, based on the first and second ratings, a second transcription generation technique for a third communication session involving the first user device and a fourth user device that is different from the second user device and the third user device, the second transcription generation technique selected to generate a third transcription of audio from the fourth user device; and

obtain, during the third communication session, the third transcription that is generated using the second transcription generation technique.

18. The system of claim 17 , wherein:

the first transcription generation technique includes using a fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a re-voicing speech recognition system to generate the third transcription.

19. The system of claim 17 , wherein:

the first transcription generation technique includes using a first fully machine based automatic speech recognition system to generate the first transcription and the second transcription; and

the second transcription generation technique includes using a second fully machine based automatic speech recognition system to generate the third transcription.

Assignments (6)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 17, 2020
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 054395/0676 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 17, 2020
From: HOLM, MICHAEL
To: CAPTIONCALL, LLC
Reel/Frame 054395/0678 →