IP Library Granted Patent US 10,192,554
Granted Patent B1
US 10,192,554 · App. 15/905,180 · Granted Jan 29, 2019

Transcription of communications using multiple speech recognition systems

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,192,554
App. No.
15/905,180
Granted
Jan 29, 2019
Kind
B1
Abstract

A method may include obtaining audio data originating at a first device during a communication session between the first device and a second device and providing the audio data to a first speech recognition system to generate a first transcript based on the audio data and directing the first transcript to the second device. The method may also include in response to obtaining a quality indication regarding a quality of the first transcript, multiplexing the audio data to provide the audio data to a second speech recognition system to generate a second transcript based on the audio data while continuing to provide the audio data to the first speech recognition system and direct the first transcript to the second device, and in response to obtaining a transfer indication that occurs after multiplexing of the audio data, directing the second transcript to the second device instead of the first transcript.

Claims (63)

1. A method to transcribe communications, the method comprising:

obtaining audio data originating at a first device during a communication session between the first device and a second device, the communication session configured for verbal communication;

providing the audio data to a first automated speech recognition system that works independent of human interaction to generate a first transcript using the audio data;

directing the first transcript to the second device;

in response to obtaining a quality indication that indicates a quality of the first transcript is below a quality threshold and while continuing to provide the audio data to the first automated speech recognition system to generate the first transcript and continuing to direct the first transcript to the second device, the method including:

multiplexing the audio data to provide the audio data to the first automated speech recognition system and a second automated speech recognition system;

broadcasting, by the second automated speech recognition system, audio based on the multiplexed audio data;

obtaining, by the second automated speech recognition system, second audio data based on a re-voicing of the broadcast audio; and

generating, by the second automated speech recognition system, a second transcript using the second audio data; and

in response to a transfer indication that occurs after multiplexing of the audio data, directing the second transcript to the second device instead of directing the first transcript to the second device and ceasing providing the audio data to the first automated speech recognition system.

2. The method of claim 1 , further comprising:

obtaining a confidence score of the first transcript from the first automated speech recognition system; and

obtaining the quality indication based on a comparison of the confidence score to the quality threshold.

3. The method of claim 1 , wherein the quality indication is obtained from the second device.

4. The method of claim 1 , wherein the transfer indication includes an occurrence of one of the following:

after multiplexing of the audio data, for a time period the audio data does not include spoken words that result in text in the first transcript,

the second transcript including text,

after multiplexing of the audio data, the first transcript includes first text and then for a second time period the audio data does not include spoken words that result in text in the first transcript, and

a last phrase of the first transcript is the same as a last phrase of the second transcript.

5. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 1 .

6. A method to transcribe communications, the method comprising:

obtaining audio data originating at a first device during a communication session between the first device and a second device, the communication session configured for verbal communication;

providing the audio data to a first speech recognition system to generate a first transcript based on the audio data;

directing the first transcript to the second device;

in response to obtaining a quality indication that indicates a quality of the first transcript is below a quality threshold, multiplexing the audio data to provide the audio data to a second speech recognition system to generate a second transcript based on the audio data while continuing to provide the audio data to the first speech recognition system to generate the first transcript and continuing to direct the first transcript to the second device; and

in response to a transfer indication that occurs after multiplexing of the audio data and that results from an occurrence of an event that indicates the second transcript is to be directed to the second device instead of the first transcript, directing the second transcript to the second device instead of the first transcript.

7. The method of claim 6 , wherein the first speech recognition system and the second speech recognition system are automated speech recognition systems that work independent of human interaction.

8. The method of claim 6 , wherein the first speech recognition system is an automated speech recognition system that works independent of human interaction and the generation of the second transcript by the second speech recognition system includes:

broadcasting audio based on the audio data; and

obtaining second audio data based on a re-voicing of the broadcast audio, wherein the second transcript is generated based on the second audio data.

9. The method of claim 6 , further comprising:

obtaining a confidence score of the first transcript from the first speech recognition system; and

obtaining the quality indication based on a comparison of the confidence score to the quality threshold.

10. The method of claim 6 , wherein the quality indication is obtained from the second device.

11. The method of claim 6 , wherein the transfer indication includes an occurrence of one of the following:

after multiplexing of the audio data, for a time period the audio data does not include spoken words that result in text in the first transcript,

the second transcript including text,

after multiplexing of the audio data, the first transcript includes first text and then for a second time period the audio data does not include spoken words that result in text in the first transcript, and

a last phrase of the first transcript is the same as a last phrase of the second transcript.

12. The method of claim 6 , further comprising in response to obtaining a transfer indication that occurs after multiplexing of the audio data, ceasing providing the audio data to the first speech recognition system.

13. At least one non-transitory computer-readable media configured to store one or more instructions that when executed by at least one processor cause or direct a system to perform the method of claim 6 .

14. A system comprising:

at least one processor; and

at least one non-transitory computer-readable media communicatively coupled to the at least one processor and configured to store one or more instructions that when executed by the at least one processor cause or direct the system to perform operations comprising:

obtain audio data originating at a first device during a communication session between the first device and a second device, the communication session configured for verbal communication;

provide the audio data to a first speech recognition system to generate a first transcript based on the audio data;

direct the first transcript to the second device;

in response to obtaining a quality indication that indicates a quality of the first transcript is below a quality threshold, multiplex the audio data to provide the audio data to a second speech recognition system to generate a second transcript based on the audio data while continuing to provide the audio data to the first speech recognition system to generate the first transcript and continuing to direct the first transcript to the second device; and

in response to a transfer indication that occurs after multiplexing of the audio data and that results from an occurrence of an event that indicates the second transcript is to be directed to the second device instead of the first transcript, direct the second transcript to the second device instead of the first transcript.

15. The system of claim 14 , wherein the first speech recognition system and the second speech recognition system are automated speech recognition systems that work independent of human interaction.

16. The system of claim 14 , wherein the first speech recognition system is an automated speech recognition system that works independent of human interaction and the generation of the second transcript by the second speech recognition system includes operations comprising:

broadcast audio based on the audio data; and

obtain second audio data based on a re-voicing of the broadcast audio, wherein the second transcript is generated based on the second audio data.

17. The system of claim 14 , wherein the operations further comprise:

obtaining a confidence score of the first transcript from the first speech recognition system; and

obtaining the quality indication based on a comparison of the confidence score to the quality threshold.

18. The system of claim 14 , wherein the quality indication is obtained from the second device.

19. The system of claim 14 , wherein the transfer indication includes an occurrence of one of the following:

after multiplexing of the audio data, for a time period the audio data does not include spoken words that result in text in the first transcript,

the second transcript including text,

after multiplexing of the audio data, the first transcript includes first text and then for a second time period the audio data does not include spoken words that result in text in the first transcript, and

a last phrase of the first transcript is the same as a last phrase of the second transcript.

20. The system of claim 14 , wherein the operations further comprise in response to obtaining a transfer indication that occurs after multiplexing of the audio data, cease providing the audio data to the first speech recognition system.

Assignments (11)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
RELEASE OF SECURITY INTEREST Recorded Dec 16, 2021
From: CORTLAND CAPITAL MARKET SERVICES LLC
To: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 058533/0467 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
LIEN Recorded Feb 11, 2020
From: SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CORTLAND CAPITAL MARKET SERVICES LLC
Reel/Frame 051894/0665 →
RELEASE OF SECURITY INTEREST Recorded May 8, 2019
From: U.S. BANK NATIONAL ASSOCIATION
To: SORENSON COMMUNICATIONS, LLC; SORENSON IP HOLDINGS, LLC; CAPTIONCALL, LLC; INTERACTIVECARE, LLC
Reel/Frame 049115/0468 →
RELEASE OF SECURITY INTEREST Recorded May 7, 2019
From: JPMORGAN CHASE BANK, N.A.
To: SORENSON COMMUNICATIONS, LLC; SORENSON IP HOLDINGS, LLC; CAPTIONCALL, LLC; INTERACTIVECARE, LLC
Reel/Frame 049109/0752 →
PATENT SECURITY AGREEMENT Recorded Apr 29, 2019
From: SORENSEN COMMUNICATIONS, LLC; CAPTIONCALL, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
Reel/Frame 050084/0793 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 22, 2018
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 045870/0598 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 26, 2018
From: BOEHME, KENNETH; HOLM, MICHAEL; ROYLANCE, SHANE
To: CAPTIONCALL, LLC
Reel/Frame 045042/0366 →
Cited By (1)
US 12,555,593