IP Library Granted Patent US 11,176,944
Granted Patent B2
US 11,176,944 · App. 16/408,826 · Granted Nov 16, 2021

Transcription summary presentation

Inventors: Scott Boekweg (South Jordan, UT); David Thomson (North Salt Lake, UT)
Assignee: Sorenson IP Holdings, LLC
G10L15/26G10L15/30G10L21/18
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,176,944
App. No.
16/408,826
Granted
Nov 16, 2021
Kind
B2
Abstract

A method to present a summary of a transcription may include obtaining, at a first device, audio directed to the first device from a second device during a communication session between the first device and the second device. Additionally, the method may include sending, from the first device, the audio to a transcription system. The method may include obtaining, at the first device, a transcription during the communication session from the transcription system based on the audio. Additionally, the method may include obtaining, at the first device, a summary of the transcription during the communication session. Additionally, the method may include presenting, on a display, both the summary and the transcription simultaneously during the communication session.

Claims (41)

1. A method comprising:

obtaining, at a first device, audio directed to the first device from a second device during a communication session between the first device and the second device;

sending, from the first device, the audio to a transcription system;

obtaining, at the first device, a transcription during the communication session from the transcription system based on the audio;

obtaining a level of user understanding of the transcription, the level of the user understanding of the transcription being determined based on behavior of the user;

in response to the level of user understanding satisfying a threshold, obtaining, at the first device, a summary of the transcription during the communication session; and

presenting, on a display, both the summary and the transcription simultaneously during the communication session.

2. The method of claim 1 , wherein obtaining the summary of the transcription includes the first device generating the summary during the communication session using the transcription of the communication session.

3. The method of claim 1 , wherein the summary and the transcription are presented on the display in a manner such that a portion of the transcription and a portion of the summary that is derived from the portion of the transcription are visually associated together.

4. The method of claim 1 , wherein the presentation includes scrolling both the transcription and the summary to move along the display as the communication session proceeds.

5. The method of claim 4 , wherein the presentation includes scrolling the summary at a first scroll rate and scrolling the transcription at a second scroll rate that is different than the first scroll rate.

6. The method of claim 5 , wherein the first scroll rate is slower than the second scroll rate.

7. The method of claim 1 , further comprising ceasing to present the summary in response to an indication of an occurrence of an event associated with the communication session.

8. The method of claim 1 , wherein the behavior of the user used to determine the level of the user understanding of the transcription includes one or more of: facial expressions of the user, audio levels of the user, words spoken by the user, and the user reading the transcription.

9. A system comprising:

a display;

a processor coupled to the display and configured to direct data to be presented on the display; and

at least one non-transitory computer-readable media communicatively coupled to the processor and configured to store one or more instructions that when executed by the processor cause or direct the system to perform operations comprising:

obtain audio directed to the system from a device during a communication session between the system and the device;

obtain a transcription during the communication session based on the audio directed to the system in the communication session;

obtain a level of user understanding of the transcription, the level of the user understanding of the transcription being determined based on behavior of the user;

in response to the level of user understanding satisfying a threshold, obtain a summary of the transcription during the communication session; and

direct presentation on the display of both the summary and the transcription simultaneously during the communication session.

10. The system of claim 9 , wherein the operation to obtain the summary of the transcription includes the device generating the summary during the communication session using the transcription of the communication session.

11. The system of claim 9 , wherein the summary and the transcription are directed to present on the display in a manner such that a portion of the transcription and a portion of the summary that is derived from the portion of the transcription are visually associated together.

12. The system of claim 9 , wherein the operation to direct the presentation includes causing scrolling of both the transcription and the summary such that both the transcription and the summary move along the display as the communication session proceeds.

13. The system of claim 12 , wherein the operation to direct the presentation includes causing scrolling of the summary at a first scroll rate and scrolling of the transcription at a second scroll rate that is different than the first scroll rate.

14. The system of claim 13 , wherein the first scroll rate is slower than the second scroll rate.

15. The system of claim 9 , wherein the operations further comprise cease directing presentation of the summary in response to an indication of an occurrence of another event associated with the communication session.

16. The system of claim 9 , wherein the behavior of the user used to determine the level of the user understanding of the transcription includes one or more of: facial expressions of the user, audio levels of the user, words spoken by the user, and the user reading the transcription.

17. A system comprising:

a processor; and

at least one non-transitory computer-readable media communicatively coupled to the processor and configured to store one or more instructions that when executed by the processor cause or direct the system to perform operations comprising:

obtain audio directed to a first device from a second device during a communication session between the first device and the second device;

obtain a transcription during the communication session based on the audio of the communication session;

provide the transcription to the first device for presentation of the transcription;

obtain a level of user understanding of the transcription, the level of the user understanding of the transcription being determined based on behavior of the user; and

in response to the level of user understanding satisfying a threshold, provide a summary of the transcription to the first device for presentation of both the summary and the transcription simultaneously during the communication session.

18. The system of claim 17 , wherein the operations further comprise cease providing the summary in response to an indication of an occurrence of another event associated with the communication session.

19. The system of claim 17 , wherein the behavior of the user used to determine the level of the user understanding of the transcription includes one or more of: facial expressions of the user, audio levels of the user, words spoken by the user, and the user reading the transcription.

20. The system of claim 17 , wherein the presentation includes scrolling the summary at a first scroll rate and scrolling the transcription at a second scroll rate that is different than the first scroll rate.

Assignments (6)
CORRECTIVE ASSIGNMENT TO CORRECT THE RECEIVING PARTY DATA THE NAME OF THE LAST RECEIVING PARTY SHOULD BE CAPTIONCALL, LLC PREVIOUSLY RECORDED ON REEL 67190 FRAME 517. ASSIGNOR(S) HEREBY CONFIRMS THE RELEASE OF SECURITY INTEREST. Recorded May 31, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONCALL, LLC
Reel/Frame 067591/0675 →
RELEASE OF SECURITY INTEREST Recorded Apr 23, 2024
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH, AS COLLATERAL AGENT
To: SORENSON IP HOLDINGS, LLC; SORENSON COMMUNICATIONS, LLC; CAPTIONALCALL, LLC
Reel/Frame 067190/0517 →
SECURITY INTEREST Recorded Apr 23, 2024
From: SORENSON COMMUNICATIONS, LLC; INTERACTIVECARE, LLC; CAPTIONCALL, LLC
To: OAKTREE FUND ADMINISTRATION, LLC, AS COLLATERAL AGENT
Reel/Frame 067573/0201 →
JOINDER NO. 1 TO THE FIRST LIEN PATENT SECURITY AGREEMENT Recorded Apr 22, 2021
From: SORENSON IP HOLDINGS, LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056019/0204 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 21, 2019
From: CAPTIONCALL, LLC
To: SORENSON IP HOLDINGS, LLC
Reel/Frame 049237/0765 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 21, 2019
From: THOMSON, DAVID; BOEKWEG, SCOTT
To: CAPTIONCALL, LLC
Reel/Frame 049237/0835 →
Cited By (4)
US 12,374,337 US 12,400,660 US 12,482,458 US 12,488,799