IP Library Granted Patent US 11,574,638
Granted Patent B2
US 11,574,638 · App. 17/739,868 · Granted Feb 7, 2023

Automated audio-to-text transcription in multi-device teleconferences

Inventors: Tomas Gorny (Scottsdale, AZ); Jean-Baptiste Martinoli (St Anaclet de Lesard, CA); Tracy Conrad (Scottsdale, AZ); Lukas Gorny (Scottsdale, AZ)
Assignee: Nextiva, Inc.
G10L15/26G10L17/00H04M3/568
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,574,638
App. No.
17/739,868
Granted
Feb 7, 2023
Kind
B2
Abstract

A system and method are disclosed for generating a teleconference space for two or more communication devices using a computer coupled with a database and comprising a processor and memory. The computer generates a teleconference space and transmits requests to join the teleconference space to the two or more communication devices. The computer stores in memory identification information, and audiovisual data associated with one or more users, for each of the two or more communication devices. The computer stores audio transcription data, transmitted to the computer by each of the two or more communication devices and associated with one or more communication device users, in the computer memory. The computer merges the audio transcription data from each of the two or more communication devices into a master audio transcript, and transmits the master audio transcript to each of the two or more communication devices.

Claims (42)

1. A system, comprising:

two or more communication devices, each of the two or more communication devices configured to store recorded local device audio data, receive inbound audio and visual data from at least one other communication device, generate a local device transcript from the recorded local device audio data and generate a selectable display of a teleconference view or a transcript view; and

a computer coupled with a database and comprising a processor and memory, the computer configured to:

generate a teleconference space;

identify two or more users based on corresponding communication device data gathered from the two or more communication devices participating in the teleconference space;

merge the local device transcript from each of the two or more communication devices into a master audio transcript;

transmit the master audio transcript to each of the two or more communication devices; and

display, in response to a selection of the selectable display, the transcript view of the master audio transcript from each of the two or more telecommunication devices or the teleconference view of the visual data.

2. The system of claim 1 , wherein the computer is further configured to:

display in the transcript view a GUI comprising a transcript column and a participant panel.

3. The system of claim 1 , wherein the computer is further configured to:

display in the teleconference view a GUI comprising a participant panel displaying a visual representation of the two or more communication devices participating in the teleconference space and a teleconference window displaying video data associated with a user that is currently speaking.

4. The system of claim 1 , wherein the communication device data further comprises identification information of the two or more users associated with each of the two or more communication devices, the identification information comprising names and addresses, company contact information, telephone numbers, email addresses or IP addresses.

5. The system of claim 1 , wherein the communication device data further comprises end user identification information, communication device identification information, or communication device MAC address information.

6. The system of claim 1 , wherein the computer is configured to update the master audio transcript once every second or once every five seconds.

7. The system of claim 1 , wherein the computer is configured to store each local device transcript separately in local device transcript data of a cloud system database.

8. A computer-implemented method, comprising:

configuring each of two or more communication devices to store recorded local device audio data, receive inbound audio and visual data from at least one other communication device, generate a local device transcript from the recorded local device audio data and generate a selectable display of a transcript view or a teleconference view;

generating, using a computer coupled with a database and comprising a processor and memory, a teleconference space in which the two or more communication devices participate;

identifying, by the computer, two or more users based on corresponding communication device data gathered from the two or more communication devices participating in the teleconference space;

merging the local device transcript from each of the two or more communication devices into a master audio transcript;

transmitting the master audio transcript to each of the two or more communication devices; and

displaying, in response to a selection of the selectable display, the transcript view of the master audio transcript from each of the two or more telecommunication devices or the teleconference view of the visual data.

9. The computer-implemented method of claim 8 , further comprising displaying in the transcript view a GUI comprising a transcript column and a participant panel.

10. The computer-implemented method of claim 8 , further comprising displaying in the teleconference view a GUI comprising a participant panel displaying a visual representation of the two or more communication devices participating in the teleconference space and a teleconference window displaying video data associated with a user that is currently speaking.

11. The computer-implemented method of claim 8 , wherein the communication device data further comprises identification information of the two or more users associated with each of the two or more communication devices, the identification information comprising names and addresses, company contact information, telephone numbers, email addresses or IP addresses.

12. The computer-implemented method of claim 8 , wherein the communication device data further comprises end user identification information, communication device identification information, or communication device MAC address information.

13. The computer-implemented method of claim 8 , wherein the computer is configured to update the master audio transcript once every second or once every five seconds.

14. The computer-implemented method of claim 8 , wherein the computer is configured to store each local device transcript separately in local device transcript data of a cloud system database.

15. A non-transitory computer-readable storage medium embodied with software, the software when executed:

configures each of two or more communication devices to store recorded local device audio data, receive inbound audio and visual data from at least one other communication device, generate a local device transcript from the recorded local device audio data and generate a selectable display of a transcript view or a teleconference view;

generates, using a computer coupled with a database and comprising a processor and memory, a teleconference space in which the two or more communication devices participate;

identifies, by the computer, two or more users based on corresponding communication device data gathered from the two or more communication devices participating in the teleconference space;

merges the local device transcript from each of the two or more communication devices into a master audio transcript;

transmits the master audio transcript to each of the two or more communication devices; and

displays, in response to a selection of the selectable display, the transcript view of the master audio transcript from each of the two or more telecommunication devices or the teleconference view of the visual data.

16. The non-transitory computer-readable storage medium of claim 15 , wherein the software when executed further:

displays in the transcript view a GUI comprising a transcript column and a participant panel.

17. The non-transitory computer-readable storage medium of claim 15 , wherein the software when executed further: displays in the teleconference view a GUI comprising a participant panel displaying a visual representation of the two or more communication devices participating in the teleconference space and a teleconference window displaying video data associated with a user that is currently speaking.

18. The non-transitory computer-readable storage medium of claim 15 , wherein the communication device data further comprises identification information of the two or more users associated with each of the two or more communication devices, the identification information comprising names and addresses, company contact information, telephone numbers, email addresses or IP addresses.

19. The non-transitory computer-readable storage medium of claim 15 , wherein the communication device data further comprises end user identification information, communication device identification information, or communication device MAC address information.

20. The non-transitory computer-readable storage medium of claim 19 , wherein the software when executed is further configured to store each local device transcript separately in local device transcript data of a cloud system database.

Assignments (3)
CORRECTIVE ASSIGNMENT TO CORRECT THE CORRECT THE PROPERTY NUMBERS PREVIOUSLY RECORDED AT REEL: 67172 FRAME: 404. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Apr 23, 2024
From: NEXTIVA, INC.; THRIO, INC.
To: WELLS FARGO BANK, NATIONAL ASSOCIATION, AS ADMINISTRATIVE AGENT
Reel/Frame 067308/0183 →
SECURITY INTEREST Recorded Apr 19, 2024
From: NEXTIVA, INC.; THRIO, INC.
To: WELLS FARGO BANK, NATIONAL ASSOCIATION, AS ADMINISTRATIVE AGENT
Reel/Frame 067172/0404 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 2, 2022
From: GORNY, TOMAS; MARTINOLI, JEAN-BAPTISTE; CONRAD, TRACY; GORNY, LUKAS
To: NEXTIVA, INC.
Reel/Frame 062052/0928 →
Continuity (3)
Continuation 16861929 · Apr 29, 2020
Provisional Application 62876401 · Jul 19, 2019
Related Publication 20220262366A1 · Aug 18, 2022