IP Library Granted Patent US 11,978,472
Granted Patent B2
US 11,978,472 · App. 17/210,108 · Granted May 7, 2024

Systems and methods for processing and presenting conversations

Inventors: Yun Fu (Cupertino, CA); Simon Lau (San Jose, CA); Kaisuke Nakajima (Sunnyvale, CA); Julius Cheng (Cupertino, CA); Gelei Chen (Mountain View, CA); Sam Song Liang (Palo Alto, CA); James Mason Altreuter (Belmont, CA); Kean Kheong Chin (Santa Clara, CA); Zhenhao Ge (Sunnyvale, CA); Hitesh Anand Gupta (Santa Clara, CA); Xiaoke Huang (Foster City, CA); James Francis McAteer (San Francisco, CA); Brian Francis Williams (San Carlos, CA); Tao Xing (San Jose, CA)
Assignee: Otter.ai, Inc.
G10L21/10G06F16/438G10L17/02G10L17/04G10L17/22H04L63/104
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,978,472
App. No.
17/210,108
Granted
May 7, 2024
Kind
B2
Abstract

A system for processing and presenting a conversation includes a sensor, a processor, and a presenter. The sensor is configured to capture an audio-form conversation. The processor is configured to automatically transform the audio-form conversation into a transformed conversation. The transformed conversation includes a synchronized text, wherein the synchronized text is synchronized with the audio-form conversation. The presenter is configured to present the transformed conversation including the synchronized text and the audio-form conversation. The presenter is further configured to present the transformed conversation to be navigable, searchable, assignable, editable, and shareable.

Claims (27)

1. A computer-implemented method for processing and presenting a conversation, the method comprising:

receiving an audio-form conversation;

processing the audio-form conversation;

automatically generating a group having one or more group members based at least in part on one of a date of capture, a time of the capture, a title of the capture, and a speaker identity associated with the processed audio-form conversation;

automatically assigning the processed audio-form conversation to the group such that access to the processed audio-form conversation is granted to the one or more group members; and

presenting the processed audio-form conversation to the group members.

2. The method of claim 1 , wherein the processing the audio-form conversation includes segmenting the audio-form conversation based on one or more speakers change detections.

3. The method of claim 2 , further comprising automatically generating one or more segments of the audio-form conversation when a speaker change occurs or a natural pause occurs such that each segment of the one or more segments of the audio-form conversation is spoken by only one speaker.

4. The method of claim 3 , further comprising automatically assigning only one speaker label to each segment of the processed audio-form conversation, each one speaker label representing one speaker.

5. The method of claim 1 , wherein the automatically generating a group having one or more group members comprises automatically generating a group by identifying one or more group members who are in a conversation associated with a calendar event of a synced calendar.

6. The method of claim 1 , wherein the automatically generating a group having one or more group members comprises automatically generating a group by identifying one or more group members based on contacts or an address book of the one or more group members.

7. The method of claim 1 , further comprising automatically assigning one or more labels to the processed audio-form conversation,

wherein the presenting the processed audio-form conversation includes presenting the processed audio-form conversation with the one or more assigned labels.

8. A system for processing and presenting a conversation, the system comprising:

a sensor configured to receive an audio-form conversation;

a processor configured to:

process the audio-form conversation;

automatically generate a group having one or more group members based at least in part on one of a date of capture, a time of the capture, a title of the capture, and a speaker identity associated with the processed audio-form conversation;

automatically assign the processed audio-form conversation to the group such that access to the processed audio-form conversation is granted to the one or more group members; and

a presenter configured to present the processed audio-form conversation to the group members.

9. The system of claim 8 , wherein the processing the audio-form conversation includes segmenting the audio-form conversation based on one or more speakers change detections.

10. The system of claim 9 , wherein the processor is further configured to automatically generate one or more segments of the audio-form conversation when a speaker change occurs or a natural pause occurs such that each segment of the one or more segments of the audio-form conversation is spoken by only one speaker.

11. The system of claim 9 , wherein the processor is further configured to automatically assigning only one speaker label to each segment of the processed audio-form conversation, each one speaker label representing one speaker.

12. The system of claim 8 , wherein to automatically generate a group having one or more group members comprises to automatically generate a group by identifying one or more group members who are in a conversation associated with a calendar event of a synced calendar.

13. The system of claim 8 , wherein to automatically generate a group having one or more group members comprises to automatically generate a group by identifying one or more group members based on contacts or an address book of the one or more group members.

14. The system of claim 8 , wherein the processor is further configured to automatically assigning one or more labels to the processed audio-form conversation,

wherein to present the processed audio-form conversation includes to present the processed audio-form conversation with the one or more assigned labels.

Assignments (2)
CHANGE OF NAME Recorded Aug 11, 2021
From: AISENSE INC.
To: OTTER.AI, INC.
Reel/Frame 057159/0577 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 10, 2021
From: LAU, SIMON; FU, YUN; ALTREUTER, JAMES MASON; WILLIAMS, BRIAN FRANCIS; HUANG, XIAOKE; XING, TAO; NAKAJIMA, KAISUKE; CHIN, KEAN KHEONG; GUPTA, HITESH ANAND; CHENG, JULIUS; LIANG, SAM SONG; CHEN, GELEI; GE, ZHENHAO; MCATEER, JAMES FRANCIS
To: AISENSE, INC.
Reel/Frame 057140/0070 →
Continuity (7)
Continuation 16276446 · Feb 14, 2019
Continuation In Part 16027511 · Jul 5, 2018
Provisional Application 62668623 · May 8, 2018
Provisional Application 62631680 · Feb 17, 2018
Provisional Application 62710631 · Feb 16, 2018
Provisional Application 62530227 · Jul 9, 2017
Related Publication 20210327454A1 · Oct 21, 2021