IP Library Granted Patent US 10,084,921
Granted Patent B2
US 10,084,921 · App. 15/653,324 · Granted Sep 25, 2018

Handling concurrent speech

Inventors: Serge Lachapelle (Vallentuna, SE); Alexander Kjeldaas (Saltsjö-Boo, SE)
Assignee: GOOGLE LLC
H04M3/568A61B17/3203A61B17/3207A61B17/3211A61B17/320725A61B17/50A61B18/245A61B90/02A61N1/056G10L21/00H04L65/403A61B2017/320044
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,084,921
App. No.
15/653,324
Granted
Sep 25, 2018
Kind
B2
Abstract

Systems and methods are provided for handling concurrent speech in which temporally overlapping first speech data and second speech data is received from respective first and second participants of a session. A speech policy applied to the speech data specifies dropping the second speech when it interrupts the first speech within a first interval of the first speech data. The first interval is temporally bounded by the beginning of the first speech and a first predetermined amount of time after the beginning of the first speech. The speech policy specifies outputting the first speech data and then outputting the second speech data when the second speech data interrupts a second interval of the first speech data. The second interval of the first speech data is temporally bounded by the end of the first speech data and a second predetermined amount of time before the end of the first speech data.

Claims (45)

1. A method comprising:

at a system comprising one or more processors and a memory storing one or more programs for execution by the one or more processors:

receiving first speech data from a first participant of a session;

receiving second speech data from a second participant of the session, wherein the second speech data temporally overlaps at least a portion of the first speech data; and

applying a speech policy to the second speech data, wherein

the speech policy specifies dropping the second speech data when the second speech data interrupts the first speech data within a first interval of the first speech data, wherein the first interval of the first speech data is temporally bounded by the beginning of the first speech data and a first predetermined amount of time after the beginning of the first speech data, and

the speech policy specifies outputting the first speech data and then outputting the second speech data when the second speech data interrupts a second portion of the first speech data, wherein the second interval of the first speech data is other than the first interval of the first speech data and is temporally bounded by the end of the first speech data and a second predetermined amount of time prior to the end of the first speech data.

2. The method of claim 1 , wherein the speech policy further specifies dropping the second speech data when a third interval of time between the beginning of the second speech and the end of the first speech data is greater than a predetermined amount of time.

3. The method of claim 1 , wherein the second participant is classified as a low priority speaker and the first participant is classified as a main speaker.

4. The method of claim 1 , wherein

the first participant is classified as a main speaker based upon a first social network status associated with the first participant, and

the second participant is classified as a low priority speaker based upon a second social network status associated with the second participant.

5. The method of claim 1 wherein the first speech and the second speech is outputted to a plurality of client devices.

6. The method of claim 1 , wherein the session comprises three or more participants and the first speech and the second speech is outputted to a user device uniquely associated with each participant in the three or more participants.

7. The method of claim 1 , wherein the second speech includes a pause and the speech policy further comprises removing the pause from the second speech when outputting the second speech.

8. The method of claim 1 , wherein the second speech includes a pause and the speech policy further comprises reducing a duration of the pause in the second speech when outputting the second speech.

9. A server system, comprising:

one or more processors;

memory; and

one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by the one or more processors, the one or more programs including instructions for:

receiving first speech data from a first participant of a session;

receiving second speech data from a second participant of the session, wherein the second speech data temporally overlaps at least a portion of the first speech data; and

applying a speech policy to the second speech data, wherein

the speech policy specifies dropping the second speech data when the second speech data interrupts the first speech data within a first interval of the first speech data, wherein the first interval of the first speech data is temporally bounded by the beginning of the first speech data and a first predetermined amount of time after the beginning of the first speech data, and

the speech policy specifies outputting the first speech data and then outputting the second speech data when the second speech data interrupts a second portion of the first speech data, wherein the second interval of the first speech data is other than the first interval of the first speech data and is temporally bounded by the end of the first speech data and a second predetermined amount of time prior to the end of the first speech data.

10. The server system of claim 9 , wherein the speech policy further specifies dropping the second speech data when a third interval of time between the beginning of the second speech and the end of the first speech data is greater than a predetermined amount of time.

11. The server system of claim 9 , wherein the second participant is classified as a low priority speaker and the first participant is classified as a main speaker.

12. The server system of claim 9 , wherein

the first participant is classified as a main speaker based upon a first social network status associated with the first participant, and

the second participant is classified as a low priority speaker based upon a second social network status associated with the second participant.

13. The server system of claim 9 , wherein the first speech and the second speech is outputted to a plurality of client devices.

14. The server system of claim 9 , wherein the session comprises three or more participants and the first speech and the second speech is outputted to a user device uniquely associated with each participant in the three or more participants.

15. The server system of claim 9 , wherein the second speech includes a pause and the speech policy further comprises removing the pause from the second speech when outputting the second speech.

16. A non-transitory computer readable storage medium storing one or more programs, the one or more programs comprising instructions, which when executed by a computer system with one or more processors, cause the computer system to:

receive first speech data from a first participant of a session;

receive second speech data from a second participant of the session, wherein the second speech data temporally overlaps at least a portion of the first speech data; and

apply a speech policy to the second speech data, wherein

the speech policy specifies dropping the second speech data when the second speech data interrupts the first speech data within a first interval of the first speech data, wherein the first interval of the first speech data is temporally bounded by the beginning of the first speech data and a first predetermined amount of time after the beginning of the first speech data, and

the speech policy specifies outputting the first speech data and then outputting the second speech data when the second speech data interrupts a second portion of the first speech data, wherein the second interval of the first speech data is other than the first interval of the first speech data and is temporally bounded by the end of the first speech data and a second predetermined amount of time prior to the end of the first speech data.

17. The non-transitory computer readable storage medium of claim 16 , wherein the speech policy further specifies dropping the second speech data when a third interval of time between the beginning of the second speech and the end of the first speech data is greater than a predetermined amount of time.

18. The non-transitory computer readable storage medium of claim 16 , wherein the second participant is classified as a low priority speaker and the first participant is classified as a main speaker.

19. The non-transitory computer readable storage medium of claim 16 , wherein

the first participant is classified as a main speaker based upon a first social network status associated with the first participant, and

the second participant is classified as a low priority speaker based upon a second social network status associated with the second participant.

20. The non-transitory computer readable storage medium of claim 16 , wherein the first speech and the second speech is outputted to a plurality of client devices.

Assignments (1)
CHANGE OF NAME Recorded Oct 20, 2017
From: GOOGLE INC.
To: GOOGLE LLC
Reel/Frame 044567/0001 →
Continuity (5)
Continuation 15336629 · Oct 27, 2016
Continuation 15059222 · Mar 2, 2016
Continuation 14027061 · Sep 13, 2013
Provisional Application 61701520 · Sep 14, 2012
Related Publication 20170318158A1 · Nov 2, 2017