IP Library Granted Patent US 12,348,327
Granted Patent B2
US 12,348,327 · App. 17/977,705 · Granted Jul 1, 2025

Participant audio stream modification within a conference

Inventor: Nick Swerdlow (Santa Clara, CA)
Assignee: Zoom Communications, Inc.
H04L12/1822H04L12/1818H04L12/1831
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,348,327
App. No.
17/977,705
Granted
Jul 1, 2025
Kind
B2
Abstract

The audio stream of a participant to a conference is modified within the conference to change a perceptible output of a characteristic of speech represented by the audio stream. An audio stream is obtained from a participant device connected to a conference. The audio stream represents speech of a user of the participant device. A user request to modify a first characteristic of the speech is initiated within the conference. The first characteristic is modified without modifying other characteristics of the speech to produce a modified audio stream, such that a second characteristic of the speech remains unmodified within the modified audio stream. An output of the modified audio stream within the conference is then caused in place of the audio stream. The audio stream modification as disclosed herein may be performed while the conference remains ongoing or during playback of a recording of the conference.

Claims (38)

1. A method, comprising:

obtaining an audio stream from a participant device connected to a conference, wherein the audio stream represents speech of a user of the participant device;

modifying, based on a user request initiated within the conference, a first characteristic of the speech to produce a modified audio stream, wherein a second characteristic of the speech remains unmodified within the modified audio stream, wherein the first characteristic corresponds to a cadence of the speech, and wherein modifying the first characteristic of the speech to produce the modified audio stream comprises removing one or more periods of silence between portions of the speech; and

causing an output, within the conference, of the modified audio stream in place of the audio stream.

2. The method of claim 1 , wherein a volume of the speech is also modified to produce the modified audio stream.

3. The method of claim 1 , comprising:

determining to modify the first characteristic based on a threshold comparison against measured values of the first characteristic; and

transmitting, based on the determination to modify the first characteristic, a prompt configured to initiate the user request.

4. The method of claim 1 , wherein the user request is obtained from the participant device.

5. The method of claim 1 , wherein the user request is obtained from a second participant device connected to the conference and used by a second user.

6. The method of claim 1 , wherein the modified audio stream is produced and output while the conference remains in-progress.

7. The method of claim 1 , wherein the modified audio stream is produced and output during playback of a recording of the conference.

8. One or more non-transitory computer readable media storing instructions operable to cause one or more processors to perform operations comprising:

obtaining an audio stream from a participant device connected to a conference, wherein the audio stream represents speech of a user of the participant device;

modifying, based on a user request initiated within the conference, a first characteristic of the speech to produce a modified audio stream, wherein a second characteristic of the speech remains unmodified within the modified audio stream, wherein the first characteristic corresponds a cadence of the speech, and wherein modifying the first characteristic of the speech to produce the modified audio stream comprises removing one or more periods of silence between portions of the speech; and

causing an output, within the conference, of the modified audio stream in place of the audio stream.

9. The one or more non-transitory computer readable media of claim 8 , wherein audio streams obtained from participant devices connected to the conference are stored in separate files associated with a recording of the conference, and wherein the operations to modify the first characteristic to produce the modified audio stream comprise:

altering one of the separate files corresponding to the audio stream.

10. The one or more non-transitory computer readable media of claim 8 , wherein the modifying is performed by software running at a server device implementing the conference.

11. The one or more non-transitory computer readable media of claim 8 , wherein the user request is initiated at the participant device and the modifying is performed by software running at the participant device.

12. The one or more non-transitory computer readable media of claim 8 , wherein the user request is initiated at a second participant device connected to the conference and the modifying is performed by software running at the second participant device.

13. A system, comprising:

one or more memories; and

one or more processors configured to execute instructions stored in the one or more memories to:

obtain an audio stream from a participant device connected to a conference, wherein the audio stream represents speech of a user of the participant device;

modify, based on a user request initiated within the conference, a first characteristic of the speech to produce a modified audio stream, wherein a second characteristic of the speech remains unmodified within the modified audio stream, wherein the first characteristic corresponds to a cadence of the speech, and wherein modifying the first characteristic of the speech to produce the modified audio stream comprises removing one or more periods of silence between portions of the speech; and

cause an output, within the conference, of the modified audio stream in place of the audio stream.

14. The system of claim 13 , wherein a pitch of the speech is also modified to produce the modified audio stream, and wherein the one or more processors are further configured to execute instructions stored in the one or more memories to:

determine that the pitch of the speech is outside of a threshold range for a threshold period of time; and

prompt, based on the determination that the pitch of the speech is outside of the threshold range for the threshold period of time, the user of the participant device to initiate the user request.

15. The system of claim 13 , wherein an accent of the speech is also modified to produce the modified audio stream, and wherein the one or more processors are further configured to execute instructions stored in the one or more memories to:

determine that the accent of the speech is different from an accent of speech of other users of participant devices connected to the conference; and

prompt, based on the determination that the accent of the speech is different from the accent of the speech of the other users, the user of the participant device to initiate the user request.

16. The system of claim 13 , wherein the one or more processors are further configured to execute instructions stored in the one or more memories to:

present, to the user of the participant device, a recommendation to apply a filter during one or more future conferences based on the modifying.

17. The system of claim 13 , wherein the output of the modified audio stream is to a first subset of participant devices connected to the conference, and wherein the audio stream is output without the modifying to a second subset of the participant devices connected to the conference.

18. The system of claim 13 , wherein the user request is initiated while the conference remains in-progress.

19. The system of claim 13 , wherein the user request is initiated during playback of a recording of the conference.

Assignments (2)
CHANGE OF NAME Recorded Jan 7, 2025
From: ZOOM VIDEO COMMUNICATIONS, INC.
To: ZOOM COMMUNICATIONS, INC.
Reel/Frame 069839/0593 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 31, 2022
From: SWERDLOW, NICK
To: ZOOM VIDEO COMMUNICATIONS, INC.
Reel/Frame 061599/0804 →