IP Library Granted Patent US 12,388,930
Granted Patent B2
US 12,388,930 · App. 17/977,725 · Granted Aug 12, 2025

Filtering sensitive topic speech within a conference audio stream

Inventor: Nick Swerdlow (Santa Clara, CA)
Assignee: Zoom Communications, Inc.
H04M3/568G10L15/08G10L25/51
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,388,930
App. No.
17/977,725
Granted
Aug 12, 2025
Kind
B2
Abstract

A portion of an audio stream representing speech of a conference participant is filtered based on a determination that the portion of the audio stream corresponds to a predefined sensitive topic. An audio stream representing speech of a user of a participant device connected to a conference is obtained. First hash values are determined for portions of the speech. A determination is made, by comparing the first hash values against second hash values for records of a data store, that a portion of the speech corresponds to a predefined sensitive topic indicated within the records. A filter is applied against the portion of the speech to produce a modified audio stream within which the predefined sensitive topic is sanitized. An output, within the conference, of the modified audio stream is then caused in place of the audio stream.

Claims (41)

1. A method, comprising:

obtaining an audio stream representing speech of a user of a participant device connected to a conference;

determining first hash values for portions of the speech;

determining, by comparing the first hash values against second hash values for records of a data store, that a portion of the speech corresponds to a predefined sensitive topic indicated within the records;

applying a filter against the portion of the speech to produce a modified audio stream within which the predefined sensitive topic is sanitized; and

causing an output, within the conference, of the modified audio stream in place of the audio stream.

2. The method of claim 1 , wherein portions of the speech other than the portion of the speech remain unfiltered within the modified audio stream.

3. The method of claim 1 , wherein the records of the data store are populated by an entity having a domain with which a conference user account used at the participant device is associated.

4. The method of claim 1 , wherein each record of the data store indicates a different predefined sensitive topic, and wherein a record of the data store is removed from the data store upon a public disclosure of an associated predefined sensitive topic.

5. The method of claim 1 , wherein the predefined sensitive topic relates to confidential information.

6. The method of claim 1 , wherein the application of the filter causes an omission of the portion of the speech from within the modified audio stream.

7. The method of claim 1 , wherein the application of the filter causes an obfuscation of the portion of the speech from within the modified audio stream.

8. The method of claim 1 , wherein the application of the filter causes a replacement of the portion of the speech from within the modified audio stream with other audible content.

9. A non-transitory computer readable medium storing instructions operable to cause one or more processors to perform operations comprising:

obtaining an audio stream representing speech of a user of a participant device connected to a conference;

determining first hash values for portions of the speech;

determining, by comparing the first hash values against second hash values for records of a data store, that a portion of the speech corresponds to a predefined sensitive topic indicated within the records;

applying a filter against the portion of the speech to produce a modified audio stream within which the predefined sensitive topic is sanitized; and

causing an output, within the conference, of the modified audio stream in place of the audio stream.

10. The non-transitory computer readable medium of claim 9 , wherein the operations for causing the output of the modified audio stream in place of the audio stream comprise:

causing the modified audio stream to be output at one or more participant devices connected to the conference based on conference user accounts used at the one or more participant devices corresponding other than to a domain associated with a conference user account used at the participant device from which the audio stream is obtained.

11. The non-transitory computer readable medium of claim 9 , wherein the operations for causing the output of the modified audio stream in place of the audio stream comprise:

transmitting the modified audio stream to one or more participant devices connected to the conference to cause the modified audio stream to be output at the one or more participant devices.

12. The non-transitory computer readable medium of claim 9 , wherein the predefined sensitive topic corresponds to a codename.

13. The non-transitory computer readable medium of claim 9 , wherein the conference is implemented by a unified communications as a service software platform.

14. An apparatus, comprising:

a memory; and

a processor configured to execute instructions stored in the memory to:

obtain an audio stream representing speech of a user of a participant device connected to a conference;

determine first hash values for portions of the speech;

determine, by comparing the first hash values against second hash values for records of a data store, that a portion of the speech corresponds to a predefined sensitive topic indicated within the records;

apply a filter against the portion of the speech to produce a modified audio stream within which the predefined sensitive topic is sanitized; and

cause an output, within the conference, of the modified audio stream in place of the audio stream.

15. The apparatus of claim 14 , wherein the processor is configured to execute the instructions to:

determine to compare the first hash values against the second hash values based on a mismatch between a first domain associated with a first conference user account used at the participant device and a second domain associated with a second conference user account used at another participant device connected to the conference.

16. The apparatus of claim 14 , wherein the processor is configured to execute the instructions to:

determine to apply the filter against the portion of the speech based on a mismatch between a first domain associated with a first conference user account used at the participant device and a second domain associated with a second conference user account used at another participant device connected to the conference.

17. The apparatus of claim 14 , wherein the modified audio stream is output to a first subset of participant devices connected to the conference and the audio stream is output to a second subset of the participant devices, wherein conference user accounts used at the first subset of the participant devices are associated with a first domain different from a second domain with which a conference user account used at the participant device is associated with, and wherein conference user accounts used at the second subset of the participant devices are associated with the second domain.

18. The apparatus of claim 14 , wherein a control policy restricts users of participant devices connected to the conference from altering the records of the data store.

19. The apparatus of claim 14 , wherein the modified audio stream is produced and output while the conference remains ongoing.

20. The apparatus of claim 14 , wherein the modified audio stream is produced and output during playback of a recording of the conference.

Assignments (2)
CHANGE OF NAME Recorded Jan 7, 2025
From: ZOOM VIDEO COMMUNICATIONS, INC.
To: ZOOM COMMUNICATIONS, INC.
Reel/Frame 069839/0593 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 31, 2022
From: SWERDLOW, NICK
To: ZOOM VIDEO COMMUNICATIONS, INC.
Reel/Frame 061599/0981 →
Continuity (1)
Related Publication 20240146846A1 · May 2, 2024
References Cited (23)
US 8121845B2 · Kirby · 2012 [cited by applicant]
US 8537978B2 · Jaiswal et al. · 2013 [cited by applicant]
US 9413891B2 · Dwyer et al. · 2016 [cited by applicant]
US 9443518B1 · Gauci · 2016 [cited by examiner]
US 10432687B1 · Hanes et al. · 2019 [cited by applicant]
US 10764534B1 · Shevchenko et al. · 2020 [cited by applicant]
US 11450334B2 · Pichaimurthy et al. · 2022 [cited by applicant]
US 11563855B1 · Spivak et al. · 2023 [cited by applicant]
US 20040263636A1 · Cutler et al. · 2004 [cited by applicant]
US 20070230372A1 · He · 2007 [cited by examiner]
US 20130139259A1 · Tegreene · 2013 [cited by applicant]
US 20130329866A1 · Mai et al. · 2013 [cited by applicant]
US 20140028784A1 · Deyerle et al. · 2014 [cited by applicant]
US 20150012270A1 · Reynolds · 2015 [cited by applicant]
US 20150149173A1 · Korycki · 2015 [cited by applicant]
US 20160063097A1 · Brown et al. · 2016 [cited by applicant]
US 20220051652A1 · Winsvold et al. · 2022 [cited by applicant]
US 20220199102A1 · Ostrand et al. · 2022 [cited by applicant]
US 20230013497A1 · Aher et al. · 2023 [cited by applicant]
US 20230117129A1 · Mouline et al. · 2023 [cited by applicant]
MyFone, 6 Popular Real-Time Voice Changers for Zoom [2022 List], Karen William, Sep. 10, 2021 (Updated Jul. 5, 2022), 8 pages. [cited by applicant]
Voicemod, Voice Changer for Video Calls: ZOOM, Hangouts, Facetime, Sep. 2022, 2 pages. [cited by applicant]
Accent Conversion using Pre-trained Model and Synthesized Data from Voice Conversion, Tuan Nam Nguyen, Ngoc Quan Pham, Alexander Waibel, Karlsruhe Institute of Technology and Carnegie Mellon University, Sep. 2022, 5 pag… [cited by applicant]