IP Library Granted Patent US 12,159,643
Granted Patent B2
US 12,159,643 · App. 18/376,716 · Granted Dec 3, 2024

Systems and methods for filtering unwanted sounds from a conference call using voice synthesis

Inventors: Rajendran Pichaimurthy (Karnataka, IN); Madhusudhan Seetharam (Karnataka, IN)
Assignee: Adeia Guides Inc.
G10L21/0208G10L15/26G10L21/0272G10L25/84H04M3/42068H04M3/568G10L2021/02087H04M2203/5072
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,159,643
App. No.
18/376,716
Granted
Dec 3, 2024
Kind
B2
Abstract

To filter unwanted sounds from a conference call, a first voice signal is captured by a first device during a conference call and converted into corresponding text, which is then analyzed to determine that a first portion of the text was spoken by a first user and a second portion of the text was spoken by a second user. If the first user is relevant to the conference call while the second user is not, the first voice signal is prevented from being transmitted into the conference call, the first portion of text is converted into a second voice signal using a voice profile of the first user to synthesize the voice of the first user, and the second voice signal is then transmitted into the conference call. The second portion of text is not converted into a voice signal, as the second user is determined not to be relevant.

Claims (54)

1. A method comprising:

capturing a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a communication session;

in response to determining, using a first voice profile of the first user, that the first voice signal includes the voice of the second user, wherein the second user is different from the first user:

preventing entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the communication session, wherein no part of the first voice signal is transmitted into the communication session;

constructing a second voice signal based on words detected in the first voice signal and attributable to the voice of the first user; and

transmitting the second voice signal into the communication session.

2. The method of claim 1 , further comprising:

generating the first voice profile of the first user based on a prior voice signal captured during a prior communication session.

3. The method of claim 1 , wherein the constructing the second voice signal based on the words detected in the first voice signal and attributable to the voice of the first user comprises:

accessing the first voice profile of the first user; and

synthesizing the words detected in the first voice signal and attributable to the voice of the first user into the second voice signal based on the first voice profile of the first user.

4. The method of claim 1 , further comprising:

generating a third voice signal based on words detected in the first voice signal and attributable to the voice of the second user; and

transmitting the third voice signal into the communication session separately from the second voice signal.

5. The method of claim 4 , further comprising:

presenting, to other participants in the communication session, an option to select whether to listen to the second voice signal or the third voice signal.

6. The method of claim 1 , further comprising muting a microphone for a predetermined period of time in response to detecting greater than one voice in the first voice signal.

7. The method of claim 1 , wherein the constructed second voice signal excludes the words detected in the first voice signal and attributable to the voice of the second user.

8. A system comprising:

input/output circuitry;

audio input circuitry configured to capture a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a communication session; and

control circuitry configured to:

in response to determining, using a first voice profile of the first user, that the first voice signal includes the voice of the second user, wherein the second user is different from the first user:

prevent entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the communication session, wherein no part of the first voice signal is transmitted into the communication session;

construct a second voice signal based on words detected in the first voice signal and attributable to the voice of the first user; and

wherein the input/output circuitry is configured to transmit the second voice signal into the communication session.

9. The system of claim 8 , wherein the control circuitry is further configured to:

generate the first voice profile of the first user based on a prior voice signal captured during a prior communication session.

10. The system of claim 8 , wherein the control circuitry is configured to construct the second voice signal based on the words detected in the first voice signal and attributable to the voice of the first user by:

accessing the first voice profile of the first user; and

synthesizing the words detected in the first voice signal and attributable to the voice of the first user into the second voice signal based on the first voice profile of the first user.

11. The system of claim 8 , wherein the control circuitry is further configured to generate a third voice signal based on words detected in the first voice signal and attributable to the voice of the second user; and

wherein the input/output circuitry is further configured to transmit the third voice signal into the communication session separately from the second voice signal.

12. The system of claim 11 , wherein the control circuitry is further configured to:

present, to other participants in the communication session, an option to select whether to listen to the second voice signal or the third voice signal.

13. The system of claim 8 , wherein the control circuitry is further configured to mute a microphone for a predetermined period of time in response to detecting greater than one voice in the first voice signal.

14. The system of claim 8 , wherein the constructed second voice signal excludes the words detected in the first voice signal and attributable to the voice of the second user.

15. A system comprising:

means for capturing a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a communication session;

means for, in response to determining, using a first voice profile of the first user, that the first voice signal includes the voice of the second user, wherein the second user is different from the first user:

preventing entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the communication session, wherein no part of the first voice signal is transmitted into the communication session;

constructing a second voice signal based on words detected in the first voice signal and attributable to the voice of the first user; and

means for transmitting the second voice signal into the communication session.

16. The system of claim 15 , further comprising:

means for generating the first voice profile of the first user based on a prior voice signal captured during a prior communication session.

17. The system of claim 15 , wherein the means for constructing the second voice signal based on the words detected in the first voice signal and attributable to the voice of the first user comprises:

means for accessing the first voice profile of the first user; and

means for synthesizing the words detected in the first voice signal and attributable to the voice of the first user into the second voice signal based on the first voice profile of the first user.

18. The system of claim 15 , further comprising:

means for generating a third voice signal based on the words detected in the first voice signal and attributable to the voice of the second user; and

means for transmitting the third voice signal into the communication session separately from the second voice signal.

19. The system of claim 18 , further comprising:

means for presenting, to other participants in the communication session, an option to select whether to listen to the second voice signal or the third voice signal.

20. The system of claim 15 , further comprising means for muting a microphone for a predetermined period of time in response to detecting greater than one voice in the first voice signal.

Assignments (2)
CHANGE OF NAME Recorded Oct 4, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069113/0392 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 15, 2023
From: PICHAIMURTHY, RAJENDRAN; SEETHARAM, MADHUSUDHAN
To: ROVI GUIDES, INC.
Reel/Frame 065565/0181 →
Continuity (3)
Continuation 17884851 · Aug 10, 2022
Continuation 17015832 · Sep 9, 2020
Related Publication 20240046943A1 · Feb 8, 2024