IP Library Granted Patent US 11,810,585
Granted Patent B2
US 11,810,585 · App. 17/884,851 · Granted Nov 7, 2023

Systems and methods for filtering unwanted sounds from a conference call using voice synthesis

Inventors: Rajendran Pichaimurthy (Karnataka, IN); Madhusudhan Seetharam (Karnataka, IN)
Assignee: Rovi Guides, Inc.
G10L21/0208G10L15/26G10L21/0272G10L25/84H04M3/42068H04M3/568G10L2021/02087H04M2203/5072
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,810,585
App. No.
17/884,851
Granted
Nov 7, 2023
Kind
B2
Abstract

To filter unwanted sounds from a conference call, a first voice signal is captured by a first device during a conference call and converted into corresponding text, which is then analyzed to determine that a first portion of the text was spoken by a first user and a second portion of the text was spoken by a second user. If the first user is relevant to the conference call while the second user is not, the first voice signal is prevented from being transmitted into the conference call, the first portion of text is converted into a second voice signal using a voice profile of the first user to synthesize the voice of the first user, and the second voice signal is then transmitted into the conference call. The second portion of text is not converted into a voice signal, as the second user is determined not to be relevant.

Claims (55)

1. A method comprising:

capturing a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a call;

preventing entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call;

converting the first voice signal into text;

generating a second voice signal based on portions of the text attributable to the voice of the first user; and

transmitting the second voice signal into the call.

2. The method of claim 1 , wherein the generated second voice signal excludes the portions of the text attributable to the voice of the second user.

3. The method of claim 1 , wherein the generating the second voice signal based on portions of the text attributable to the voice of the first user comprises:

accessing a voice profile of the first user; and

synthesizing the portions of the text attributable to the voice of the first user into the second voice signal based on the voice profile of the first user.

4. The method of claim 1 , further comprising:

generating a third voice signal based on portions of the text attributable to the voice of the second user;

transmitting the third voice signal into the call separately from the second voice signal.

5. The method of claim 4 , further comprising:

presenting, to other participants in the call, an option to select whether to listen to the second voice signal or the third voice signal.

6. The method of claim 1 , wherein the preventing the entirety of the first voice signal, the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call, further comprises:

preventing transmission into the call of all voices present in the first voice signal.

7. The method of claim 1 , further comprising muting the microphone for a predetermined period of time in response to detecting greater than one voice in the first voice signal.

8. A system comprising:

audio input circuitry configured to capture voice signals; and

control circuitry configured to:

capture, using the audio input circuitry, a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a call;

prevent the entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call;

convert the first voice signal into text;

generate a second voice signal based on portions of the text attributable to the voice of the first user; and

transmit the second voice signal into the call.

9. The system of claim 8 , wherein the generated second voice signal excludes the portions of the text attributable to the voice of the second user.

10. The system of claim 8 , wherein the control circuitry is configured to generate the second voice signal based on portions of the text attributable to the voice of the first user by:

accessing a voice profile of the first user; and

synthesizing the portions of the text attributable to the voice of the first user into the second voice signal based on the voice profile of the first user.

11. The system of claim 8 , wherein the control circuitry is further configured to:

generate a third voice signal based on portions of the text attributable to the voice of the second user;

transmit the third voice signal into the call separately from the second voice signal.

12. The system of claim 11 , wherein the control circuitry is further configured to:

present, to other participants in the call, an option to select whether to listen to the second voice signal or the third voice signal.

13. The system of claim 8 , wherein the control circuitry, when preventing the entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call, is further configured to:

prevent transmission into the call of all voices present in the first voice signal.

14. The system of claim 8 , wherein the control circuitry is further configured to mute the microphone for a predetermined period of time in response to detecting greater than one voice in the first voice signal.

15. A system comprising:

means for capturing a first voice signal, the first voice signal comprising a voice of a first user and a voice of a second user, by a first device, during a call;

means for preventing the entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call;

means for converting the first voice signal into text;

means for generating a second voice signal based on portions of the text attributable to the voice of the first user; and

means for transmitting the second voice signal into the call.

16. The system of claim 15 , wherein the generated second voice signal excludes the portions of the text attributable to the voice of the second user.

17. The system of claim 15 , wherein the means for generating the second voice signal based on portions of the text attributable to the voice of the first user comprises:

means for accessing a voice profile of the first user; and

means for synthesizing the portions of the text attributable to the voice of the first user into the second voice signal based on the voice profile of the first user.

18. The system of claim 15 , further comprising:

means for generating a third voice signal based on portions of the text attributable to the voice of the second user;

means for transmitting the third voice signal into the call separately from the second voice signal.

19. The system of claim 18 , further comprising:

means for presenting, to other participants in the call, an option to select whether to listen to the second voice signal or the third voice signal.

20. The system of claim 15 , wherein the means for preventing the entirety of the first voice signal, the entirety of the first voice signal comprising the voice of the first user and the voice of the second user, from being transmitted into the call, wherein no part of the first voice signal is transmitted into the call, further comprises:

means for preventing transmission into the call of all voices present in the first voice signal.

Assignments (3)
CHANGE OF NAME Recorded Oct 4, 2024
From: ROVI GUIDES, INC.
To: ADEIA GUIDES INC.
Reel/Frame 069113/0392 →
SECURITY INTEREST Recorded May 3, 2023
From: ADEIA GUIDES INC.; ADEIA IMAGING LLC; ADEIA MEDIA HOLDINGS LLC; ADEIA MEDIA SOLUTIONS INC.; ADEIA SEMICONDUCTOR ADVANCED TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR BONDING TECHNOLOGIES INC.; ADEIA SEMICONDUCTOR INC.; ADEIA SEMICONDUCTOR SOLUTIONS LLC; ADEIA SEMICONDUCTOR TECHNOLOGIES LLC; ADEIA SOLUTIONS LLC
To: BANK OF AMERICA, N.A., AS COLLATERAL AGENT
Reel/Frame 063529/0272 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 10, 2022
From: PICHAIMURTHY, RAJENDRAN; SEETHARAM, MADHUSUDHAN
To: ROVI GUIDES, INC.
Reel/Frame 060769/0880 →
Continuity (2)
Continuation 17015832 · Sep 9, 2020
Related Publication 20220383888A1 · Dec 1, 2022