IP Library › Granted Patent US 12,266,378
Granted Patent B2
US 12,266,378 · App. 17/615,691 · Granted Apr 1, 2025

Sound modification based on frequency composition

Inventors: Joseph Verbeke (San Francisco, CA); Stefan Marti (Oakland, CA)
Assignee: Harman International Industries, Incorporated
G10L21/034G06F3/04847G06F3/165G10L25/18G10L25/51H03F3/183H03F2200/03
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,266,378
App. No.
17/615,691
Granted
Apr 1, 2025
Kind
B2
Abstract

In various embodiments, a sound modification application selectively modifies one or more sounds included in one or more audio signals. In operation, the sound modification application determines classifications associated with multiple sounds included in one or more audio signals. The sound modification application selects a first frequency sub-band of a first sound included in the multiple sounds based on a first classification associated with the first sound. The sound modification application then modifies the first frequency sub-band of the first sound, without modifying at least a second frequency sub-band of the first sound, to generate a modified audio signal.

Claims (42)

1. A computer-implemented method for modifying a sound included in an audio signal, comprising:

determining, for each sound included in a plurality of sounds included in at least one audio signal, one or more classifications associated with the sound;

selecting a first frequency sub-band of a first sound included in the plurality of sounds based on a first classification associated with the first sound; and

modifying the first frequency sub-band of the first sound, to generate a modified first frequency sub-band; and

generating a second audio signal by combining the modified first frequency sub-band and an unmodified version of a second frequency sub-band of the first sound.

2. The method of claim 1 , further comprising:

selecting a third frequency sub-band of a second sound included in the plurality of sounds based on at least one of a second classification associated with the second sound, an analysis of the second sound, a frequency range of the first frequency sub-band, and a center frequency of the first frequency sub-band; and

modifying the third frequency sub-band of the second sound, without modifying at least a fourth frequency sub-band of the second sound.

3. The method of claim 1 , wherein modifying the first frequency sub-band of the first sound comprises performing parametric equalization on the first frequency sub-band.

4. The method of claim 1 , further comprising receiving a user input, wherein the modifying is performed based on the user input.

5. The method of claim 4 , further comprising displaying a user interface, the user interface comprising a control object for each sound included in the plurality of sounds, wherein the user input is received via a control object corresponding to the first sound.

6. The method of claim 1 , wherein selecting the first frequency sub-band of the first sound comprises:

obtaining, from a database, characteristic frequency information associated with the one or more classifications associated with the first sound; and

selecting the first frequency sub-band based on the information.

7. The method of claim 1 , wherein the modifying is performed in response to a determination that the first sound perceptually competes with a second sound included in the plurality of sounds.

8. One or more non-transitory computer readable storage media storing instructions, that, when executed by at least one processor, cause the at least one processor to perform the steps of:

determining, for each sound included in a plurality of sounds included in at least one audio signal, one or more classifications associated with the sound;

selecting a first frequency sub-band of a first sound included in the plurality of sounds based on a first classification associated with the first sound;

modifying the first frequency sub-band of the first sound to generate a first modified frequency sub-band;

selecting a second frequency sub-band of a second sound included in the plurality of sounds;

modifying the second frequency sub-band of the second sound to generate a second modified frequency sub-band; and

generating at least a second audio signal by combining the first modified frequency sub-band and a first unmodified version of a third frequency sub-band of the first sound and by combining the second modified frequency sub-band and a second unmodified version of a fourth frequency sub-band of the second sound.

9. The one or more computer-readable storage media of claim 8 , wherein modifying the first frequency sub-band of the first sound comprises performing parametric equalization on the first frequency sub-band.

10. The one or more computer-readable storage media of claim 8 , further comprising receiving a user input, wherein modifying the first frequency sub-band of the first sound comprises increasing or decreasing an amplitude of the first frequency sub- band based on the user input.

11. The one or more computer-readable storage media of claim 8 , further comprising displaying a user interface, the user interface comprising a control object for each sound included in the plurality of sounds.

12. The one or more computer-readable storage media of claim 8 , wherein selecting the second frequency sub-band of the second sound comprises selecting the second frequency sub-band based on at least one of a second classification associated with the second sound, an analysis of the second sound, a frequency range of the first frequency sub-band, and a center frequency of the first frequency sub-band.

13. The one or more computer-readable storage media of claim 8 , wherein selecting the first frequency sub-band of the first sound comprises:

obtaining, from a database, characteristic frequency information associated with the one or more classifications associated with the first sound; and

selecting the first frequency sub-band based on the information.

14. A system, comprising:

a memory; and

at least one processor coupled to the memory and configured to:

detect a plurality of sounds included in at least one audio signal;

determine, for each sound included in the plurality of sounds, one or more classifications associated with the sound;

select a first frequency sub-band of a first sound included in the plurality of sounds based on a first classification associated with the first sound; and

modify the first frequency sub-band of the first sound to generate a modified first frequency sub-band; and

generate a second audio signal by combining the modified first frequency sub-band and an unmodified version of a second frequency sub-band of the first sound.

15. The system of claim 14 , wherein the at least one processor is further configured to:

select a third frequency sub-band of a second sound included in the plurality of sounds based on at least one of a second classification associated with the second sound, an analysis of the second sound, a frequency range of the first frequency sub-band, and a center frequency of the first frequency sub-band; and

modify the third frequency sub-band of the second sound, without modifying at least a fourth frequency sub-band of the second sound.

16. The system of claim 14 , wherein the first classification is one of a human voice, an animal sound, or an object sound.

17. The system of claim 14 , further comprising a database, wherein the database comprises at least one mapping of the first classification to one or more characteristic frequency sub-bands, and wherein the one or more characteristic frequency sub-bands include the first frequency sub-band.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 9, 2021
From: VERBEKE, JOSEPH; MARTI, STEFAN
To: HARMAN INTERNATIONAL INDUSTRIES, INCORPORATED
Reel/Frame 058345/0558 →
Continuity (1)
Related Publication 20220246161A1 · Aug 4, 2022
References Cited (13)
US 7590543B2 · Kjorling · 2009 [cited by examiner]
US 9609383B1 · Hirst · 2017 [cited by examiner]
US 10014002B2 · Koretzky · 2018 [cited by examiner]
US 11373672B2 · Mesgarani · 2022 [cited by examiner]
US 11606663B2 · Vetter · 2023 [cited by examiner]
US 20130097510A1 · Maling, III · 2013 [cited by examiner]
US 20160165336A1 · Di Censo · 2016 [cited by examiner]
US 20170316792A1 · Chaudhuri · 2017 [cited by examiner]
US 20170374478A1 · Jones · 2017 [cited by examiner]
US 20180122403A1 · Koretzky · 2018 [cited by examiner]
US 20200329322A1 · Ramsay · 2020 [cited by examiner]
CN 105679302A · 2016 [cited by applicant]
International Search Report and Written Opinion, PCT/IB2019/054648, Mar. 4, 2020, 13 pages. [cited by applicant]