IP Library Granted Patent US 12,407,972
Granted Patent B2
US 12,407,972 · App. 17/672,461 · Granted Sep 2, 2025

Auditory augmented reality using selective noise cancellation

Inventors: Stefan Marti (Oakland, CA); Joseph Verbeke (San Francisco, CA)
Assignee: Harman International Industries, Incorporated
H04R1/1083G10K11/178H04R2460/01
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,407,972
App. No.
17/672,461
Granted
Sep 2, 2025
Kind
B2
Abstract

Techniques for auditory augmented reality using selective noise cancellation include receiving an input signal capturing an ambient auditory environment; separating the input signal into a set of audio signals that includes first and second component signals; and in response to generating the set of audio signals: generating a context-sensitive user interface that displays a plurality of first controls for modifying the first component signal and a plurality of second controls for modifying the second component signal, the plurality of first controls being independent from the plurality of second controls; receiving, from a user, a selection to remove the first component signal; and in response to receiving the selection: removing the first component signal from the set of audio signals to generate a modified set of audio signals that includes the second component signal; and driving an audio output device to generate sound based on the modified set of audio signals.

Claims (69)

1. A computer-implemented method comprising:

receiving an input signal capturing an ambient auditory environment of a user;

separating the input signal into a set of audio signals that includes a first component signal of a first sound type and a second component signal of a second sound type; and

in response to generating the set of audio signals:

generating a context-sensitive user interface that concurrently displays a plurality of first controls and a plurality of second controls, the plurality of first controls being independent from the plurality of second controls;

receiving, from the user, a selection to remove the first component signal; and

in response to receiving the selection to remove the first component signal:

removing, using a first noise cancellation technique, the first component signal from the set of audio signals to generate a modified set of audio signals;

determining that a portion of the removed first component signal bleeds through an audio output device as a bleeding acoustic sound;

removing the bleeding acoustic sound using a second noise cancellation technique different from the first noise cancellation technique, wherein the second noise cancellation technique comprises generating an anti-noise signal to cancel the bleeding acoustic sound; and

driving the audio output device to generate sound based on the modified set of audio signals and the anti-noise signal.

2. The computer-implemented method of claim 1 , wherein: the plurality of first controls comprises controls for canceling the first component signal, attenuating the first component signal, reproducing the first component signal unchanged, and amplifying the first component signal.

3. The computer-implemented method of claim 1 , wherein the plurality of second controls comprises a continuum of sound pressure levels or a continuum of gain levels for adjusting a loudness of the second component signal.

4. The computer-implemented method of claim 1 , further comprising:

receiving, from the user, a selection to attenuate the second component signal; and

in response to receiving the selection to attenuate the second component signal, reducing a power level of the second component signal in the modified set of audio signals prior to driving the audio output device based on the modified set of audio signals.

5. The computer-implemented method of claim 1 , further comprising:

receiving, from the user, a selection to amplify the second component signal; and

in response to receiving the selection to amplify the second component signal, increasing a power level of the second component signal in the modified set of audio signals prior to driving the audio output device based on the modified set of audio signals.

6. The computer-implemented method of claim 1 , further comprising:

comparing the first component signal to a plurality of audio recordings, wherein each audio recording included in the plurality of audio recordings is associated with a sound source type;

identifying a first audio recording that is substantially similar to the first component signal;

determining that a sound source type associated with the first audio recording is the first sound type; and

associating the first component signal with the first sound type.

7. The computer-implemented method of claim 1 , wherein the first component signal comprises first ambient sounds originating from a first direction, and wherein the first direction is determined based on a direction and an angular width specified by the user.

8. The computer-implemented method of claim 1 , wherein generating the modified set of audio signals further comprises:

combining a first audio channel associated with the first component signal and a second audio channel associated with the second component signal into a combined audio channel that does not include the removed first component signal.

9. The computer-implemented method of claim 1 , wherein the first noise cancellation technique comprises applying active noise cancellation to non-periodic sounds included in the first component signal.

10. The computer-implemented method of claim 1 , wherein the first noise cancellation technique comprises attenuating specific frequency values of the first sound type.

11. The computer-implemented method of claim 1 , wherein the first noise cancellation technique comprises applying a bandpass filter to attenuate a pre-defined frequency range associated with the first sound type.

12. A system comprising:

at least one audio sensor that acquires sound from an environment of a user and produces an input signal;

an audio output device; and

a first computing device comprising one or more processors coupled to the at least one audio sensor that performs the steps of:

receiving the input signal;

separating the input signal into a set of audio signals that includes a first component signal of a first sound type and a second component signal of a second sound type; and

in response to generating the set of audio signals:

generating a context-sensitive user interface that concurrently displays a plurality of first controls and a plurality of second controls, the plurality of first controls being independent from the plurality of second controls;

receiving, from the user, a selection to remove the first component signal; and

in response to receiving the selection to remove the first component signal:

removing, using a first noise cancellation technique, the first component signal from the set of audio signals to generate a modified set of audio signals;

determining that a portion of the removed first component signal bleeds through the audio output device as a bleeding acoustic sound;

removing the bleeding acoustic sound using a second noise cancellation technique different from the first noise cancellation technique, wherein the second noise cancellation technique comprises generating an anti-noise signal to cancel the bleeding acoustic sound; and

driving the audio output device to generate sound based on the modified set of audio signals and the anti-noise signal.

13. The system of claim 12 , wherein the at least one audio sensor includes an array of audio sensors or a set of headphones.

14. The system of claim 12 , wherein the one or more processors further perform the step of providing the context-sensitive user interface to a second computing device for presentation to the user, the second computing device being separate from the first computing device.

15. The system of claim 12 , wherein generating the context-sensitive user interface comprises:

determining a location of the environment, and loading the context-sensitive user interface that is associated with the location.

16. The system of claim 12 , wherein the one or more processors further perform the steps of:

receiving an audio source signal from an audio source; and

adding the audio source signal to the modified set of audio signals prior to driving the audio output device based on the modified set of audio signals.

17. The system of claim 12 , wherein the context-sensitive user interface comprises a continuum of levels for adjusting a loudness of the second component signal, and wherein the one or more processors further perform the steps of:

receiving, from the user via the context-sensitive user interface, a first level from the continuum of levels; and

amplifying or reducing a power level of the second component signal based on the first level.

18. One or more non-transitory computer-readable media including instructions that, when executed by one or more processors, cause the one or more processors to perform steps comprising:

receiving an input signal capturing an ambient auditory environment of a user;

separating the input signal into a set of audio signals that includes a first component signal of a first sound type and a second component signal of a second sound type; and

in response to generating the set of audio signals:

generating a context-sensitive user interface that concurrently displays a plurality of first controls and a plurality of second controls, the plurality of first controls being independent from the plurality of second controls;

receiving, from the user, a selection to remove the first component signal; and

in response to receiving the selection to remove the first component signal:

removing, using a first noise cancellation technique, the first component signal from the set of audio signals to generate a modified set of audio signals;

determining that a portion of the removed first component signal bleeds through an audio output device as a bleeding acoustic sound;

removing the bleeding acoustic sound using a second noise cancellation technique different from the first noise cancellation technique, wherein the second noise cancellation technique comprises generating an anti-noise signal to cancel the bleeding acoustic sound; and

driving the audio output device to generate sound based on the modified set of audio signals and the anti-noise signal.

19. The one or more non-transitory computer-readable media of claim 18 , wherein removing the first component signal from the set of audio signals comprises applying a bandpass filter to the set of audio signals.

20. The one or more non-transitory computer-readable media of claim 18 , wherein the steps further comprise:

generating, based on the first component signal, an inverse signal that is a polar inverse of the first component signal; and

adding the inverse signal to the modified set of audio signals prior to driving the audio output device based on the modified set of audio signals.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 15, 2022
From: MARTI, STEFAN; VERBEKE, JOSEPH
To: HARMAN INTERNATIONAL INDUSTRIES, INCORPORATED
Reel/Frame 059020/0157 →
Continuity (2)
Continuation 16907063 · Jun 19, 2020
Related Publication 20220174395A1 · Jun 2, 2022
References Cited (27)
US 5027410A · Williamson et al. · 1991 [cited by applicant]
US 6989744B2 · Proebsting · 2006 [cited by applicant]
US 7512247B1 · Odinak et al. · 2009 [cited by applicant]
US 8081780B2 · Goldstein et al. · 2011 [cited by applicant]
US 9622013B2 · Di Censo et al. · 2017 [cited by applicant]
US 9716939B2 · Di Censo · 2017 [cited by examiner]
US 9727129B2 · Di Censo · 2017 [cited by examiner]
US 10186248B2 · Zukowski et al. · 2019 [cited by applicant]
US 10388297B2 · Di Censo et al. · 2019 [cited by applicant]
US 10575117B2 · Di Censo · 2020 [cited by examiner]
US 20050036637A1 · Janssen · 2005 [cited by applicant]
US 20050226425A1 · Polk, Jr. · 2005 [cited by applicant]
US 20080267416A1 · Goldstein et al. · 2008 [cited by applicant]
US 20100076793A1 · Goldstein · 2010 [cited by examiner]
US 20110069843A1 · Cohen et al. · 2011 [cited by applicant]
US 20110107216A1 · Bi · 2011 [cited by applicant]
US 20120215519A1 · Park et al. · 2012 [cited by applicant]
US 20130273967A1 · Dave et al. · 2013 [cited by applicant]
US 20150071457A1 · Burciu · 2015 [cited by applicant]
US 20180302738A1 · Di Censo · 2018 [cited by examiner]
US 20200382859A1 · Woodruff · 2020 [cited by examiner]
US 20210195278A1 · Philipp · 2021 [cited by examiner]
US 20210264931A1 · Leider · 2021 [cited by examiner]
WO 2008103925A1 · 2008 [cited by applicant]
WO WO2011118595A1 · 2011 [cited by examiner]
Basu et al., “Smart Headphones”, http://alumni.media.mil.edu/-sbasu/papers/chi2001.pdf, Proceedings of CHI 2001, Seattle, WA, 2001, pp. 267-268. [cited by applicant]
Mueller, et al., “Transparent Hearing”, http://lloydmueller.com/projects/transparent_hearing/, CHI, 2002, 1 page. [cited by applicant]