IP Library Granted Patent US 12,425,782
Granted Patent B2
US 12,425,782 · App. 19/060,336 · Granted Sep 23, 2025

Ear-worn device with neural network for noise reduction and/or spatial focusing using multiple input audio signals

Inventors: Igor Lovchinsky (New York, NY); Israel Malkin (Manhattan Beach, CA); Jonathan Macoskey (Pittsburgh, PA); Philip Meyers, IV (San Francisco, CA); Andrew Casper (Inver Grove Heights, MN); Nicholas Morris (Brooklyn, NY); Matthew de Jonge (Brooklyn, NY)
Assignee: Fortell Research Inc.
H04R25/507H04R25/405H04R25/407H04S7/303H04S2400/11H04S2400/13H04S2400/15
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,425,782
App. No.
19/060,336
Granted
Sep 23, 2025
Kind
B2
Abstract

An ear-worn device may include two or more microphones configured to generate time-domain audio signals, each of the two or more microphones configured to generate one of the time-domain audio signals; processing circuitry comprising analog processing circuitry, digital processing circuitry, beamforming circuitry, and short-time Fourier transformation (STFT) circuitry, the processing circuitry configured to generate, from the time-domain audio signals, one or more frequency-domain non-beamformed audio signals and one or more frequency-domain beamformed signals; and enhancement circuitry comprising neural network circuitry configured to receive multiple frequency-domain input audio signals originating from the one or more frequency-domain non-beamformed audio signals and the one or more frequency-domain beamformed signals, and implement a single neural network trained to generate, based on the multiple frequency-domain input audio signals, a noise-reduced and spatially-focused output audio signal or an output for generating a noise-reduced and spatially-focused output audio signal.

Claims (52)

1. An ear-worn device, comprising:

two or more microphones configured to generate audio signals, each of the two or more microphones configured to generate one of the audio signals;

wherein the two or more microphones include a front microphone and a back microphone, and

wherein the front microphone is configured to generate a front time-domain audio signal and the back microphone is configured to generate a back time-domain audio signal;

processing circuitry comprising beamforming circuitry, the processing circuitry configured to generate, from the audio signals, one or more non-beamformed audio signals originating from the front time-domain audio signal and the back time-domain audio signal and one or more beamformed audio signals, the one or more non-beamformed audio signals each comprising a speech portion and a noise portion, and the one or more beamformed audio signals each comprising a speech portion and a noise portion; and

neural network circuitry downstream of the beamforming circuitry, the neural network circuitry configured to:

receive multiple input signals originating from the one or more non-beamformed audio signals and the one or more beamformed audio signals; and

implement one or more neural networks trained to generate, based on the multiple input signals:

a noise-reduced and/or spatially-focused output audio signal; or

an output configured to generate the noise-reduced and/or spatially-focused output audio signal.

2. The ear-worn device of claim 1 , wherein the two or more microphones comprise exactly two or exactly three microphones.

3. The ear-worn device of claim 1 , wherein:

the output audio signal is spatially focused; and

the generating of the output audio signal or the output configured to generate the output audio signal is based on a wearer selection of a size of a front-facing spatial region.

4. A system comprising:

the ear-worn device of claim 3 ; and

a processing device in communication with the ear-worn device and configured to receive the wearer selection of the size of the front-facing spatial region.

5. The system of claim 4 , wherein the processing device is configured to display multiple options for the size of the front-facing spatial region.

6. The system of claim 5 , wherein the processing device is configured to display exactly two, exactly three, or exactly four options for the size of the front-facing spatial region.

7. The system of claim 5 , wherein the processing device is configured, when displaying the multiple options for the size of the front-facing spatial region, to display graphical representations of the multiple options for the size of the front-facing spatial region.

8. The ear-worn device of claim 3 , wherein the output audio signal uses a mapping of gains to respective spatial regions relative to a wearer of the ear-worn device.

9. The ear-worn device of claim 8 , wherein the mapping of the gains to the respective spatial regions comprises mapping more than two spatial regions each to a different gain, and one or more of the spatial regions are processed with gains not equal to 1 or 0.

10. The ear-worn device of claim 1 , wherein the one or more beamformed audio signals comprise multiple beamformed audio signals each having a different directional pattern, and at least one of the multiple beamformed audio signals has a dipole, hypercardioid, supercardioid, or cardioid directional pattern.

11. The ear-worn device of claim 1 , wherein the neural network circuitry is configured to output a single output based on the multiple input signals.

12. An ear-worn device, comprising:

two or more microphones, wherein:

the two or more microphones comprise at least a front microphone and a back microphone; and

the front microphone is configured to generate a front time-domain audio signal and the back microphone is configured to generate a back time-domain audio signal;

processing circuitry comprising beamforming circuitry, the processing circuitry configured to generate, from the front time-domain audio signal and the back time-domain audio signal, multiple beamformed signals each having a different directional pattern; and

neural network circuitry downstream of the beamforming circuitry, the neural network circuitry configured to:

receive the multiple beamformed signals; and

implement one or more neural networks trained to generate, based on processing together the multiple beamformed signals, a mask configured to generate a spatially-focused output audio signal;

wherein the spatially-focused output audio signal has a mapping of gain to direction-of-arrival different from any directional pattern of the multiple beamformed signals.

13. The ear-worn device of claim 12 , wherein the two or more microphones comprise exactly two or exactly three microphones.

14. The ear-worn device of claim 12 , wherein:

the generating of the spatially-focused output audio signal is based on a wearer selection of a size of a front-facing spatial region.

15. A system comprising:

the ear-worn device of claim 14 ; and

a processing device in communication with the ear-worn device and configured to receive the wearer selection of the size of the front-facing spatial region.

16. The system of claim 15 , wherein the processing device is configured to display multiple options for the size of the front-facing spatial region.

17. The system of claim 16 , wherein the processing device is configured to display exactly two, exactly three, or exactly four options for the size of the front-facing spatial region.

18. The ear-worn device of claim 12 , wherein the mapping of gain to direction-of-arrival is relative to a wearer of the ear-worn device.

19. The ear-worn device of claim 18 , wherein the mapping of gain to direction-of-arrival comprises mapping more than two spatial regions each to a different gain, and one or more of the spatial regions are processed with gains not equal to 1 or 0.

20. The ear-worn device of claim 12 , wherein at least one of the multiple beamformed signals has a dipole, hypercardioid, supercardioid, or cardioid directional pattern.

21. The ear-worn device of claim 12 , wherein the neural network circuitry is configured to output a single output based on processing together the multiple beamformed signals.

22. The ear-worn device of claim 1 , wherein the noise-reduced and/or spatially-focused output audio signal comprises a spatially-focused output audio signal.

23. The ear-worn device of claim 22 , wherein the spatially-focused output audio signal has a mapping of gain to direction-of-arrival different from any directional pattern of the one or more beamformed audio signals.

24. The ear-worn device of claim 22 , wherein the spatially-focused output audio signal comprises speech components with different gains based on their different directions-of-arrival.

25. The ear-worn device of claim 1 , wherein the one or more beamformed audio signals comprise multiple beamformed audio signals each having a different directional pattern.

26. The ear-worn device of claim 1 , wherein the one or more beamformed audio signals comprise a front-facing beamformed signal and a back-facing beamformed signal.

27. The ear-worn device of claim 12 , wherein the spatially-focused output audio signal comprises speech components with different gains based on their different directions-of-arrival.

28. The ear-worn device of claim 12 , wherein the multiple beamformed signals comprise a front-facing beamformed signal and a back-facing beamformed signal.

Assignments (1)
CHANGE OF NAME Recorded Aug 18, 2025
From: CHROMATIC INC.
To: FORTELL RESEARCH INC.
Reel/Frame 072556/0785 →
Continuity (4)
Continuation 18592720 · Mar 1, 2024
Continuation 18477087 · Sep 28, 2023
Provisional Application 63517755 · Aug 4, 2023
Related Publication 20250193611A1 · Jun 12, 2025
References Cited (31)
US 10304475B1 · Wang et al. · 2019 [cited by applicant]
US 11134351B1 · Lunner · 2021 [cited by examiner]
US 11646009B1 · Chhetri · 2023 [cited by applicant]
US 11678111B1 · Messingher Lang et al. · 2023 [cited by applicant]
US 11711648B2 · Lopatka et al. · 2023 [cited by applicant]
US 20030063759A1 · Brennan et al. · 2003 [cited by applicant]
US 20080212810A1 · Pedersen · 2008 [cited by examiner]
US 20100027820A1 · Kates · 2010 [cited by applicant]
US 20100123785A1 · Chen · 2010 [cited by applicant]
US 20120250916A1 · Hain et al. · 2012 [cited by applicant]
US 20130070935A1 · Hui et al. · 2013 [cited by applicant]
US 20140219471A1 · Deshpande · 2014 [cited by applicant]
US 20160111113A1 · Cho · 2016 [cited by applicant]
US 20170213565A1 · Makinen · 2017 [cited by applicant]
US 20170295439A1 · Xu · 2017 [cited by applicant]
US 20180063654A1 · Kuriger · 2018 [cited by applicant]
US 20180146285A1 · Benattar · 2018 [cited by examiner]
US 20190208317A1 · Woodruff · 2019 [cited by applicant]
US 20190320260A1 · Alders · 2019 [cited by examiner]
US 20190335288A1 · Latypov · 2019 [cited by applicant]
US 20200322735A1 · Mosgaard et al. · 2020 [cited by applicant]
US 20210125625A1 · Huang · 2021 [cited by applicant]
US 20210281958A1 · Diehl et al. · 2021 [cited by applicant]
US 20220058773A1 · Chou et al. · 2022 [cited by applicant]
US 20220070585A1 · Chen · 2022 [cited by applicant]
US 20220124444A1 · Andersen · 2022 [cited by applicant]
US 20220232331A1 · Jelcicova · 2022 [cited by applicant]
US 20230034525A1 · Scheller · 2023 [cited by applicant]
WO 2023136835A1 · 2023 [cited by applicant]
Notification of Transmittal of The International Search Report and The Written Opinion of The International Searching Authority, or The Declaration, dated Jan. 15, 2025. [cited by applicant]
D. Markovic, et al. “Implicit Neural Spatial Filtering for Multichannel Source Separation in the Waveform Domain”, Meta, Reality Labs Research, Pittsburgh PA, USA, Metal AI Research, Paris, France, Jun. 30, 2022, 5 page… [cited by applicant]