IP Library › Granted Patent US 12,432,498
Granted Patent B2
US 12,432,498 · App. 18/353,282 · Granted Sep 30, 2025

Selective modification of stereo or spatial audio

Inventors: Lasse Juhani Laaksonen (Tampere, FI); Miikka Tapani Vilermo (Tampere, FI); Arto Juhani Lehtiniemi (Tampere, FI)
Assignee: NOKIA TECHNOLOGIES OY
H04R5/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,432,498
App. No.
18/353,282
Granted
Sep 30, 2025
Kind
B2
Abstract

An apparatus, method and computer program is described comprising: providing a stereo or spatial audio signal produced using individual signals from respective microphones of a user device, the stereo or spatial audio signal representing an audio scene; detecting unwanted noise in at least one of the individual signals; responsive to detecting the unwanted noise, modifying the stereo or spatial audio signal based on a determined level of spatial interest in the audio scene meeting a predetermined condition; and providing the modified audio signal for output via one or more speakers.

Claims (42)

1. An apparatus comprising:

at least one processor; and

at least one memory storing instructions that, when executed by the at least one processor, cause the apparatus at least to:

provide a stereo or spatial audio signal produced using individual signals from respective microphones of a user device, the stereo or spatial audio signal representing an audio scene;

detect unwanted noise in at least one of the individual signals;

responsive to detecting the unwanted noise, modify the stereo or spatial audio signal based on a determined level of spatial interest in the audio scene meeting a predetermined condition, wherein the determined level of spatial interest in the audio scene is a value based at least in part on a number of significant audio sources in the audio scene; and

provide the modified audio signal for output via one or more speakers.

2. The apparatus of claim 1 , wherein the modifying further comprises;

produce at least one of a reduced directional or spatial representation of the audio scene.

3. The apparatus of claim 2 , wherein the audio scene is represented by a stereo signal and wherein the modifying comprises producing a monaural version of the audio scene.

4. The apparatus of claim 3 , wherein the modifying comprises producing a monaural version of the audio scene and providing the monaural version on first and second channels for stereo output via at least two speakers.

5. The apparatus of claim 2 , wherein the audio scene is represented by a spatial audio signal and wherein the modifying further comprises; produce a stereo or monaural version of the audio scene.

6. The apparatus of claim 2 , wherein the modifying further comprises; produce the at least one of reduced directional or spatial representation of the audio scene by suppressing the at least one individual signal in which the unwanted noise is detected.

7. The apparatus of claim 6 , wherein the modifying further comprises; suppress the at least one individual signal by disabling the at least one of respective microphones which produce the at least one individual signal in which the unwanted noise is detected.

8. The apparatus of claim 1 , wherein the modifying further comprises; modify the stereo or spatial audio signal until the unwanted noise is at least no longer detected in at least one of the individual signals.

9. The apparatus of claim 1 , wherein the unwanted noise is wind noise.

10. The apparatus of claim 1 , wherein the apparatus is further caused to:

identify significant audio sources in the audio scene, wherein the predetermined condition is met if the value is below a predetermined threshold.

11. The apparatus of claim 10 , wherein the significant audio sources are identified based on respective properties of one or more audio sources in the audio scene, the respective properties comprising one or more of:

frequency band;

energy level;

type of audio source;

temporal activity over a predetermined time period;

direction relative to a reference direction of the user device; or

direction relative to a gaze direction of a user of the user device.

12. The apparatus of claim 10 , wherein the significant audio sources are identified based at least in part on one or more of the audio sources in the audio scene being speech-type audio source.

13. The apparatus of claim 10 , wherein the significant audio sources are identified based at least in part on one or more of the audio sources in the audio scene having a direction within a predetermined angle of a reference direction of the user device or gaze direction of the user of the user device.

14. The apparatus of claim 13 , wherein the reference direction of the user device corresponds with a direction of a camera of the user device.

15. The apparatus of claim 13 , wherein the predetermined angle is substantially 180 degrees or less.

16. A method comprising:

providing a stereo or spatial audio signal produced using individual signals from respective microphones of a user device, the stereo or spatial audio signal representing an audio scene;

detecting unwanted noise in at least one of the individual signals;

responsive to detecting the unwanted noise, modifying the stereo or spatial audio signal based on a determined level of spatial interest in the audio scene meeting a predetermined condition, wherein the determined level of spatial interest in the audio scene is a value based at least in part on a number of significant audio sources in the audio scene; and

providing the modified audio signal for output via one or more speakers.

17. The method of claim 16 , wherein the providing the modified audio signal further comprises producing at least one of a reduced directional or spatial representation of the audio scene.

18. The method of claim 17 , wherein the audio scene is represented by a stereo signal and wherein the providing the modified audio signal further comprises producing a monaural version of the audio scene.

19. The method of claim 17 , wherein the audio scene is represented by a spatial audio signal and wherein the providing the modified audio signal further comprises producing a stereo or monaural version of the audio scene.

20. A non-transitory computer readable medium comprising program instructions stored thereon for performing at least the following:

providing a stereo or spatial audio signal produced using individual signals from respective microphones of a user device, the stereo or spatial audio signal representing an audio scene;

detecting unwanted noise in at least one of the individual signals;

responsive to detecting the unwanted noise, modifying the stereo or spatial audio signal based on a determined level of spatial interest in the audio scene meeting a predetermined condition, wherein the determined level of spatial interest in the audio scene is a value based at least in part on a number of significant audio sources in the audio scene; and

providing the modified audio signal for output via one or more speakers.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 3, 2023
From: JUHANI LAAKSONEN, LASSE; TAPANI VILERMO, MIIKKA; JUHANI LEHTINIEMI, ARTO
To: NOKIA TECHNOLOGIES OY
Reel/Frame 064783/0961 →
Priority Claims (1)
EP 22190219 · Aug 12, 2022 · regional
Continuity (1)
Related Publication 20240056734A1 · Feb 15, 2024
References Cited (14)
US 11217264B1 · Yang et al. · 2022 [cited by applicant]
US 11258940B2 · Kasugai · 2022 [cited by examiner]
US 20070021958A1 · Visser · 2007 [cited by examiner]
US 20080226098A1 · Haulick et al. · 2008 [cited by applicant]
US 20100123785A1 · Chen · 2010 [cited by examiner]
US 20130342731A1 · Lee · 2013 [cited by examiner]
US 20190253795A1 · Ozcan · 2019 [cited by examiner]
US 20220021970A1 · Vilermo et al. · 2022 [cited by applicant]
US 20220068290A1 · Mate et al. · 2022 [cited by applicant]
US 20230319469A1 · Vilkamo · 2023 [cited by examiner]
JP H03106299A · 1991 [cited by applicant]
“IVAS Design Constraints (IVAS-4)”, 3GPP TSG SA WG4#118-e meeting, S4-220551, Agenda: 15.2, Editor, Apr. 6-14, 2022, pp. 1-15. [cited by applicant]
Extended European Search Report received for corresponding European Patent Application No. 22190219.0, dated Feb. 2, 2023, 11 pages. [cited by applicant]
“AirPods redefine the personal audio experience”, Apple, Retrieved on Jul. 13, 2023, Webpage available at : https://www.apple.com/newsroom/2023/06/airpods-redefine-the-personal-audio-experience/. [cited by applicant]