IP Library Granted Patent US 12,342,149
Granted Patent B2
US 12,342,149 · App. 17/689,299 · Granted Jun 24, 2025

Method and apparatus with abnormal channel of microphone array detection and compensation signal generation

Inventors: Nam Soo Kim (Seoul, KR); Ji Won Yoon (Seoul, KR)
Assignees: Samsung Electronics Co., Ltd.; Seoul National University R&DB Foundation
H04S7/30G10L19/008G10L19/24H04R3/005H04R5/027H04S3/008G10L15/02H04S2400/01H04S2400/15
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,342,149
App. No.
17/689,299
Granted
Jun 24, 2025
Kind
B2
Abstract

Method and apparatus for detecting abnormal channel of microphone array detection and generating compensation signal are provided. A method includes receiving multi-channel sound source signals from a microphone array, synchronizing the multi-channel sound source signals based on spatial information of the microphone array, and detecting an abnormal channel of the microphone array by inputting the synchronized sound source signals and first conditional information to a neural network model configured to perform an inverse operation. The method further includes generating a compensation signal corresponding to an abnormal channel using a neural network model in response to an abnormal channel being detected.

Claims (59)

1. A method, comprising:

receiving multi-channel sound source signals from a microphone array;

synchronizing the multi-channel sound source signals based on spatial information of the microphone array; and

detecting an abnormal channel of the microphone array by inputting the synchronized sound source signals and first conditional information to a neural network model configured to perform an inverse operation, further including generating an intermediate compensation signal by inputting sampled data, sampled from an output of the neural network, and second conditional information to the neural network model and shifting the intermediate compensation signal based on the spatial information of the microphone array.

2. The method of claim 1 , further comprising:

generating a compensation signal corresponding to the abnormal channel using the neural network model in response to the abnormal channel being detected.

3. The method of claim 1 , further comprising:

determining a sound source signal of a reference channel of the microphone array from among the multi-channel sound source signals; and

determining the sound source signal of the reference channel to be the first conditional information.

4. The method of claim 3 , wherein the synchronizing further comprises shifting the multi-channel sound source signals based on the sound source signal of the reference channel.

5. The method of claim 1 , wherein the detecting of the abnormal channel comprises:

determining an output vector by inputting the synchronized sound source signals and the first conditional information to the neural network model;

determining a probability value corresponding to the output vector; and

detecting the abnormal channel of the microphone array by comparing the probability value to a threshold.

6. The method of claim 1 , wherein the spatial information of the microphone array comprises either one or both of shape information of the microphone array and distance information between channels included in the microphone array.

7. The method of claim 2 , wherein the generating of the compensation signal comprises:

sampling an arbitrary vector in a probability distribution corresponding to the output of the neural network model;

generating the intermediate compensation signal by inputting the arbitrary vector and the second conditional information to the neural network model; and

generating, as the compensation signal, a final compensation signal by shifting the intermediate compensation signal based on the spatial information of the microphone array.

8. The method of claim 7 , wherein the second conditional information comprises a sound source signal corresponding to one of the channels other than the abnormal channel.

9. A non-transitory computer-readable storage medium storing instructions that, when executed by one or more processors, configure the one or more processors to perform the method of claim 1 .

10. An apparatus, comprising:

one or more processors configured to:

receive multi-channel sound source signals from a microphone array;

synchronize the multi-channel sound source signals based on spatial information of the microphone array; and

detect an abnormal channel of the microphone array by inputting the synchronized sound source signals and first conditional information to a neural network model configured to perform an inverse operation, further including generating an intermediate compensation signal by inputting sampled data, sampled from an output of the neural network model, and second conditional information to the neural network model and shifting the intermediate compensation signal based on the spatial information of the microphone array.

11. The apparatus of claim 10 , wherein the one or more processors are further configured to generate a compensation signal corresponding to the abnormal channel using the neural network model in response to the abnormal channel being detected.

12. The apparatus of claim 10 , wherein the one or more processors are further configured to:

determine a sound source signal of a reference channel of the microphone array from among the multi-channel sound source signals; and

determine the sound source signal of the reference channel to be the first conditional information.

13. The apparatus of claim 12 , wherein, for the synchronizing, the one or more processors are further configured to shift the multi-channel sound source signals based on the sound source signal of the reference channel.

14. The apparatus of claim 10 , wherein, for the detecting of the abnormal channel, the one or more processors are further configured to:

determine an output vector by inputting the synchronized sound source signals and the first conditional information to the neural network model;

determine a probability value corresponding to the output vector; and

detect the abnormal channel of the microphone array by comparing the probability value to a threshold.

15. The apparatus of claim 10 , wherein the spatial information of the microphone array comprises either one or both of shape information of the microphone array and distance information between channels included in the microphone array.

16. The apparatus of claim 11 , wherein, for the generating of the compensation signal, the one or more processors are further configured to:

sample an arbitrary vector in a probability distribution corresponding to an output of the neural network model;

generate an intermediate compensation signal by inputting the arbitrary vector and second conditional information to the neural network model; and

generate, as the compensation signal, a final compensation signal by shifting the intermediate compensation signal based on the spatial information of the microphone array.

17. The apparatus of claim 16 , wherein the second conditional information comprises a sound source signal corresponding to one of the channels other than the abnormal channel.

18. An electronic device, comprising:

a microphone array configured to receive multi-channel sound source signals; and

one or more processors configured to:

synchronize the multi-channel sound source signals based on spatial information of the microphone array; and

detect an abnormal channel of the microphone array by inputting the synchronized sound source signals and first conditional information to a neural network model configured to perform an inverse operation, further including generating an intermediate compensation signal by inputting sampled data, sampled from an output of the neural network model, and second conditional information to the neural network model and shifting the intermediate compensation signal based on the spatial information of the microphone array.

19. The electronic device of claim 18 , wherein the one or more processors are further configured to generate a compensation signal corresponding to the abnormal channel using the neural network model in response to the abnormal channel being detected.

20. A method, comprising:

sampling an arbitrary vector in a probability distribution corresponding to an output of a neural network model configured to perform an inverse operation;

generating an intermediate compensation signal by inputting the arbitrary vector and first conditional information to the neural network model; and

generating a compensation signal corresponding to a detected abnormal channel of a microphone array by shifting the intermediate compensation signal based on spatial information of the microphone array.

21. The method of claim 20 , further comprising:

receiving multi-channel sound source signals from the microphone array;

synchronizing the multi-channel sound source signals based on the spatial information; and

detecting the abnormal channel by inputting the synchronized sound source signals and second conditional information to the neural network model.

22. The method of claim 21 , wherein the synchronizing further comprises:

determining delay times among the channels of the microphone array based on the spatial information of the microphone array; and

synchronizing the multi-channel sound source signals by shifting the multi-channel sound source signals based on the delay times.

23. The method of claim 20 , wherein the spatial information of the microphone array comprises information of an angle formed by the microphone array.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 8, 2022
From: KIM, NAM SOO; YOON, JI WON
To: SAMSUNG ELECTRONICS CO., LTD.; SEOUL NATIONAL UNIVERSITY R&DB FOUNDATION
Reel/Frame 059196/0453 →
Priority Claims (1)
KR 10-2021-0132102 · Oct 6, 2021 · national
Continuity (1)
Related Publication 20230104123A1 · Apr 6, 2023
References Cited (21)
US 10405115B1 · Landron et al. · 2019 [cited by applicant]
US 10959029B2 · Soto · 2021 [cited by examiner]
US 20170127180A1 · Shields et al. · 2017 [cited by applicant]
US 20170188138A1 · Makinen et al. · 2017 [cited by applicant]
US 20190043491A1 · Kupryjanow et al. · 2019 [cited by applicant]
US 20200328789A1 · Pritsker et al. · 2020 [cited by applicant]
US 20200366994A1 · Arteaga et al. · 2020 [cited by applicant]
CN 206931362U · 2018 [cited by examiner]
CN 110798790A · 2020 [cited by examiner]
CN 112348052A · 2021 [cited by examiner]
JP 2017090606A · 2017 [cited by examiner]
JP 2019200091A · 2019 [cited by applicant]
KR 101015102B1 · 2011 [cited by applicant]
KR 1020170050908A · 2017 [cited by applicant]
KR 1020190098981A · 2019 [cited by applicant]
KR 102199158B1 · 2021 [cited by applicant]
WO WO2019160070A1 · 2019 [cited by examiner]
WO WO2021041623A1 · 2021 [cited by examiner]
Harsh Purohit, Ryo Tanabe, Kenji Ichige, Takashi Endo, Yuki Nikaido, Kaori Suefusa, and Yohei Kawaguchi, MIMII Dataset: Sound Dataset for Malfunctioning Industrial Machine Investigation and Inspection, arXiv:1909.09347,… [cited by examiner]
Kim, Jinsung, et al. “Fault Detection in a Microphone Array by Intercorrelation of Features in Voice Activity Detection.” [cited by applicant]
Kirichenko, Polina, et al. “Why Normalizing Flows Fail to Detect Out-of-Distribution Data.” [cited by applicant]