IP Library › Granted Patent US 12,373,701
Granted Patent B2
US 12,373,701 · App. 18/573,826 · Granted Jul 29, 2025

Method and apparatus for detecting and explaining anomalies

Inventors: Razieh Abbasi Ghalehtaki (Québec, CA); Fetahi Wuhib (Québec, CA); Amin Ebrahimzadeh (Québec, CA); Roch Glitho (Québec, CA)
Assignee: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
G06N3/088G06F11/08
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,373,701
App. No.
18/573,826
Granted
Jul 29, 2025
Kind
B2
Abstract

Embodiments described herein relate to a method and apparatus for detecting and explaining anomalies in data obtained from an environment using an encoder-decoder machine learning model. A state of the environment is represented by a plurality of features, and the machine learning model is trained with a first set of data samples. Each data sample in the first set of data samples comprises values for each of the plurality of features. The method comprises determining a respective first threshold for each of the plurality of features based on respective maximum reconstruction errors for each feature found during training of the encoder-decoder machine learning model; obtaining an anomalous data sample; determining respective reconstruction errors for each feature in the anomalous data sample using the trained encoder-decoder machine learning model; and determining one or more features in the anomalous data sample that are responsible for the anomalous data sample being anomalous responsive to the reconstruction errors associated with the one or more features being greater than or equal to the respective first thresholds.

Claims (45)

1. A method of detecting and explaining anomalies in data obtained from an environment using an encoder-decoder machine learning model, wherein a state of the environment is represented by a plurality of features, and wherein the encoder-decoder machine learning model is trained with a first set of data samples, and each data sample in the first set of data samples comprises values for each of the plurality of features, the method comprising:

determining a respective first threshold for each of the plurality of features based on respective maximum reconstruction errors for each feature found during training of the encoder-decoder machine learning model;

obtaining an anomalous data sample;

determining respective reconstruction errors for each feature in the anomalous data sample using the trained encoder-decoder machine learning model; and

determining one or more features in the anomalous data sample that are responsible for the anomalous data sample being anomalous responsive to the reconstruction errors associated with the one or more features being greater than or equal to the respective first thresholds.

2. The method of claim 1 , wherein the first set of data samples is assumed to contain no anomalies.

3. The method of claim 1 , further comprising training the encoder-decoder machine learning model with the first set of data samples.

4. The method of claim 1 , wherein the respective first threshold for each of the plurality of features is equal to the maximum reconstruction error for the respective feature.

5. The method of claim 1 , wherein the respective first threshold for each of the plurality of features is greater than 95th percentile of the maximum reconstruction error for the respective feature.

6. The method of claim 1 , wherein the respective first threshold for each of the plurality of features is set to the Nth percentile of the maximum reconstruction error for the respective feature responsive to the first set of data samples comprising (100−N) % anomalies, where N is an numerical value.

7. The method of claim 1 , further comprising training a sensitivity analysis model, wherein the sensitivity analysis model comprises a mapping function, by:

a. adding respective known noise signals to one or more samples in the first set of data samples to generate a noisy set of data samples;

b. for each sample in the noisy set of data samples:

i. inputting the sample into the trained encoder-decoder machine learning model to determine reconstruction errors for each feature in the sample; and

ii. determining which one or more features in the sample comprise noise signals by comparing the respective reconstruction errors for each feature to the respective first threshold; and

c. determining the mapping function such that the mapping function maps the reconstruction errors for features comprising a noise signal to the respective known noise signals added in step a.

8. The method of claim 7 , wherein the step of determining the mapping function comprises using multiple linear regression to determine a correlation between the reconstruction errors for features comprising a noise signal and two variables, wherein the two variables comprise the known noise signals and a number of noisy features in each data sample in the noisy set of data samples.

9. The method of claim 8 , further comprising:

for each of the one or more features that are responsible for the anomalous data sample being anomalous:

determining a noise level associated with the feature by using the mapping function to map a reconstruction error for the feature to a noise level.

10. The method of claim 1 , further comprising:

determining a second threshold based on a maximum average sample reconstruction error for the first set of data samples found during the training step, wherein an average sample reconstruction error for a sample is the average of the reconstruction errors for the features in said sample.

11. The method of claim 10 , wherein the step of obtaining an anomalous data sample comprises:

obtaining a first data sample in a second set of data samples;

inputting the first data sample into the encoder-decoder machine learning model to determine a first average sample reconstruction error for the first data sample;

comparing the first average sample reconstruction error to the second threshold; and

responsive to the first average sample reconstruction error being greater than the second threshold setting the first data sample as the anomalous data sample.

12. The method of claim 11 , further comprising:

responsive to the first average sample reconstruction error being less than or equal to the second threshold setting the first data sample as a normal data sample.

13. The method of claim 11 , further comprising:

calculating a first score value as the difference between the first average sample reconstruction error and the second threshold.

14. The method of claim 11 , wherein

the step of obtaining an anomalous data sample comprises:

determining a reconstruction error for each feature in the first data sample based on an output of the encoder-decoder machine learning model;

comparing each reconstruction error to the respective first thresholds for each feature; and

responsive to at least one of the reconstruction errors being greater than the respective first threshold, setting the first data sample as the anomalous data sample.

15. The method of claim 14 , further comprising:

responsive to all of the reconstruction errors being less than or equal to the respective first thresholds, setting the first data sample as a normal data sample.

16. The method of claim 14 , further comprising:

calculating a second score values for each feature as the difference between the reconstruction errors and the respective first thresholds;

sorting the one or more features in the anomalous data sample that are responsible for the anomalous data sample being anomalous based on the second score values associated with each of the one or more features.

17. The method of claim 1 , wherein the first set of data samples and the anomalous sample have been normalized.

18. The method of claim 1 , wherein the first set of data samples is calculated based on a rate of change of an original set of data samples.

19. An apparatus for detecting and explaining anomalies in data obtained from an environment using an encoder-decoder machine learning model, wherein a state of the environment is represented by a plurality of features, and wherein the machine learning model is trained with a first set of data samples, and each data sample in the first set of data samples comprises values for each of the plurality of features, the apparatus comprising processing circuitry configured to cause the apparatus to perform the method of claim 1 .

20. A non-transitory computer readable storage medium storing a computer program comprising instructions which, when executed on at least one processor, cause the at least one processor to carry out the method of claim 1 .

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 27, 2024
From: GHALEHTAKI, RAZIEH ABBASI; EBRAHIMZADEH, AMIN; GLITHO, ROCH; WUHIB, FETAHI
To: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
Reel/Frame 066569/0259 →
Continuity (1)
Related Publication 20240289209A1 · Aug 29, 2024
References Cited (14)
US 10574512B1 · Mermoud et al. · 2020 [cited by applicant]
US 20220303288A1 · Wang · 2022 [cited by examiner]
US 20230018848A1 · Eline · 2023 [cited by examiner]
WO 2018133924A1 · 2018 [cited by applicant]
Islam et al., Anomaly detection in a large-scale cloud platform, IEEE, pp. 150 to 159. (Year: 2021). [cited by examiner]
International Search Report and Written Opinion issued in International Application No. PCT/IB2021/055868 dated Mar. 29, 2022 (15 pages). [cited by applicant]
Dai, L. et al., “SDFVAE: Static and Dynamic Factorized VAE for Anomaly Detection of Multivariate CDN KPIs”, Proceedings of the 7th ACM Conference on Information-Centric Networking, ACMPUB27, Apr. 19, 2021 (11 pages). [cited by applicant]
Khan, S. et al., “Robustness of AI-based prognostic and systems health management”, Annual Reviews in Control, vol. 51, Jan. 1, 2021 (23 pages). [cited by applicant]
Li, J. et al., “Cognitive visual anomaly detection with constrained latent representations for industrial inspection robot”, Applied Soft Computing Journal, vol. 95, Jul. 16, 2020 (11 pages). [cited by applicant]
Hale, J., “Scale, Standardize, or Normalize with Scikit-Learn”, Mar. 14, 2019 (14 pages). [cited by applicant]
Su, Y. et al., “Robust Anomaly Detection for Multivariate Time Series through Stochastic Recurrent Neural Network”, Applied Data Science Track Paper, KDD '19 Aug. 4-8, 2019 (10 pages). [cited by applicant]
Provotar, O. I. et al., “Unsupervised Anomaly Detection in Time Series Using LSTM-Based Autoencoders”, IEEE, Jun. 6, 2021 (5 pages). [cited by applicant]
Malhotra, P. et al., “LSTM-based Encoder-Decoder for Multi-sensor Anomaly Detection”, Jul. 11, 2016 (5 pages). [cited by applicant]
“What is Kubernetes?”, Jul. 23, 2021 (4 pages). [cited by applicant]