IP Library Granted Patent US 12,399,985
Granted Patent B2
US 12,399,985 · App. 18/005,761 · Granted Aug 26, 2025

DPOD: differentially private outsourcing of anomaly detection

Inventors: Meisam Mohammady (Montreal, CA); Mengyuan Zhang (Hong Kong, CN); Yosr Jarraya (Montreal, CA); Makan Pourzandi (Montreal, CA); Han Wang (Chicago, IL); Yuan Hong (Saunderstown, RI); Lingyu Wang (Montreal, CA); Suryadipta Majumdar (Montreal, CA); Mourad Debbabi (Dollard des Ormeaux, CA)
Assignee: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
G06F21/554G06F21/6245G06F2221/033
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,399,985
App. No.
18/005,761
Granted
Aug 26, 2025
Kind
B2
Abstract

A method, system and apparatus are disclosed. According to one or more embodiments, a data node is provided. The data node includes processing circuitry configured to: receive an anomaly estimation for a first privatized dataset, the first private dataset being based on a dataset and a first noise profile, apply a second noise profile to the dataset to generate a second privatized dataset, the second noise profile being based at least on the anomaly estimation, and optionally cause transmission of the second privatized dataset for anomaly estimation.

Claims (68)

1. A data node configured to communicate with a data analyzer node and a plurality of data owners, the data node comprising processing circuitry configured to:

receive from each data owner of the plurality of data owners, a dataset having a respective privacy requirement and a sensitivity;

apply to each dataset a respective one of a plurality of first noise profiles, each first noise profile configured to add noise to a respective dataset according to the respective privacy requirements of the dataset to provide a plurality of first privatized datasets;

release the plurality of first privatized data sets to the data analyzer node;

receive from the data analyzer node an anomaly estimation for each first privatized dataset, the anomaly estimation being indicative of outlier data in a respective first privatized dataset;

update the respective sensitivity for each first privatized dataset based at least in part on the anomaly estimations;

apply a second noise profile configured to reduce noise of outlier data of each first privatized dataset to generate respective second privatized datasets, the second noise profile being selected based at least in part on the respective updated sensitivity.

2. The data node of claim 1 , wherein the processing circuitry is further configured to determine an updated sensitivity value for the dataset based at least on the anomaly estimation and a privacy budget.

3. The data node of claim 1 , wherein a received dataset includes outlier data and non-outlier data, the second noise profile being configured to reduce noise applied to outlier data when compared to a noise applied to outlier data by the first noise profile.

4. The data node of claim 1 , wherein a received dataset includes outlier data and non-outlier data, the second noise profile being configured to increase noise applied to non-outlier data when compared to noise applied to the non-outlier data by the first noise profile.

5. The data node of claim 1 , where the processing circuitry is further configured to apply the first noise profile to outlier data in a received dataset; and

applying the second noise profile to a first privatized dataset to generate the second privatized dataset includes applying the second noise profile to non-outlier data in the dataset.

6. The data node of claim 1 , wherein the anomaly estimation indicates an anomaly score for a respective privatized dataset.

7. The data node of claim 1 , wherein the processing circuitry is further configured to receive a sensitivity estimation that is based on the first privatized dataset, the sensitivity estimation indicating whether to modify the first noise profile, the second noise profile being further based at least on the sensitivity estimation.

8. The data node of claim 7 , wherein the sensitivity estimation indicates to reduce privacy for outlier data of the dataset while at least maintaining privacy for non-outlier data of the dataset.

9. The data node of claim 1 , wherein a first noise profile provides a first data sensitivity value for a first privatized dataset; and

the second noise profile provides a second data sensitivity value for the second privatized dataset, the second data sensitivity value being different from the first data sensitivity value.

10. The data node of claim 1 , wherein respective noise profiles of the plurality of first noise profiles each correspond to a respective differential privacy mechanism configured to quantify a privacy level provided to data of the respective received dataset.

11. The data node of claim 1 , wherein a first privatized dataset is a first privatized histogram of the respective received dataset and the second privatized dataset is a second privatized histogram of the first privatized dataset.

12. A data analyst node configured to communicate with a data node, the data analyst node comprising processing circuitry configured to:

receive from the data node a plurality of first privatized data sets, each first privatized data set being based on application of a respective first noise profile of a plurality of noise profiles configured to add noise to a respective dataset received by the data node from a different data owner of a plurality of data owners, each received respective dataset having a respective privacy requirement and a sensitivity;

determine an anomaly estimation for each first privatized dataset, the anomaly estimation being indicative of outlier data in a respective first privatized dataset;

transmit to the data node the anomaly estimation for a first privatized dataset;

receive a second privatized dataset, the second privatized dataset being based on application of a second noise profile configured to reduce noise of outlier data of a first privatized dataset based at least in part on an updated sensitivity for the first privatized database, the second noise profile being based at least on the anomaly estimation; and

perform anomaly estimation for the second privatized dataset.

13. The data analyst node of claim 12 , wherein a received first privatized dataset includes outlier data and non-outlier data, the second noise profile being configured to reduce noise applied to outlier data when compared to a noise applied to outlier data by the first noise profile.

14. The data analyst node of claim 12 , wherein a received first privatized dataset includes outlier data and non-outlier data, the second noise profile being configured to increase noise applied to non-outlier data when compared to noise applied to the non-outlier data by the first noise profile.

15. The data analyst node of claim 12 , wherein the anomaly estimation indicates an anomaly score for a respective privatized dataset.

16. The data analyst node of claim 12 , wherein a respective first noise profile provides a first data sensitivity value for a respective first privatized dataset; and

the second noise profile provides a second data sensitivity value for the second privatized dataset, the second data sensitivity value being different from the first data sensitivity value.

17. The data analyst node of claim 12 , wherein the processing circuitry is further configured to determine a sensitivity estimation that is based on a first privatized dataset, the sensitivity estimation indicating whether to modify the respective first noise profile, the second noise profile being further based at least on the sensitivity estimation.

18. The data analyst node of claim 12 , wherein the sensitivity estimation indicates to reduce privacy for outlier data of the dataset while at least maintaining privacy for non-outlier data of the dataset.

19. The data analyst node of claim 12 , wherein the processing circuitry is further configured to cause transmission of the anomaly estimation for the second privatized dataset.

20. The data analyst node of claim 12 , wherein a first privatized dataset is a first privatized histogram of the first privatized dataset and the second privatized dataset is a second privatized histogram of the dataset.

21. A method implemented by a data node configured to communicate with a data analyzer node and a plurality of data owners, the method comprising:

receiving from each data owner of the plurality of data owners, a dataset having a respective privacy requirement and a sensitivity;

applying to each dataset a respective one of a plurality of first noise profiles, each first noise profile configured to add noise to a respective dataset according to the respective privacy requirements of the dataset to provide a plurality of first privatized datasets;

releasing the plurality of first privatized data sets to the data analyzer node;

receiving from the data analyzer node an anomaly estimation for each first privatized dataset, the anomaly estimation being indicative of outlier data in a respective first privatized dataset;

updating the respective sensitivity for each first privatized dataset based at least in part on the anomaly estimations;

applying a second noise profile to the configured to reduce noise of outlier data of each first privatized dataset to generate respective second privatized datasets, the second noise profile being selected based at least in part on respective updated sensitivity.

22. The method of claim 21 , further comprising determining an updated sensitivity value for the dataset based at least on the anomaly estimation and privacy.

23. The method of claim 21 , wherein a received dataset includes outlier data and non-outlier data, the second noise profile being configured to reduce noise applied to outlier data when compared to a noise applied to outlier data by the first noise profile.

24. The method of claim 21 , wherein a received dataset includes outlier data and non-outlier data, the second noise profile being configured to increase noise applied to non-outlier data when compared to noise applied to the non-outlier data by the first noise profile.

25. The method of claim 21 , further comprising applying the first noise profile to outlier data in a received dataset; and

applying of the second noise profile to a first privatized dataset to generate the second privatized dataset includes applying the second noise profile to non-outlier data in the dataset.

26. The method of claim 21 , wherein the anomaly estimation indicates an anomaly score for a respective privatized dataset.

27. The method of claim 21 , further comprising receiving a sensitivity estimation that is based on a first privatized dataset, the sensitivity estimation indicating whether to modify a first noise profile, the second noise profile being further based at least on the sensitivity estimation.

28. The method of claim 27 , wherein the sensitivity estimation indicates to reduce privacy for outlier data of a second privatized dataset while at least maintaining privacy for non-outlier data of the dataset.

29. The method of claim 21 , wherein a first noise profile provides a first data sensitivity value for a first privatized dataset; and

the second noise profile provides a second data sensitivity value for the second privatized dataset, the second data sensitivity value being different from the first data sensitivity value.

30. The method of claim 21 , wherein the respective noise profiles of the plurality of first noise profiles each correspond to a respective differential privacy mechanism configured to quantify a privacy level provided to data of the respective received dataset.

31. The method of claim 21 , wherein a first privatized dataset is a first privatized histogram of the respective received dataset and the second privatized dataset is a second privatized histogram of the first privatized dataset.

32. A method implemented by a data analyst node configured to communicate with a data node, the method comprising:

receiving from the data node a plurality of first privatized data sets, each first privatized data set being based on application of a respective first noise profile of a plurality of noise profiles configured to add noise to a respective dataset received by the data node from a different data owner of a plurality of data owners, each received respective dataset having a respective privacy requirement and a sensitivity;

determining an anomaly estimation for each first privatized dataset, the anomaly estimation being indicative of outlier data in a respective first privatized dataset;

transmit the anomaly estimation for a first privatized dataset;

receiving a second privatized dataset the second privatized dataset being based on application of a second noise profile configured to reduce noise of outlier data of a first privatized dataset based at least in part on an updated sensitivity for the first privatized database, the second noise profile being based at least on the anomaly estimation; and

performing anomaly estimation for the second privatized dataset.

33. The method of claim 32 , wherein a received first privatized dataset includes outlier data and non-outlier data, the second noise profile being configured to reduce noise applied to outlier data when compared to a noise applied to outlier data by the first noise profile.

34. The method of claim 32 , wherein a received first privatized dataset includes outlier data and non-outlier data, the second noise profile being configured to increase noise applied to non-outlier data when compared to noise applied to the non-outlier data by the first noise profile.

35. The method of claim 32 , wherein the anomaly estimation indicates an anomaly score for a respective privatized dataset.

36. The method of claim 32 , wherein a respective first noise profile provides a first data sensitivity value for a respective first privatized dataset; and

the second noise profile provides a second data sensitivity value for the second privatized dataset, the second data sensitivity value being different from the first data sensitivity value.

37. The method of claim 32 , further comprising determining a sensitivity estimation that is based on a first privatized dataset, the sensitivity estimation indicating whether to modify the respective first noise profile, the second noise profile being further based at least on the sensitivity estimation.

38. The method of claim 32 , wherein the sensitivity estimation indicates to reduce privacy for outlier data of the second privatized dataset while at least maintaining privacy for non-outlier data of the second privatized dataset.

39. The method of claim 32 , further comprising causing transmission of the anomaly estimation for the second privatized dataset.

40. The method of claim 32 , wherein a first privatized dataset is a first privatized histogram of the first privatized dataset and the second privatized dataset is a second privatized histogram of the dataset.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 27, 2023
From: MOHAMMADY, MEISAM; ZHANG, MENGYUAN; JARRAYA, YOSR; POURZANDI, MAKAN; WANG, HAN; HONG, YUAN; WANG, LINGYU; MAJUMDAR, SURYADIPTA; DEBBABI, MOURAD
To: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
Reel/Frame 065368/0217 →
Continuity (2)
Provisional Application 63053231 · Jul 17, 2020
Related Publication 20230273994A1 · Aug 31, 2023
References Cited (55)
US 20100060350A1 · Zhang · 2010 [cited by examiner]
US 20110103673A1 · Rosenstengel · 2011 [cited by examiner]
US 20170316346A1 · Park · 2017 [cited by examiner]
US 20170353855A1 · Joy · 2017 [cited by applicant]
US 20180173894A1 · Boehler · 2018 [cited by examiner]
US 20180181878A1 · Kasiviswanathan · 2018 [cited by examiner]
US 20190026489A1 · Nerurkar · 2019 [cited by examiner]
US 20190372859A1 · Mermoud · 2019 [cited by examiner]
US 20200401726A1 · Lim · 2020 [cited by examiner]
US 20220138316A1 · Chen · 2022 [cited by examiner]
EP 3340101A1 · 2018 [cited by applicant]
WO 2017187207A1 · 2017 [cited by applicant]
Mohammady et al., DPOAD: Differentially Private Outsourcing of Anomaly Detection through Iterative Sensitivity Learning, arXiv: 2206.13046, pp. 1-13 (Jun. 2022) (Year: 2022). [cited by examiner]
Meisam Mohammady, Novel Approaches to Preserving Utility in Privacy Enhancing Technologies, Diss. Concordia University, pp. 114-152 (Sep. 2020) (Year: 2020). [cited by examiner]
International Search Report and Written Opinion dated Oct. 18, 2021 issued in PCT Application No. PCT/IB2021/056462 filed Jul. 16, 2021, consisting of 18 pages. [cited by applicant]
Chowdhury et al., Cryptϵ: Crypto-Assisted Differential Privacy on Untrusted Servers; Cornell University Library; Feb. 20, 2019, consisting of 24 pages. [cited by applicant]
Mohan et al., GUPT: Privacy Preserving Data Analysis Made Easy; Human Factors in Computing Systems, ACM Press/Addison-Wesley Publishing Co.; Scottsdale, AZ, May 20-24, 2012, consisting of 12 pages. [cited by applicant]
Okada et al., Differentially Private Analysis of Outliers; Graduate School of SIE, University of Tsukuba; Advances in Intelligent Data Analysis XIX; Springer International Publishing, Aug. 29, 2015, consisting of 22 pag… [cited by applicant]
Rattanavipanon et al., Releasing ARP Data with Differential Privacy Guarantees for LAN Anomaly Detection; 2021 18th International Conference on Electrical Engineering/Electronics, Computer, Telecommunications and Inform… [cited by applicant]
Ukil et al., Privacy for IoT: Involuntary Privacy Enablement for Smart Energy Systems; 2015 IEEE International Conference on Communications (ICC), Jun. 8, 2015, consisting of 6 pages. [cited by applicant]
Bhaduri et al., Privacy-Preserving Outlier Detection Through Random Nonlinear Data Distortion, IEEE Transactions on Systems, MAN, and Cybernetics—Part B: Cybernetics; 2010, consisting of 13 pages. [cited by applicant]
Boggs et al., Cross-domain Collaborative Anomaly Detection: So Far Yet So Close, ResearchGate; Conference Paper—Sep. 2011, consisting of 21 pages. [cited by applicant]
Burkhart et al., The Role of Network Trace Anonymization Under Attack; ACM SIGCOMM Computer Communication Review; vol. 40, No. 1, Jan. 2010, consisting of 6 pages. [cited by applicant]
Burkhart et al., SEPIA: Privacy-Preserving Aggregation of Multi-Domain Network Events and Statistics; Proceedings of the 19th USENIX Conference on Security (Berkeley, CA, 2010), USENIX Security '10, USENIX Association, … [cited by applicant]
Cheetancheri et al., A Distributed Host-based Worm Detection System; SIGCOMM '06 Workshops Sep. 11-15, 2006, Pisa, Italy; Copyright 2006 ACM, consisting of 7 pages. [cited by applicant]
Chen et al., Collaborative Detection of DDOS Attacks over Multiple Network Domains; IEEE Transactions on Parallel and Distributed Systems TPDS-0228-0806, Jan. 2008, consisting of 15 pages. [cited by applicant]
The Dung et al., A Distributed Solution for Privacy Preserving Outlier Detection; IEEE Computer Society; 2011 Third International Conference on Knowledge and Systems Engineering, consisting of 6 pages. [cited by applicant]
Dwork, Differential Privacy: A Survey of Results, Microsoft Research; Springer-Verlag Berlin, Heidelberg 2008, consisting of 19 pages. [cited by applicant]
Dwork et al., Calibrating Noise to Sensitivity in Private Data Analysis; Microsoft Research, Silicon Valley; Theory of Cryptography (Berlin, Heidelberg, 2006), Springer Berlin Heidelberg, consisting of 20 pages. [cited by applicant]
Erfani et al., Privacy-Preserving Collaborative Anomaly Detection for Participatory Sensing; Advances in Knowledge Discovery and Data Mining; Cham, Switzerland, 2014, Springer International Publishing, consisting of 13 … [cited by applicant]
Hong et al., Collaborative Search Log Sanitization: Toward Differential Privacy and Boosted Utility; IEEE Transactions on Dependable and Secure Computing, vol. 12, No. 5, Sep./Oct. 2015, consisting of 15 pages. [cited by applicant]
Jung et al., Fast Portscan Detection Using Sequential Hypothesis Testing; IEEE Symposium on Security and Privacy, 2004; Proceedings May 2004, consisting of 15 pages. [cited by applicant]
King et al., A Taxonomy and Adversarial Model for Attacks against Network Log Anonymization; Proceedings of the 2009 ACM Symposium on Applied Computing; New York, NY 2009, SAC '09, ACM, consisting of 8 pages. [cited by applicant]
Machiraju et al., Verifying Global Invariants in Multi-Provider Distributed Systems; Proc. 3rd Workshop on Hot Topics in Networks (HotNets-III), New York, NY: The Association for Computing Machinery, Inc., Nov. 2004 con… [cited by applicant]
McSherry et al., Differentially-Private Network Trace Analysis; Microsoft Research; SIGCOMM '10, Aug. 30-Sep. 3, 2010, New Delhi, India, consisting of 12 pages. [cited by applicant]
McSherry, Privacy Integrated Queries: An Extensible Platform for Privacy-Preserving Data Analysis; Communications of the ACM; ACM SIGMOD International Conference on Management of Data, New York, NY, Sep. 2010, vol. 53 N… [cited by applicant]
Pang et al., The Devil and Packet Trace Anonymization; ACM SIGCOMM Computer Communication Review; vol. 36, No. 1, Jan. 2006, consisting of 10 pages. [cited by applicant]
Parekh et al., Privacy-Preserving Payload-Based Correlation for Accurate Malicious Traffic Detection; SIGCOMM '06 Workshops Sep. 11-15, 2006, Pisa, Italy, consisting of 8 pages. [cited by applicant]
Riboni et al., Obfuscation of Sensitive Data in Network Flows; Proceedings IEEE INFOCOM, Mar. 2012, consisting of 9 pages. [cited by applicant]
Simmross-Wattenberg et al., Anomaly Detection in Network Traffic Based on Statistical Inference and Alpha-Stable Modeling; ResearchGate; IEEE Transactions on Dependable and Secure Computing, vol. 8, No. 4, Jul./Aug. 201… [cited by applicant]
Simpson et al., Neti@home: A Distributed Approach to Collecting End-to-End Network Performance Measurements; Passive and Active Network Measurement (Berlin, Heidelberg, 2004), consisting of 7 pages. [cited by applicant]
Bin Tariq et al., Detecting Network Neutrality Violations with Causal Inference; CoNEXT '09, Dec. 1-4, 2009, Rome, Italy, consisting of 12 pages. [cited by applicant]
Terrovitis et al., Privacy-preserving Anonymization of Set-valued Data; PVLDB '08, Aug. 23-28, 2008, Auckland, New Zealand, consisting of 11 pages. [cited by applicant]
Yegneswaran et al., Global Intrusion Detection in the DOMINO Overlay System; Proceedings of Network and Distributed System Security Symposium (NDSS (2004), consisting of 17 pages. [cited by applicant]
Goldstein et al., Histogram-based Outlier Score (HBOS): A fast Unsupervised Anomaly Detection Algorithm; Poster and Demo Track of the 35th German Conference on Artificial Intelligence (KI-2012), Sep. 24-27, 2012, consis… [cited by applicant]
Asif et al., How to Accurately and Privately Identify Anomalies; CCS '19, London, United Kingdom; Association for Computing Machinery, Nov. 11-15, 2019, consisting of 18 pages. [cited by applicant]
Bittner et al., Using Noisy Binary Search for Differentially Private Anomaly Detection; International Symposium on Cyber Security Cryptography and Machine Learning; Springer International Publishing AG, 2018, consisting… [cited by applicant]
Asif et al., Collaborative Differentially Private Outlier Detection for Categorical Data; 2016 IEEE 2nd International Conference on Collaboration and Internet Computing, consisting of 10 pages. [cited by applicant]
Asif et al., Differentially Private Outlier Detection in a Collaborative Environment; International Journal of Cooperative Information Systems 27, No. 03; Sep. 2018, consisting of 44 pages. [cited by applicant]
Bohler et al., Privacy-Preserving Outlier Detection for Data Streams; Data and Applications Security and Privacy; Lecture Notes in Computer Science, vol. 10359; Springer, Cham, Switzerland; Jan. 15, 2018, consisting of … [cited by applicant]
Okada et al., Differentially Private Analysis of Outliers; Machine Learning and Knowledge Discovery in Databases; .Lecture Notes in Computer Science, vol. 9285. Springer International Publishing, Cham, Switzerland; 2015… [cited by applicant]
Ding et al., Outsourcing Internet Security: Economic Analysis of Incentives for Managed Security Service Providers; International Workshop on Internet and Network Economics; Springer, Berlin, Heidelberg, 2005, consistin… [cited by applicant]
Thorp, Portfolio Choice and the Kelly Criterion; Stochastic optimization models in Finance; Business and Economics Statistics Section of Proceedings of the American Statistical Association; Aug. 13, 2016, consisting of … [cited by applicant]
Rubinstein et al., Pain-Free Random Differential Privacy with Sensitivity Sampling; Proceedings of the 34th International Conference on Machine Learning, Sydney, Australia, Jun. 8, 2017, consisting of 12 pages. [cited by applicant]
Goldstein et al., Histogram-based Outlier Score (HBOS): A fast Unsupervised Anomaly Detection Algorithm; German Research Center for Artificial Intelligence (DFKI) Kaiserslautern; 2012, consisting of 1 page. [cited by applicant]