IP Library Granted Patent US 12,299,084
Granted Patent B2
US 12,299,084 · App. 17/153,852 · Granted May 13, 2025

Artificial intelligence optimization platform

Inventors: Chaitra Kallianpur (Westborough, MA); Kalapriya Kannan (Bangalore, IN); Suparna Bhattacharya (Bangalore, IN)
Assignee: Hewlett Packard Enterprise Development LP
G06F18/285G06F18/214G06F18/22G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,299,084
App. No.
17/153,852
Granted
May 13, 2025
Kind
B2
Abstract

Systems and methods are provided for reusing machine learning models. For example, the applicability of prior models may be compared using one or more assessment values, including a similarity threshold and/or an accuracy threshold. The similarity threshold may identify a similarity of data between a first data set used to generate a first model and a new data set that is received by the system. When the similarity between these two data sets is exceeded, the system may reuse a model with the highest similarity value. When an accuracy value of the data set does not exceed an accuracy threshold, the system may initiate a retraining process to generate a second ML model associated with the second data.

Claims (67)

1. A computing platform comprising:

a set of edge devices comprising a first edge device and a second edge device that generate time series data sets in near real time,

wherein the first edge device is configured to:

process a first processing workload in a first environment where the first edge device is located, and

generate a first time series data set associated with the first processing workload;

wherein the second edge device is configured to:

process a second processing workload in a second environment where the second edge device is located, and

generate a second time series data set associated with the second processing workload; and

a central computing system comprising:

a memory; and

one or more processors configured to access the memory and execute machine readable instructions stored to:

receive a first machine learning (ML) model, the first ML model being trained using the first time series data set;

receive the second time series data set, the second time series data set being generated by the second edge device;

determine that the second time series data set comprises a data characteristic;

compare the second time series data set with the first time series data set to generate a similarity value associated with the data characteristic;

when the similarity value exceeds a similarity threshold:

reuse the first ML model to process the second time series data set,

generate an output from the first ML model, and

measure an accuracy value based on the output;

compare the accuracy value to an accuracy threshold; and

when the accuracy value does not exceed the accuracy threshold or when the similarity value does not exceed the similarity threshold, initiate a retraining process to generate a second ML model associated with the second time series data set.

2. The computing platform of claim 1 , wherein the instructions further to:

when the accuracy value exceeds the accuracy threshold, run forecasting predictions with the first or second ML model.

3. The computing platform of claim 1 , wherein when the accuracy value does not exceed the accuracy threshold, the similarity threshold is adjusted for the first ML model.

4. The computing platform of claim 1 , wherein the similarity threshold and the accuracy threshold are different values.

5. The computing platform of claim 1 , wherein the accuracy value exceeds the accuracy threshold based on a drift-based determination.

6. The computing platform of claim 1 , wherein the instructions further to:

store the second ML model with the first ML model in a model data store.

7. The computing platform of claim 1 , wherein a determination to reuse the first ML model to process the second time series data set is executed in near real time.

8. The computing platform of claim 1 , wherein the first time series data set and the second time series data set are transmitted from the first edge device and the second edge device, respectively, as data matrices where rows represent time instances and columns represent a dimension being ingested.

9. A computer-implemented method comprising:

processing, by a first edge device of a computing platform, a first processing workload in a first environment where the first edge device is located;

generating, by the first edge device, a first time series data set associated with the first processing workload;

processing, by a second edge device of the computing platform, a second processing workload in a second environment where the second edge device is located;

generating, by the second edge device, a second time series data set associated with the second processing workload;

receiving, by a central computing system of a computing platform, a first machine learning (ML) model, the first ML model being trained using the first time series data set;

receiving, by the central computing system, the second time series data set, the second time series data set being generated by the second edge device;

determining, by the central computing system, that the second time series data set comprises a data characteristic;

comparing the second time series data set with the first time series data set to generate an assessment value associated with the data characteristic;

when the assessment value exceeds a similarity threshold:

reusing the first ML model to process the second time series data set,

generating an output from the first ML model, and

measuring an accuracy value based on the output;

comparing the accuracy value to an accuracy threshold; and

when the accuracy value does not exceed the accuracy threshold or when the assessment value does not exceed the similarity threshold, initiating a retraining process to generate a second ML model associated with the second time series data set.

10. The computer-implemented method of claim 9 , further comprising:

when the accuracy value exceeds the accuracy threshold, run forecasting predictions with the first or second ML model.

11. The computer-implemented method of claim 9 , wherein the first data set and the second data set are time series data sets.

12. The computer-implemented method of claim 9 , wherein the similarity threshold and the accuracy threshold are different values.

13. The computer-implemented method of claim 9 , wherein the first data set and the second data set originate from different sources.

14. The computer-implemented method of claim 9 , wherein the accuracy value exceeds the accuracy threshold based on a drift-based determination.

15. The computer-implemented method of claim 9 , further comprising:

storing the second ML model with the first ML model in a model data store.

16. A non-transitory computer-readable storage medium storing a plurality of instructions executable by one or more processors, the plurality of instructions when executed by the one or more processors cause the one or more processors to:

receive a second data set comprising a data characteristic, the second data set being generated by a second edge device or node that collects the second data set from a second processing workload from a second environment where the second edge device is located;

compare the second data set with a first data set to generate a similarity value associated with the data characteristic, the first data set being generated by a first edge device or node that collects the first data set from a first processing workload from a first environment where the first edge device is located, wherein the first data set was used to generate a first machine learning (ML) model;

when the similarity value exceeds a similarity threshold:

reuse the first ML model to process the second data set,

generate an output from the first ML model, and

measure an accuracy value based on the output;

compare the accuracy value to an accuracy threshold; and

when the accuracy value does not exceed the accuracy threshold or when the similarity value does not exceed the similarity threshold, initiate a retraining process to generate a second ML model associated with the second data.

17. The computer-readable storage medium of claim 16 , wherein the plurality of instructions further cause the one or more processors to:

when the accuracy value exceeds the accuracy threshold, run forecasting predictions with the first or second ML model.

18. The computer-readable storage medium of claim 16 , wherein the similarity threshold and the accuracy threshold are different values.

19. The computer-readable storage medium of claim 16 , wherein the first data set and the second data set originate from different sources.

20. The computer-readable storage medium of claim 16 , wherein the accuracy value exceeds the accuracy threshold based on a drift-based determination.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 20, 2021
From: KALLIANPUR, CHAITRA; KANNAN, KALAPRIYA; BHATTACHARYA, SUPARNA
To: HEWLETT PACKARD ENTERPRISE DEVELOPMENT LP
Reel/Frame 054974/0514 →
Continuity (1)
Related Publication 20220230024A1 · Jul 21, 2022
References Cited (28)
US 12052260B2 · Kozhaya · 2024 [cited by examiner]
US 20200364551A1 · Dwivedi · 2020 [cited by examiner]
US 20210110310A1 · Guim Bernat · 2021 [cited by examiner]
Anderson, Robert, et al. “Recurring concept meta-learning for evolving data streams.” Expert Systems with Applications 138 (2019): 112832. (Year: 2019). [cited by examiner]
Zhou, Ling, Tsuyoshi Tanaka, and Daisuke Tashiro. “Data Similarity Estimation for AI Model Reuse.” (2020) (Year: 2020). [cited by examiner]
Guo, Peizhen, et al. “Foggycache: Cross-device approximate computation reuse.” Proceedings of the 24th annual international conference on mobile computing and networking. 2018. (Year: 2018). [cited by examiner]
Mastorakis, Spyridon, et al. “Icedge: When edge computing meets information-centric networking.” IEEE Internet of Things Journal 7.5 (2020): 4203-4217. (Year: 2020). [cited by examiner]
Samarah, Samer, et al. “Transferring activity recognition models in FOG computing architecture.” Journal of Parallel and Distributed Computing 122 (2018): 122-130. (Year: 2018). [cited by examiner]
Bharde et al., “Store-Edge RippleStream: Versatile Infrastructure for IoT Data Transfer”, HotEdge, 2018, 6 pages. [cited by applicant]
Brownlee, Jason, “How to Check-Point Deep Learning Models in Keras”, Machine Learning Mastery Pty. Ltd., available online at <https://machinelearningmastery.com/check-point-deep-learning-models-keras/>, 2020, 44 pages. [cited by applicant]
Chen et al., “Deep Learning Training Times Get Significant Reduction”, IBM, available online at <ibm.com/blogs/research/2018/02/deep-learning-training/>, Feb. 5, 2018, 7 pages. [cited by applicant]
Das et al., “Multi-criteria online frame subset selection for autonomous vehicle videos”, Pattern Recognition Letters, vol. 133, 2020, pp. 349-355. [cited by applicant]
Data Science Stack Exchange, “Should a model be re-trained if new observations are available?”, available online at <https://datascience.stackexchange.com/questions/12761/should-a-model-be-re-trained-if-new-observations… [cited by applicant]
DataRobot, Inc., “DataRobot”, available online at <datarobot.com>, 2021, 8 pages. [cited by applicant]
Eck et al., “Scalable Deployment of AI Time-series Models for IoT”, 2020, 7 pages. [cited by applicant]
Garc'ia et al., “Early Drift Detection Method”, 2005, 10 pages. [cited by applicant]
Guo et al., “FoggyCache: Cross-Device Approximate Computation Reuse”, ACM, 2018, 16 pages. [cited by applicant]
Guo et al., “Potluck: Cross-Application Approximate Deduplication for Computation-Intensive Mobile Applications”, Association for Computing Machinery, 2018, 14 pages. [cited by applicant]
H2O AI, “H2O AI Hybrid Cloud”, available online at <https://www.h2o.ai/>, 2021, 8 pages. [cited by applicant]
Huang et al., “Scalable AutoML for Time Series Forecasting using Ray”, Intel, Jul. 9, 2020, 24 pages. [cited by applicant]
Johnson et al., “Billionscale similarity search with GPUs”, CoRR, 2017, 11 pages. [cited by applicant]
Kannan et al., “SEeSAW—Similarity Exploiting Storage for Accelerating Analytics Workflows”, HotStorage'16, 2016, 5 pages. [cited by applicant]
Luigi, “The Ultimate Guide to Model Retraining”, Deployment TagsMonitoring, Retraining, available online at <https://mlinproduction.com/model-retraining/>, Jun. 10, 2019, 9 pages. [cited by applicant]
Makkinejad et al., “Reduction in training time of a deep learning model in detection of lesions in CT”, SPIE Medical Imaging, 2018, 13 pages. [cited by applicant]
Miranda et al., “Reducing the Training Time of Neural Networks by Partitioning”, ICLR, 2016, 10 pages. [cited by applicant]
ScienceDaily, “New technique cuts AI training time by more than 60 percent”, available online at <https://www.sciencedaily.com/releases/2019/04/190408114322.htm>, Apr. 8, 2019, 4 pages. [cited by applicant]
Seif, George, “3 ways to improve your Machine Learningresults without more data”, Towards Data Science, available online at <https://towardsdatascience.com/3-ways-to-improve-your-machine-learning-results-without-more-da… [cited by applicant]
Wong et al., “Transfer Learning with Neural AutoML”, Jan. 28, 2019, 17 pages. [cited by applicant]