IP Library › Granted Patent US 12,417,410
Granted Patent B2
US 12,417,410 · App. 17/369,524 · Granted Sep 16, 2025

Systems and methods for generating optimal data predictions in real-time for time series data signals

Inventor: Ibrahim F. Ghalyan (New York, NY)
Assignee: THE BANK OF NEW YORK MELLON
G06N20/20G06F18/211G06F18/217G06F18/2193G06F18/285G06F2123/02
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,417,410
App. No.
17/369,524
Filed
Jul 7, 2021
Granted
Sep 16, 2025
Kind
B2
Art Unit
2124
USPC
706/12
Abstract

Methods and systems are disclosed for generating optimal data predictions in time series data signals based on empirically-optimized model selection, noise filtering, and window size selection using machine learning models. For example, the system may receive a first subset of time series data. The system may receive a prediction horizon. The system may generate a feature input based on the first subset of time series data and the prediction horizon. The system may input the feature input into a machine learning model, wherein the machine learning model includes multiple components. The system may receive an output from the machine learning model. The system may generate for display, on a user interface, a prediction for the first subset of time series data at the prediction horizon based on the output.

Claims (51)

1. A system for generating optimal data predictions in real-time for time series data signals related to load processing in disparate computer networks based on empirically-optimized model selection, noise filtering, and window size selection using machine learning models, the system comprising:

cloud-based storage circuitry configured to store:

an ensemble of base models; and

a machine learning model, wherein the machine learning model includes:

a first model component, wherein the first model component is trained to select, for given time series data and given prediction horizons, an optimal base model collection for the ensemble of base models and select an optimal parameter set for the optimal base model collection, and wherein the optimal base model collection and optimal parameter set are selected by evaluating multiple candidate model structures and associated parameters to minimize prediction error and maximize a predictive performance metric;

a second model component, wherein the second model component is trained to select a filtering parameter for the given time series data and given prediction horizons; and

a third model component, wherein the third model component is trained to select an optimal window size for the given time series data and given prediction horizons;

cloud-based control circuitry configured to:

receive a first subset of time series data, wherein the first subset of times series data comprises values for load amounts for processing requests on computer nodes;

receive a prediction horizon, wherein the prediction horizon comprises a future date for predicting a load amount for processing requests;

generate a feature input based on the first subset of time series data and the prediction horizon, wherein the feature input comprises a vectorized representation of the first subset of time series data and the prediction horizon;

input the feature input into the machine learning model; and

receive an output from the machine learning model; and

cloud-based input/output circuitry configured to generate for display, on a user interface, a prediction for the first subset of time series data at the prediction horizon based on the output, wherein the prediction comprises a graphical representation of a predicted set of values for time series values and a confidence level for each of the predicted set of values for time series values.

2. A method for generating optimal data predictions in real-time for time series data signals based on empirically-optimized model selection, noise filtering, and window size selection using machine learning models, the method comprising:

receiving a first subset of time series data, wherein the first subset of times series data comprises values for load amounts for processing requests on computer nodes;

receiving a prediction horizon, wherein the prediction horizon comprises a future date for predicting a load amount for processing requests;

generating a feature input based on the first subset of time series data and the prediction horizon, wherein the feature input comprises a vectorized representation of the first subset of time series data and the prediction horizon;

inputting the feature input into a machine learning model, wherein the machine learning model includes:

a first model component, wherein the first model component is trained to select, for given time series data and given prediction horizons, an optimal base model collection for an ensemble of base models and select an optimal parameter set for the optimal base model collection, and wherein the optimal base model collection and optimal parameter set are selected by evaluating multiple candidate model structures and associated parameters to minimize prediction error and maximize a predictive performance metric;

a second model component, wherein the second model component is trained to select a filtering parameter for the given time series data and given prediction horizons; and

a third model component, wherein the third model component is trained to select an optimal window size for the given time series data and given prediction horizons;

receiving an output from the machine learning model; and

generating for display, on a user interface, a prediction for the first subset of time series data at the prediction horizon based on the output, wherein the prediction comprises a graphical representation of a predicted set of values for time series values and a confidence level for each of the predicted set of values for time series values.

3. The method of claim 2 , wherein the ensemble of base models comprises individual models and models based on combinations of the individual models.

4. The method of claim 2 , wherein the ensemble of base models comprises all permutations of individual models and models based on combinations of all the permutations of the individual models.

5. The method of claim 2 , wherein the second model component comprises an empirically-optimized Gaussian smoothing filter, wherein empirically-optimized the Gaussian smoothing filter comprises optimizing the filter parameter and the optimal parameter set using a single function.

6. The method of claim 2 , further comprising training the machine learning model using four-fold cross-validation.

7. The method of claim 2 , wherein variables for the first model component include values of the first subset of time series data, the prediction horizon, or the optimal window size.

8. The method of claim 2 , wherein variables for the second model component include values of the first subset of time series data, the prediction horizon, or the optimal window size.

9. The method of claim 2 , wherein variables for the third model component include values of the first subset of time series data, the prediction horizon, or the optimal parameter set.

10. The method of claim 2 , wherein the second model component is trained to filter both stationary and non-stationary noise from the given time series data.

11. The method of claim 2 , wherein the prediction includes a predicted set of values for time series values and a confidence level for each of the predicted set of values for time series values.

12. A non-transitory, computer readable medium for generating optimal data predictions in real-time for time series data signals based on empirically-optimized model selection, noise filtering, and window size selection using machine learning models comprising instructions that when executed by one or more processors, cause operations comprising:

receiving a first subset of time series data, wherein the first subset of times series data comprises values for load amounts for processing requests on computer nodes;

receiving a prediction horizon, wherein the prediction horizon comprises a future date for predicting a load amount for processing requests;

generating a feature input based on the first subset of time series data and the prediction horizon, wherein the feature input comprises a vectorized representation of the first subset of time series data and the prediction horizon;

inputting the feature input into a machine learning model, wherein the machine learning model includes:

a first model component, wherein the first model component is trained to select, for given time series data and given prediction horizons, an optimal base model collection for an ensemble of base models and select an optimal parameter set for the optimal base model collection, and wherein the optimal base model collection and optimal parameter set are selected by evaluating multiple candidate model structures and associated parameters to minimize prediction error and maximize a predictive performance metric;

a second model component, wherein the second model component is trained to select a filtering parameter for the given time series data and given prediction horizons; and

a third model component, wherein the third model component is trained to select an optimal window size for the given time series data and given prediction horizons;

receiving an output from the machine learning model; and

generating for display, on a user interface, a prediction for the first subset of time series data at the prediction horizon based on the output, wherein the prediction comprises a graphical representation of a predicted set of values for time series values and a confidence level for each of the predicted set of values for time series values.

13. The non-transitory, computer readable medium of claim 12 , wherein the ensemble of base models comprises individual models and models based on combinations of the individual models.

14. The non-transitory, computer readable medium of claim 12 , wherein the ensemble of base models comprises all permutations of individual models and models based on combinations of all the permutations of the individual models.

15. The non-transitory, computer readable medium of claim 12 , wherein the second model component comprises an empirically-optimized Gaussian smoothing filter, wherein empirically-optimized the Gaussian smoothing filter comprises optimizing the filtering parameter and the optimal parameter set using a single function.

16. The non-transitory, computer readable medium of claim 12 , further comprising training the machine learning model using four-fold cross-validation.

17. The non-transitory, computer readable medium of claim 12 , wherein variables for the first model component include values of the first subset of time series data, the prediction horizon, or the optimal window size.

18. The non-transitory, computer readable medium of claim 12 , wherein variables for the second model component include values of the first subset of time series data, the prediction horizon, or the optimal window size.

19. The non-transitory, computer readable medium of claim 12 , wherein variables for the third model component include values of the first subset of time series data, the prediction horizon, or the optimal parameter set.

20. The non-transitory, computer readable medium of claim 12 , wherein the second model component is trained to filter both stationary and non-stationary noise from the given time series data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 7, 2021
From: GHALYAN, IBRAHIM F.
To: THE BANK OF NEW YORK MELLON
Reel/Frame 056779/0851 →
Continuity (1)
Related Publication 20230012177A1 · Jan 12, 2023
References Cited (29)
US 11556789B2 · Gugulothu · 2023 [cited by examiner]
US 20020183931A1 · Anno · 2002 [cited by applicant]
US 20190265971A1 · Behzadi et al. · 2019 [cited by applicant]
US 20200097810A1 · Hetherington et al. · 2020 [cited by applicant]
US 20200242483A1 · Shashikant Rao · 2020 [cited by examiner]
US 20210081492A1 · Higginson · 2021 [cited by examiner]
US 20220172038A1 · Chen · 2022 [cited by examiner]
US 20220180207A1 · Liang · 2022 [cited by examiner]
US 20230018125A1 · Lim · 2023 [cited by examiner]
US 20240023884A1 · Liu · 2024 [cited by examiner]
WO WO2021126500A1 · 2021 [cited by examiner]
WO WO2022257187A1 · 2022 [cited by examiner]
Ghalyan et al., “Gaussian Smoothing Filter for Improved EMG Signal Modeling” Mar. 17, 2020, Chapter 6 Signal Processing in Medicine and Biology, pp. 161-204. (Year: 2020). [cited by examiner]
Wang et al., “ConceptExplorer: Visual Analysis of Concept Drifts in Multi-source Time-series Data” Jan. 4, 2021, pp. 1-11. (Year: 2021). [cited by examiner]
Fawaz et al., “Deep Neural Network Ensembles for Time Series Classification” Apr. 26, 2019, arXiv: 1903.066602v2, pp. 1-6. (Year: 2019). [cited by examiner]
Javeri et al., “Improving Neural Networks for Time-Series Forecasting using Data Augmentation and AutoML” May 7, 2021, arXiv: 2103.01992v3, pp. 1-8. (Year: 2021). [cited by examiner]
Middlehurst et al., “HIVE-COTE 2.0: a new meta ensemble for time series classification” Apr. 15, 2021, arXiv: 2104.07551v1, pp. 1-35. (Year: 2021). [cited by examiner]
Shi et al., “For2For: Learning to forecast from forecasts” Jan. 14, 2020, arXiv: 2001.04601v1, pp. 1-13. (Year: 2020). [cited by examiner]
Elsayed et al., “Do We Really Need Deep Learning Models for Time Series Forecasting?” Jan. 6, 2021, arXiv: 2101.02118v1, pp. 1-14. (Year: 2021). [cited by examiner]
International Search Report & Written Opinion of the International Searching Authority dated Nov. 3, 2022, issued in corresponding International Application No. PCT/US2022/035252 (15 pgs.). [cited by applicant]
International Preliminary Report on Patentability dated Jan. 18, 2024, issued in corresponding International Patent Application No. PCT/US2022/035252 (8 pgs.). [cited by applicant]
Ibrahim Fahad Jasim Ghalyan, “Force-Controlled Robotic Assembly Processes of Rigid and Flexible Objects: Methodologies and Applications”, Springer, Cham. Switzerland, 2016, 17 pgs. [cited by applicant]
Elhanan Elboher et al., “Efficient and Accurate Gaussian Image Filtering Using Running Sums ,” Computing Research Repository, vol. abs/1107.4958, 2011. http://arxiv.org/abs/1107.4958. [cited by applicant]
I.F. Ghalyan et al., “Gaussian Filtering of EMG Signals for Improved Hand Gesture Classification”, Proc. 2018 IEEE Signal Processing in Medicine and Biology Symposium (SPMB), Philadelphia, PA, 2018, pp. 1-6, doi: 10.110… [cited by applicant]
Ibrahim F.J. Ghalyan et al., “Gaussian Smoothing Filter for Improved EMG Signal Modeling ,” Proc. 2018 IEEE Signal Processing in Medicine and Biology Symposium (SPMB), Philadelphia, PA, 2018, pp. 1-6, doi: 10.1109/SPMB.… [cited by applicant]
Pascal Gwosdek et al., “Theoretical Foundations of Gaussian Convolution by Extended Box Filtering”, International Conference on Scale Space and Variational Methods in Computer Vision, pp. 447-458, 2011. http://dx.doi.or… [cited by applicant]
Richard A. Haddad et al., “A Class of Fast Gaussian Binomial Filters for Speech and Image Processing”, IEEE Transactions on Signal Processing, vol. 39, No. 3, Mar. 1991, pp. 723-727. [cited by applicant]
Mark Nixon et al., “Feature Extraction and Image Processing” Academic Press, 2008, pp. 1-100. [cited by applicant]
Linda Shapiro et al., “Computer Vision” (1st ed.). Prentice Hall PTR, Upper Saddle River, NJ, USA, 2001, pp. 1-609. [cited by applicant]