IP Library Granted Patent US 12,293,302
Granted Patent B2
US 12,293,302 · App. 17/000,503 · Granted May 6, 2025

Multidimensional hierarchy level recommendation for forecasting models

Inventor: Mokrane Amzal (Courbevoie, FR)
Assignee: SAP SE
G06N5/04G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,293,302
App. No.
17/000,503
Granted
May 6, 2025
Kind
B2
Abstract

Provided is a system and method for identifying and recommending the best hierarchical levels for training a predictive model such as a time-series forecasting model. In one example, the method may include receiving an identification of a measure of multidimensional data, generating a plurality of training data sets that comprise different combinations of hierarchical dimension granularities of aggregation, training a plurality of instances of a machine learning model based on the plurality of training data sets, respectively, and determining and outputting predictive accuracy values of the plurality of instances of the trained machine learning model.

Claims (46)

1. A computing system comprising:

a processor configured to:

receive an identification of a measure of multidimensional data;

extract metadata from the multidimensional data, the extracted metadata identifying a dimension of the measure and a plurality of different granularities applicable to the dimension;

generate, based on the extracted metadata, a plurality of multidimensional queries that comprise a plurality of different granularities of aggregation for the measure identified in the extracted metadata;

query the multidimensional data via the plurality of multidimensional queries to generate a plurality of training data sets;

train a plurality of different instances of a machine learning model via execution of a machine learning algorithm on the plurality of training data sets corresponding to the plurality of different granularities of aggregation, respectively;

determine predictive accuracy values of the plurality of different instances of the machine learning model corresponding to the plurality of different granularities of aggregation;

generate a ranked ordering of the plurality of different instances of the machine learning model based on the predictive accuracy values; and

display, in accordance with the ranked ordering, the predictive accuracy values via a user interface.

2. The computing system of claim 1 , wherein the multidimensional data comprises a time dimension.

3. The computing system of claim 1 , wherein the processor is further configured to identify restrictions on feature generation from the metadata, and generate the plurality of training data sets based on the identified restrictions.

4. The computing system of claim 1 , wherein the processor is configured to generate at least two training data sets from a same cube of the multidimensional data using a different aggregation granularity for time, respectively.

5. The computing system of claim 1 , wherein the processor is configured to generate at least two training data sets from a same cube of the multidimensional data using a different aggregation granularity for the measure, respectively.

6. The computing system of claim 1 , wherein the processor is configured to, for a training data set, execute the machine learning algorithm on the respective training data set to generate a corresponding instance among the plurality of different instances of the machine learning model.

7. The computing system of claim 1 , wherein the processor is configured to determine a predictive accuracy of an instance of the machine learning model based on a comparison of a forecasted output of the instance of the machine learning model with an actual output.

8. The computing system of claim 1 , wherein the processor is further configured to arrange identifiers of the plurality of different instances of the machine learning model from most accurate to least accurate based on the predictive accuracy values, and display the arranged identifiers via the user interface.

9. A method comprising:

receiving an identification of a measure of multidimensional data;

extracting metadata from the multidimensional data, the extracted metadata identifying a dimension of the measure and a plurality of different granularities applicable to the dimension;

generating, based on the extracted metadata, a plurality of multidimensional data sets that comprise a plurality of different granularities of aggregation for the measure identified in the extracted metadata;

querying the multidimensional data via the plurality of multidimensional queries to generate a plurality of training data sets;

training a plurality of different instances of a machine learning model via execution of a machine learning algorithm on the plurality of training data sets corresponding to the plurality of different granularities of aggregation, respectively;

determining predictive accuracy values of the plurality of different instances of the machine learning model corresponding to the plurality of different granularities of aggregation;

generating a ranked ordering of the plurality of different instances of the machine learning model based on the predictive accuracy values; and

displaying, in accordance with the ranked ordering, the predictive accuracy values via a user interface.

10. The method of claim 9 , wherein the multidimensional data comprises a time dimension.

11. The method of claim 9 , wherein the method further comprises identifying restrictions on feature generation from the metadata, and generating the plurality of training data sets based on the identified restrictions.

12. The method of claim 9 , wherein the generating comprises generating at least two training data sets from a same cube of the multidimensional data using a different aggregation granularity for time, respectively.

13. The method of claim 9 , wherein the generating comprises generating at least two training data sets from a same cube of the multidimensional data using a different aggregation granularity for the measure, respectively.

14. The method of claim 9 , wherein the training comprises, for a training data set, executing the machine learning algorithm on the respective training data set to generate a corresponding instance among the plurality of different instances of the machine learning model.

15. The method of claim 9 , wherein the determining comprises determining a predictive accuracy of an instance of the machine learning model based on a comparison of a forecasted output of the instance of the machine learning model with actual output.

16. The method of claim 9 , wherein the outputting comprises arranging identifiers of the plurality of different instances of the machine learning model from most accurate to least accurate based on the predictive accuracy values, and displaying the arranged identifiers via the user interface.

17. A non-transitory computer-readable medium comprising instructions which when read by a processor cause a computer to perform a method comprising:

receiving an identification of a measure of multidimensional data;

extracting metadata from the multidimensional data, the extracted metadata identifying a dimension of the measure and a plurality of different granularities applicable to the dimension;

generating, based on the extracted metadata, a plurality of multidimensional queries that comprise a plurality of different granularities of aggregation for the measure identified in the extracted metadata;

querying the multidimensional data via the plurality of multidimensional queries to generate a plurality of training data sets;

training a plurality of different instances of a machine learning model via execution of a machine learning algorithm on the plurality of training data sets corresponding to the plurality of different granularities of aggregation, respectively;

determining predictive accuracy values of the plurality of different instances of the machine learning model corresponding to the plurality of different granularities of aggregation;

generating a ranked ordering of the plurality of different instances of the machine learning model based on the predictive accuracy values; and

displaying, in accordance with the ranked ordering, the predictive accuracy values via a user interface.

18. The non-transitory computer-readable medium of claim 17 , wherein the multidimensional data comprises a time dimension.

19. The non-transitory computer-readable medium of claim 17 ,

wherein the method further comprises identifying restrictions on feature generation from the metadata, and generating the plurality of training data sets based on the identified restrictions.

20. The non-transitory computer-readable medium of claim 17 , wherein the generating comprises generating at least two training data sets from a same cube of the multidimensional data using a different aggregation granularity for time, respectively.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Aug 24, 2020
From: AMZAL, MOKRANE
To: SAP SE
Reel/Frame 053572/0680 →
Continuity (1)
Related Publication 20220058499A1 · Feb 24, 2022
References Cited (9)
US 9530228B1 · Fermum · 2016 [cited by examiner]
US 20040181519A1 · Anwar · 2004 [cited by examiner]
US 20090228436A1 · Pasumansky · 2009 [cited by examiner]
US 20180081953A1 · Goedken · 2018 [cited by examiner]
Lusis et al., Short-term residential load forecasting: Impact of calendar effects and forecast granularity, Applied Energy 205 (2017), 654-669 (Year: 2017). [cited by examiner]
Perez et al., Hierarchical Aggregation Prediction Model, JMLRL Workshop and Conference Proceedings 11, KDD CUP (2010), 1-14 (Year: 2010). [cited by examiner]
Jo et al., Improving Protein Fold Recognition by Deep Learning Networks, Scientific Reports (2015), 1-11 (Year: 2015). [cited by examiner]
Bauer et al., An Automated Forecasting Framework based on Method Recommendation for Seasonal Time Series, ICPE (2020), 48-55 (Year: 2020). [cited by examiner]
Yang et al., Machine Learning Differential Privacy With Multifunctional Aggregation in a Fog Computing Architecture, Special Section on Cyber-Physical-Social Computing and Networking (2018), 17119-17129 (Year: 2018). [cited by examiner]