IP Library › Granted Patent US 12,493,826
Granted Patent B2
US 12,493,826 · App. 17/956,120 · Granted Dec 9, 2025

Automatic machine learning feature backward stripping

Inventor: Jacques Doan Huu (Montigny le Bretonneux, FR)
Assignee: SAP SE
G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,493,826
App. No.
17/956,120
Granted
Dec 9, 2025
Kind
B2
Abstract

Features are used to train one or more ML models in a modelling layer. In a feature selection layer, each generated ML model is analyzed to determine, for each input feature, a degree of importance of the feature on the results generated by the ML model. Features with low importance are identified and the information is propagated backward to the data source and feature engineering layers. In response, the data source and feature engineering layers refrain from gathering or generating the unimportant features. Based on a confidence measure of the determination that each feature is important or unimportant, a number of periods between reevaluation of the feature importance is determined. After the number of periods has elapsed, a removed feature is restored to the pipeline.

Claims (40)

1 . A method comprising:

training, by one or more processors, a first set of machine learning (ML) models based on a current subset of a set of features;

determining, by the one or more processors, that a feature of the current subset of features has low importance;

based on the low importance of the feature, modifying the current subset of the set of features by removing the feature;

determining, for the feature, a confidence measure in the low importance of the feature;

based on the confidence measure, determining a number of iterations to delay reintroduction of the feature to the current subset of the set of features; and

training a second set of ML models based on the modified current subset of the set of features.

2 . The method of claim 1 , wherein the determining of the confidence measure is based on a fraction of the first set of ML models in which the feature was important.

3 . The method of claim 1 , wherein the number of iterations to delay reintroduction of the feature is larger when the confidence measure is larger.

4 . The method of claim 1 , wherein the training of the first set of ML models comprises training a set of random forest models.

5 . The method of claim 1 , wherein the training of the first set of ML models comprises training a set of gradient boosting decision trees.

6 . The method of claim 1 , wherein the first set of ML models are generated at a predetermined rate.

7 . The method of claim 1 , wherein the set of features comprises raw data and engineered data.

8 . A system comprising:

a memory that stores instructions; and

one or more processors configured by the instructions to perform operations comprising:

training a first set of machine learning (ML) models based on a current subset of a set of features;

determining that a feature of the current subset of features has low importance;

based on the low importance of the feature, modifying the current subset of the set of features by removing the feature;

determining, for the feature, a confidence measure in the low importance of the feature;

based on the confidence measure, determining a number of iterations to delay reintroduction of the feature to the current subset of the set of features; and

training a second set of ML models based on the modified current subset of the set of features.

9 . The system of claim 8 , wherein the determining of the confidence measure is based on a fraction of the first set of ML models in which the feature was important.

10 . The system of claim 8 , wherein the number of iterations to delay reintroduction of the feature is larger when the confidence measure is larger.

11 . The system of claim 8 , wherein the training of the first set of ML models comprises training a set of random forest models.

12 . The system of claim 8 , wherein the training of the first set of ML models comprises training a set of gradient boosting decision trees.

13 . The system of claim 8 , wherein the first set of ML models are generated at a predetermined rate.

14 . The system of claim 8 , wherein the set of features comprises raw data and engineered data.

15 . A non-transitory machine-readable medium that stores instructions that, when executed by one or more processors, cause the one or more processors to perform operations comprising:

training a first set of machine learning (ML) models based on a current subset of a set of features;

determining that a feature of the current subset of features has low importance;

based on the low importance of the feature, modifying the current subset of the set of features by removing the feature;

determining, for the feature, a confidence measure in the low importance of the feature;

based on the confidence measure, determining a number of iterations to delay reintroduction of the feature to the current subset of the set of features; and

training a second set of ML models based on the modified current subset of the set of features.

16 . The non-transitory machine-readable medium of claim 15 , wherein the determining of the confidence measure is based on a fraction of the first set of ML models in which the feature was important.

17 . The non-transitory machine-readable medium of claim 15 , wherein the number of iterations to delay reintroduction of the feature is larger when the confidence measure is larger.

18 . The non-transitory machine-readable medium of claim 15 , wherein the training of the first set of ML models comprises training a set of random forest models.

19 . The non-transitory machine-readable medium of claim 15 , wherein the training of the first set of ML models comprises training a set of gradient boosting decision trees.

20 . The non-transitory machine-readable medium of claim 15 , wherein the first set of ML models are generated at a predetermined rate.

Assignments (2)
CORRECTIVE ASSIGNMENT TO CORRECT THE THE INVENTOR'S NAME PREVIOUSLY RECORDED AT REEL: 061300 FRAME: 0654. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded Dec 13, 2022
From: DOAN HUU, JACQUES
To: SAP SE
Reel/Frame 062117/0181 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Oct 4, 2022
From: HUU, JACQUES DOAN
To: SAP SE
Reel/Frame 061300/0654 →
Continuity (2)
Continuation 16868145 · May 6, 2020
Related Publication 20230026391A1 · Jan 26, 2023
References Cited (10)
US 11062400B1 · Mccall et al. · 2021 [cited by applicant]
US 11315030B2 · Cataltepe · 2022 [cited by examiner]
US 11599826B2 · Khurana · 2023 [cited by examiner]
US 20210319354A1 · Raz · 2021 [cited by examiner]
US 20210319363A1 · Gillberg et al. · 2021 [cited by applicant]
US 20210342949A1 · Kim et al. · 2021 [cited by applicant]
US 20210350273A1 · Huu · 2021 [cited by applicant]
Loyola, et al, Learning Feature Representations from Change Dependency Graphs for Defect Prediction, Retrieved from Internet:<https://ieeexplore.ieee.org/abstract/document/8109101> (Year: 2017). [cited by examiner]
Singh, et al, Feature Selection Effects on Classification Algorithms, Retrieved from Internet: <chrome-extension://efaidnbmnnnibpcajpcglclefindmkaj/https://www.ijert.org/research/feature-selection-effects-on-classificat… [cited by examiner]
“U.S. Appl. No. 16/868,145, Notice of Allowance mailed Jun. 29, 2022”, 10 pgs. [cited by applicant]