IP Library › Granted Patent US 12,481,727
Granted Patent B2
US 12,481,727 · App. 17/203,921 · Granted Nov 25, 2025

User acceptance test system for machine learning systems

Inventors: Atreya Biswas (Bangalore, IN); Denny Jee King Gee (Singapore, SG); Srivatsan Santhanam (Bangalore, IN)
Assignee: SAP SE
G06F18/285G06F18/2178G06N20/00G06V10/751
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,481,727
App. No.
17/203,921
Granted
Nov 25, 2025
Kind
B2
Abstract

Methods, systems, and computer-readable storage media for receiving, by a ML application executing within a cloud platform, a first inference request, the first inference request including first inference data, transmitting, by the ML application, the first inference data to the UAT system within the cloud platform, retrieving, by the UAT system, a first ML model in response to the inference request, the first ML model being in an inactive state, providing, by the UAT system, a first inference based on the first inference data using the first ML model, providing a first accuracy evaluation at least partially based on the first inference, and transitioning the first ML model from the inactive state to an active state, the first ML model being used for production in the active state.

Claims (63)

1 . A computer-implemented method for selectively deploying machine learning (ML) models to production using a user acceptance test (UAT) system, the method comprising:

receiving, by a ML application executing within a cloud platform, a first inference request, the first inference request comprising first inference data and an identifier that identifies one or more of an application, a system, a user, and an enterprise for which inferencing is to be executed;

transmitting, by the ML application, the first inference data to the UAT system within the cloud platform;

retrieving, by the UAT system, a first ML model from a set of ML models, the first ML model being associated with the identifier and being specific to the one or more of the application, the system, the user, and the enterprise, the first ML model being in an inactive state, in which the first ML model is not in production use to perform enterprise operations;

postponing execution of a passive inference using the first ML model and the first inference data until completion of an active inference using the first inference data, the postponing being in response to availability of computing resources for executing the passive inference and the active inference;

providing, as the passive inference and by the UAT system, a first inference based on the first inference data using the first ML model;

providing a first accuracy evaluation at least partially based on the first inference; and

transitioning the first ML model from the inactive state to an active state, the first ML model being used for production in the active state, in which the first ML model is in production use to perform enterprise operations.

2 . The method of claim 1 , further comprising:

generating, by the ML application, a second inference based on the first inference data using a second ML model in parallel with generating the first inference, the second ML model being in the active state; and

replacing the second ML model with the first ML model for subsequent production use in response to transitioning the first ML model to the active state.

3 . The method of claim 2 , wherein the first ML model is an updated version of the second ML model.

4 . The method of claim 1 , wherein the first accuracy evaluation comprises:

determining an accuracy of the first ML model that represents correct inferences of the first ML model; and

comparing the accuracy of the first ML model to a threshold accuracy.

5 . The method of claim 1 , wherein providing a first accuracy evaluation is executed in response to occurrence of a polling condition.

6 . The method of claim 1 , wherein the first inference data comprises production data.

7 . The method of claim 1 , further comprising:

retrieving, by the UAT system, a second ML model in response to a second inference request, the second ML model being in an inactive state;

providing, by the UAT system, a second inference based on second inference data of the second inference request using the second ML model;

determining a second accuracy evaluation at least partially based on the second inference; and

transmitting an alert regarding the second ML model in response to the second accuracy evaluation.

8 . A non-transitory computer-readable storage medium coupled to one or more processors and having instructions stored thereon which, when executed by the one or more processors, cause the one or more processors to perform operations for selectively deploying machine learning (ML) models to production using a user acceptance test (UAT) system, the operations comprising:

receiving, by a ML application executing within a cloud platform, a first inference request, the first inference request comprising first inference data and an identifier that identifies one or more of an application, a system, a user, and an enterprise for which inferencing is to be executed;

transmitting, by the ML application, the first inference data to the UAT system within the cloud platform;

retrieving, by the UAT system, a first ML model from a set of ML models, the first ML model being associated with the identifier and being specific to the one or more of the application, the system, the user, and the enterprise, the first ML model being in an inactive state, in which the first ML model is not in production use to perform enterprise operations;

postponing execution of a passive inference using the first ML model and the first inference data until completion of an active inference using the first inference data, the postponing being in response to availability of computing resources for executing the passive inference and the active inference;

providing, as the passive inference and by the UAT system, a first inference based on the first inference data using the first ML model;

providing a first accuracy evaluation at least partially based on the first inference; and

transitioning the first ML model from the inactive state to an active state, the first ML model being used for production in the active state, in which the first ML model is in production use to perform enterprise operations.

9 . The non-transitory computer-readable storage medium of claim 8 , wherein operations further comprise:

generating, by the ML application, a second inference based on the first inference data using a second ML model in parallel with generating the first inference, the second ML model being in the active state; and

replacing the second ML model with the first ML model for subsequent production use in response to transitioning the first ML model to the active state.

10 . The non-transitory computer-readable storage medium of claim 9 , wherein the first ML model is an updated version of the second ML model.

11 . The non-transitory computer-readable storage medium of claim 8 , wherein the first accuracy evaluation comprises:

determining an accuracy of the first ML model that represents correct inferences of the first ML model; and

comparing the accuracy of the first ML model to a threshold accuracy.

12 . The non-transitory computer-readable storage medium of claim 8 , wherein providing a first accuracy evaluation is executed in response to occurrence of a polling condition.

13 . The non-transitory computer-readable storage medium of claim 8 , wherein the first inference data comprises production data.

14 . The non-transitory computer-readable storage medium of claim 8 , wherein operations further comprise:

retrieving, by the UAT system, a second ML model in response to a second inference request, the second ML model being in an inactive state;

providing, by the UAT system, a second inference based on second inference data of the second inference request using the second ML model;

determining a second accuracy evaluation at least partially based on the second inference; and

transmitting an alert regarding the second ML model in response to the second accuracy evaluation.

15 . A system, comprising:

a computing device; and

a computer-readable storage device coupled to the computing device and having instructions stored thereon which, when executed by the computing device, cause the computing device to perform operations for selectively deploying machine learning (ML) models to production using a user acceptance test (UAT) system, the operations comprising:

receiving, by a ML application executing within a cloud platform, a first inference request, the first inference request comprising first inference data and an identifier that identifies one or more of an application, a system, a user, and an enterprise for which inferencing is to be executed;

transmitting, by the ML application, the first inference data to the UAT system within the cloud platform;

retrieving, by the UAT system, a first ML model from a set of ML models, the first ML model being associated with the identifier and being specific to the one or more of the application, the system, the user, and the enterprise, the first ML model being in an inactive state, in which the first ML model is not in production use to perform enterprise operations;

postponing execution of a passive inference using the first ML model and the first inference data until completion of an active inference using the first inference data, the postponing being in response to availability of computing resources for executing the passive inference and the active inference;

providing, as the passive inference and by the UAT system, a first inference based on the first inference data using the first ML model;

providing a first accuracy evaluation at least partially based on the first inference; and

transitioning the first ML model from the inactive state to an active state, the first ML model being used for production in the active state, in which the first ML model is in production use to perform enterprise operations.

16 . The system of claim 15 , wherein operations further comprise:

generating, by the ML application, a second inference based on the first inference data using a second ML model in parallel with generating the first inference, the second ML model being in the active state; and

replacing the second ML model with the first ML model for subsequent production use in response to transitioning the first ML model to the active state.

17 . The system of claim 16 , wherein the first ML model is an updated version of the second ML model.

18 . The system of claim 15 , wherein the first accuracy evaluation comprises:

determining an accuracy of the first ML model that represents correct inferences of the first ML model; and

comparing the accuracy of the first ML model to a threshold accuracy.

19 . The system of claim 15 , wherein providing a first accuracy evaluation is executed in response to occurrence of a polling condition.

20 . The system of claim 15 , wherein the first inference data comprises production data.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Mar 17, 2021
From: BISWAS, ATREYA; GEE, DENNY JEE KING; SANTHANAM, SRIVATSAN
To: SAP SE
Reel/Frame 055618/0188 →
Continuity (1)
Related Publication 20220300754A1 · Sep 22, 2022
References Cited (12)
US 10853693B2 · Eberlein et al. · 2020 [cited by applicant]
US 11341605B1 · Singh · 2022 [cited by examiner]
US 11853401B1 · Nookula · 2023 [cited by examiner]
US 20100242030A1 · Talbert · 2010 [cited by examiner]
US 20120036339A1 · Frazier · 2012 [cited by examiner]
US 20180349191A1 · Dorsey · 2018 [cited by examiner]
US 20190156247A1 · Faulhaber, Jr. · 2019 [cited by examiner]
US 20190294927A1 · Guttmann · 2019 [cited by examiner]
US 20190327259A1 · DeFelice · 2019 [cited by examiner]
US 20200401491A1 · Mohamed et al. · 2020 [cited by applicant]
Extended European Search Report in European Appln. No. 22159851.9, dated Aug. 23, 2022, 11 pages. [cited by applicant]
Office Action in European Appln. No. 22159851.9, mailed on Jan. 30, 2025, 5 pages. [cited by applicant]