IP Library Patent Application 17236625
Patent Application
App. No. 17/236,625

REINFORCEMENT LEARNING AND ACCUMULATORS MAP-BASED PIPELINE FOR WORKLOAD PLACEMENT

Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US None
App. No.
17/236,625
Abstract

One example method includes running multiple iterations of a computing workload, for each iteration of the computing workload, for each iteration of the computing workload, using a reinforcement learning process to generate an initial infrastructure allocation for the computing workload, and a reward function of the reinforcement learning process generates a respective reward for each initial infrastructure allocation, running an accumulator map voting process to generate a total reward for each initial infrastructure allocation, and identifying the initial infrastructure allocation with the largest total reward and assigning that initial infrastructure allocation to the computing workload.

Claims (28)

1 . A method, comprising:

running multiple iterations of a computing workload;

for each iteration of the computing workload, using a reinforcement learning process to generate an initial infrastructure allocation for the computing workload, and a reward function of the reinforcement learning process generates a respective reward for each initial infrastructure allocation;

running an accumulator map voting process to generate a total reward for each initial infrastructure allocation; and

identifying the initial infrastructure allocation with the largest total reward and assigning that initial infrastructure allocation to the computing workload.

2 . The method as recited in claim 1 , wherein the reinforcement learning process comprises a Deep Q-Network reinforcement learning process.

3 . The method as recited in claim 1 , wherein one of the rewards has a value that indicates a relation between an execution time of the computing workload and an execution time specified by a service level agreement, and the execution time of the computing workload is the time taken for execution of the computing workload by the initial infrastructure allocation to which the reward value corresponds.

4 . The method as recited in claim 1 , wherein running the reward function identifies, for each of the initial infrastructure allocations, one of: computing resource wastage; a service level agreement violation; or, conformance with a service level agreement requirement.

5 . The method as recited in claim 1 , wherein a reward value of zero for an initial infrastructure allocation indicates that computing resources included in that initial infrastructure allocation exceed the computing resources needed to execute the computing workload in a manner that meets requirements of a service level agreement.

6 . The method as recited in claim 1 , wherein a negative reward value for an initial infrastructure allocation indicates that computing resources included in that initial infrastructure allocation are inadequate to execute the computing workload in a manner that meets requirements of a service level agreement.

7 . The method as recited in claim 1 , wherein a reward value for an initial infrastructure allocation is at a maximum at a point between a zero reward value and a negative reward value.

8 . The method as recited in claim 1 , wherein inputs to the reward function comprise an execution time x of an epoch of the workload, a service level agreement value μ, and a parameter o that defines how quickly a reward curve decays.

9 . The method as recited in claim 1 , wherein a plot of the reward function comprises a reward band that includes a range of reward values, and each of the reward values in the reward band corresponds to an initial infrastructure allocation that is capable of executing the computing workload according to a requirement specified in a service level agreement.

10 . The method as recited in claim 9 , wherein the reward band includes a positive reward value, a maximum reward value, and a negative reward value.

11 . A computer readable storage medium having stored therein instructions that are executable by one or more hardware processors to perform operations comprising:

running multiple iterations of a computing workload;

for each iteration of the computing workload, using a reinforcement learning process to generate an initial infrastructure allocation for the computing workload, and a reward function of the reinforcement learning process generates a respective reward for each initial infrastructure allocation;

running an accumulator map voting process to generate a total reward for each initial infrastructure allocation; and

identifying the initial infrastructure allocation with the largest total reward and assigning that initial infrastructure allocation to the computing workload.

12 . The computer readable storage medium as recited in claim 11 , wherein the reinforcement learning process comprises a Deep Q-Network reinforcement learning process.

13 . The computer readable storage medium as recited in claim 11 , wherein one of the rewards has a value that indicates a relation between an execution time of the computing workload and an execution time specified by a service level agreement, and the execution time of the computing workload is the time taken for execution of the computing workload by the initial infrastructure allocation to which the reward value corresponds.

14 . The computer readable storage medium as recited in claim 11 , wherein running the reward function identifies, for each of the initial infrastructure allocations, one of: computing resource wastage; a service level agreement violation; or, conformance with a service level agreement requirement.

15 . The computer readable storage medium as recited in claim 11 , wherein a reward value of zero for an initial infrastructure allocation indicates that computing resources included in that initial infrastructure allocation exceed the computing resources needed to execute the computing workload in a manner that meets requirements of a service level agreement.

16 . The computer readable storage medium as recited in claim 11 , wherein a negative reward value for an initial infrastructure allocation indicates that computing resources included in that initial infrastructure allocation are inadequate to execute the computing workload in a manner that meets requirements of a service level agreement.

17 . The computer readable storage medium as recited in claim 11 , wherein a reward value for an initial infrastructure allocation is at a maximum at a point between a zero reward value and a negative reward value.

18 . The computer readable storage medium as recited in claim 11 , wherein inputs to the reward function comprise an execution time x of an epoch of the workload, a service level agreement value μ, and a parameter o that defines how quickly a reward curve decays.

19 . The computer readable storage medium as recited in claim 11 , wherein a plot of the reward function comprises a reward band that includes a range of reward values, and each of the reward values in the reward band corresponds to an initial infrastructure allocation that is capable of executing the computing workload according to a requirement specified in a service level agreement.

20 . The computer readable storage medium as recited in claim 19 , wherein the reward band includes a positive reward value, a maximum reward value, and a negative reward value.

Assignments (10)
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (056295/0280) Recorded Jun 10, 2022
From: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
To: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
Reel/Frame 062022/0255 →
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (056295/0124) Recorded Jun 10, 2022
From: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
To: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
Reel/Frame 062022/0012 →
RELEASE OF SECURITY INTEREST IN PATENTS PREVIOUSLY RECORDED AT REEL/FRAME (056295/0001) Recorded Jun 10, 2022
From: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
To: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
Reel/Frame 062021/0844 →
RELEASE OF SECURITY INTEREST Recorded Nov 2, 2021
From: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
To: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
Reel/Frame 058297/0332 →
SECURITY INTEREST Recorded May 19, 2021
From: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
To: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
Reel/Frame 056295/0124 →
SECURITY INTEREST Recorded May 19, 2021
From: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
To: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
Reel/Frame 056295/0001 →
SECURITY INTEREST Recorded May 19, 2021
From: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
To: THE BANK OF NEW YORK MELLON TRUST COMPANY, N.A., AS NOTES COLLATERAL AGENT
Reel/Frame 056295/0280 →
CORRECTIVE ASSIGNMENT TO CORRECT THE MISSING PATENTS THAT WERE ON THE ORIGINAL SCHEDULED SUBMITTED BUT NOT ENTERED PREVIOUSLY RECORDED AT REEL: 056250 FRAME: 0541. ASSIGNOR(S) HEREBY CONFIRMS THE ASSIGNMENT. Recorded May 17, 2021
From: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056311/0781 →
SECURITY AGREEMENT Recorded May 14, 2021
From: DELL PRODUCTS L.P.; EMC IP HOLDING COMPANY LLC
To: CREDIT SUISSE AG, CAYMAN ISLANDS BRANCH
Reel/Frame 056250/0541 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 21, 2021
From: SOUSA, EDUARDO VERA; OLIVEIRA, ANA CRISTINA BERNARDO DE
To: EMC IP HOLDING COMPANY LLC
Reel/Frame 055992/0728 →