IP Library Granted Patent US 9,959,138
Granted Patent B1
US 9,959,138 · App. 14/852,249 · Granted May 1, 2018

Adaptive self-maintenance scheduler

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 9,959,138
App. No.
14/852,249
Granted
May 1, 2018
Kind
B1
Abstract

Embodiments presented herein disclose adaptive techniques for scheduling self-maintenance processes. A load predictor estimates, based on a current state of a distributed storage system, an amount of resources of the system required to perform each of a plurality of self-maintenance processes. A maintenance process scheduler estimates, based on one or more inputs, an amount of resources of the distributed system available to perform one or more of the self-maintenance processes during at least a first time period. The maintenance process scheduler determines a schedule for the one or more of the self-maintenance processes to perform during the first time period, based on the estimated amount of resources required and available.

Claims (45)

1. A method, comprising:

estimating, based on a current state of a distributed storage system comprising a primary storage server and a plurality of secondary storage servers, an amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of a plurality of self-maintenance processes wherein the plurality of secondary storage servers are configured to perform one or more backup processes and the plurality of self-maintenance processes, wherein the plurality of self-maintenance processes include at least one of garbage collection and correcting inconsistencies associated with the distributed file system, wherein the plurality of secondary storage servers form a distributed file system providing backup storage services to the primary storage server;

estimating, based on one or more inputs, an amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes during at least a first time period of a plurality of time periods;

determining which of the plurality of self-maintenance processes to perform during the at least first time period based on the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the plurality of self-maintenance processes and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes wherein one or more of the self-maintenance processes are determined to be performed during the at least first time period; and

scheduling, based on the estimated amount of computing resources required and on the estimated amount of computing resources available, the determined one or more self-maintenance processes to perform during the first time period.

2. The method of claim 1 , wherein the one or more inputs includes at least one of a plurality of current activities of the primary storage server and the plurality of secondary storage servers, the current state of the primary storage server and the state of the plurality of secondary storage servers, a plurality of external events, and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the self-maintenance processes.

3. The method of claim 1 , further comprising:

performing each of the scheduled self-maintenance processes during the first time period; and

collecting execution statistics of the performance of each of the scheduled self-maintenance processes.

4. The method of claim 3 , wherein the amount of computing resources of the primary storage server and the plurality of secondary storage servers required and the amount of computing resources of the primary storage server and the plurality of secondary storage servers available are further estimated based on the execution statistics.

5. The method of claim 4 , further comprising:

upon determining that the estimated amount of computing resources available is within a specified range of an actual amount of computing resources available based on the execution statistics, reinforcing the one or more inputs using a machine learning algorithm; and

upon determining that the estimated amount of computing resources available is not within the specified range of the actual amount of computing resources available based on the execution statistics, readjusting the one or more inputs using the machine learning algorithm.

6. The method of claim 1 , wherein the amount of computing resources required and amount of computing resources available include at least one of an amount of I/O resources, network resources, storage capacity, and processing resources.

7. A non-transitory computer-readable storage medium having instructions, which, when executed on a processor, perform an operation comprising:

estimating, based on a current state of a distributed storage system comprising a primary storage server and a plurality of secondary storage servers, an amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of a plurality of self-maintenance processes wherein the plurality of secondary storage servers are configured to perform one or more backup processes and the plurality of self-maintenance processes, wherein the plurality of self-maintenance processes include at least one of garbage collection and correcting inconsistencies associated with the distributed file system, wherein the plurality of secondary storage servers form a distributed file system providing backup storage services to the primary storage server;

estimating, based on one or more inputs, an amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes during at least a first time period of a plurality of time periods;

determining which of the plurality of self-maintenance processes to perform during the at least first time period based on the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the plurality of self-maintenance processes and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes, wherein one or more of the self-maintenance processes are determined to be performed during the at least first time period; and

scheduling, based on the estimated amount of computing resources required and on the estimated amount of computing resources available, the determined one or more self-maintenance processes to perform during the first time period.

8. The computer-readable storage medium of claim 7 , wherein the one or more inputs includes at least one of a plurality of current activities of the primary storage server and the plurality of secondary storage servers, the current state of the primary storage server and the state of the plurality of secondary storage servers, a plurality of external events, and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the self-maintenance processes.

9. The computer-readable storage medium of claim 7 , wherein the operation further comprises:

performing each of the scheduled self-maintenance processes during the first time period; and

collecting execution statistics of the performance of each of the scheduled self-maintenance processes.

10. The computer-readable storage medium of claim 9 , wherein the amount of computing resources of the primary storage server and the plurality of secondary storage servers required and the amount of computing resources of the primary storage server and the plurality of secondary storage servers available are further estimated based on the execution statistics.

11. The computer-readable storage medium of claim 10 , the operation further comprising:

upon determining that the estimated amount of computing resources available is within a specified range of an actual amount of computing resources available based on the execution statistics, reinforcing the one or more inputs using a machine learning algorithm; and

upon determining that the estimated amount of computing resources available is not within the specified range of the actual amount of computing resources available based on the execution statistics, readjusting the one or more inputs using the machine learning algorithm.

12. The computer-readable storage medium of claim 7 , wherein

the amount of computing resources required and amount of computing resources available include at least one of an amount of I/O resources, network resources, storage capacity, and processing resources.

13. A system, comprising:

a processor; and

a memory containing program code, which, when executed on the processor, performs an operation comprising:

estimating, based on a current state of a distributed storage system comprising a primary storage server and a plurality of secondary storage servers, an amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of a plurality of self-maintenance processes wherein the plurality of secondary storage servers are configured to perform one or more backup processes and the plurality of self-maintenance processes, wherein the plurality of self-maintenance processes include at least one of garbage collection and correcting inconsistencies associated with the distributed file system, wherein the plurality of secondary storage servers form a distributed file system providing backup storage services to the primary storage server;

estimating, based on one or more inputs, an amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes during at least a first time period of a plurality of time periods;

determine which of the plurality of self-maintenance processes to perform during the at least first time period based on the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the plurality of self-maintenance processes and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers available to perform one or more of the plurality of self-maintenance processes, wherein one or more of the self-maintenance processes are determined to be performed during the at least first time period; and

scheduling, based on the estimated amount of computing resources required and on the estimated amount of computing resources available, the determined one or more self-maintenance processes to perform during the first time period.

14. The system of claim 13 , wherein the one or more inputs includes at least one of a plurality of current activities of the primary storage server and the plurality of secondary storage servers, the current state of the primary storage server and the state of the plurality of secondary storage servers, a plurality of external events, and the estimated amount of computing resources of the primary storage server and the plurality of secondary storage servers required to perform each of the self-maintenance processes.

15. The system of claim 13 , wherein the operation further comprises:

performing each of the scheduled self-maintenance processes during the first time period; and

collecting execution statistics of the performance of each of the scheduled self-maintenance processes.

16. The system of claim 15 , wherein the amount of computing resources of the primary storage server and the plurality of secondary storage servers required and the amount of computing resources of the primary storage server and the plurality of secondary storage servers available are further estimated based on the execution statistics.

17. The system of claim 16 , the operation further comprising:

upon determining that the estimated amount of computing resources available is within a specified range of an actual amount of computing resources available based on the execution statistics, reinforcing the one or more inputs using a machine learning algorithm; and

upon determining that the estimated amount of computing resources available is not within the specified range of the actual amount of computing resources available based on the execution statistics, readjusting the one or more inputs using the machine learning algorithm.

18. The system of claim 13 , wherein the amount of computing resources required and amount of computing resources available include at least one of an amount of I/O resources, network resources, storage capacity, and processing resources.

Assignments (4)
TERMINATION AND RELEASE OF INTELLECTUAL PROPERTY SECURITY AGREEMENT Recorded Dec 10, 2024
From: FIRST-CITIZENS BANK & TRUST COMPANY (AS SUCCESSOR TO SILICON VALLEY BANK)
To: COHESITY, INC.
Reel/Frame 069584/0498 →
SECURITY INTEREST Recorded Dec 9, 2024
From: VERITAS TECHNOLOGIES LLC; COHESITY, INC.
To: JPMORGAN CHASE BANK. N.A.
Reel/Frame 069890/0001 →
SECURITY INTEREST Recorded Sep 23, 2022
From: COHESITY, INC.
To: SILICON VALLEY BANK, AS ADMINISTRATIVE AGENT
Reel/Frame 061509/0818 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 11, 2015
From: VAISH, TARANG; DUTTAGUPTA, ANIRVAN; MADDURI, SASHI
To: COHESITY, INC.
Reel/Frame 036547/0388 →