IP Library › Granted Patent US 11,366,730
Granted Patent B2
US 11,366,730 · App. 16/677,647 · Granted Jun 21, 2022

Determining an availability score based on available resources at a first server to determine whether to direct processing from a first server to a second server

Inventors: Herve G. P. Andre (Orlando, FL); Matthew D. Carson (Encino, CA); Rashmi Chandra (Santa Clara, CA); Clint A. Hardy (Tucson, AZ); Larry Juarez (Tucson, AZ); Tony Leung (Tucson, AZ); Todd C. Sorenson (Tucson, AZ)
Assignee: INTERNATIONAL BUSINESS MACHINES CORPORATION
G06F11/2094G06F11/008G06F11/203G06F11/2069G06F11/3034G06F11/3495G06F11/2071G06F2201/805
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,366,730
App. No.
16/677,647
Granted
Jun 21, 2022
Kind
B2
Abstract

Provided are a computer program product, system, and method for a computer program product, system, and method for determining an availability score based on available resources of different resource types in a distributed computing environment of storage servers to determine whether to perform a failure operation for one of the storage servers. A health status monitor program deployed in the storage servers performs: maintaining information indicating availability of a plurality of storage server resources for a plurality of resource types; calculating an availability score as a function of a number of available resources of the resource types; and transmitting information on the availability score to a management program. The management program uses the transmitted information to determine whether to migrate services from the storage server from which the availability score is received to at least one of the other storage servers in the distributed computing environment.

Claims (32)

1. A computer program product for determining a health status of a first server in a network including a second server, wherein the computer program product comprises a computer readable storage medium having program code executed in the first server to perform operations, the operations comprising:

receiving an error message for a resource;

calculating an availability score as a function of a number of available resources at the first server and a number of recovery events at the first server; and

performing a failover from the first server to the second server in response to the availability score having a severity level indicating a system failure.

2. The computer program product of claim 1 , wherein the operations further comprise:

transmitting information on the availability score to a management program in a management system in the network to use to determine whether to perform the failover.

3. The computer program product of claim 1 , wherein the operations further comprise:

performing a failback from the second server to the first server in response to the failover from the first server to the second server and the availability score not comprising the severity level indicating the system failure.

4. The computer program product of claim 1 , wherein the function for calculating the availability score additionally considers a total number of resources at the first server and a total number of allowed recovery events at the first server.

5. The computer program product of claim 4 , wherein the function calculates the availability score by combining a percentage of the number of available resources to the total number of resources and a percentage of the number of recovery events divided by the total number of allowed recovery events.

6. A system in communication with a remote server over a network, comprising:

a processor; and

a computer readable storage medium having program code executed to perform operations, the operations comprising:

receiving an error message for a resource;

calculating an availability score as a function of a number of available resources at the system and a number of recovery events at the system; and

performing a failover from the system to the remote server in response to the availability score having a severity level indicating a system failure.

7. The system of claim 6 , wherein the operations further comprise:

transmitting information on the availability score to a management program in a management system in the network to use to determine whether to perform the failover.

8. The system of claim 6 , wherein the operations further comprise:

performing a failback from the remote server to the system in response to the failover from the system to the remote server and the availability score not comprising the severity level indicating the system failure.

9. The system of claim 6 , wherein the function for calculating the availability score additionally considers a total number of resources at the system and a total number of allowed recovery events at the system.

10. The system of claim 9 , wherein the function calculates the availability score by combining a percentage of the number of available resources to the total number of resources and a percentage of the number of recovery events divided by the total number of allowed recovery events.

11. A method for determining a health status of a first server in a network including a second server, comprising:

receiving an error message for a resource;

calculating an availability score as a function of a number of available resources at the first server and a number of recovery events at the first server; and

performing a failover from the first server to the second server in response to the availability score having a severity level indicating a system failure.

12. The method of claim 11 , further comprising:

transmitting information on the availability score to a management program in a management system in the network to use to determine whether to perform the failover.

13. The method of claim 11 , further comprising:

performing a failback from the second server to the first server in response to the failover from the first server to the second server and the availability score not comprising the severity level indicating the system failure.

14. The method of claim 11 , wherein the function for calculating the availability score additionally considers a total number of resources at the first server and a total number of allowed recovery events at the first server.

15. The method of claim 14 , wherein the function calculates the availability score by combining a percentage of the number of available resources to the total number of resources and a percentage of the number of recovery events divided by the total number of allowed recovery events.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 8, 2019
From: ANDRE, HERVE G.P.; CARSON, MATTHEW D.; CHANDRA, RASHMI; HARDY, CLINT A.; JUAREZ, LARRY; LEUNG, TONY; SORENSON, TODD C.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 050962/0869 →
Continuity (4)
Continuation 15792751 · Oct 25, 2017
Continuation 15206093 · Jul 8, 2016
Continuation 14289333 · May 28, 2014
Related Publication 20200073772A1 · Mar 5, 2020