IP Library › Granted Patent US 11,307,949
Granted Patent B2
US 11,307,949 · App. 15/813,953 · Granted Apr 19, 2022

Decreasing downtime of computer systems using predictive detection

Inventors: Rares Ioan Almasan (Phoenix, AZ); Jeffery Freed (Scottsdale, AZ); Kai Wang (Stony Brook, NY)
Assignee: American Express Travel Related Services Company, Inc.
G06F11/302G05B23/0218G06F11/008G06F11/22G06F11/261G06F11/30G06F11/34G06F11/3438G06F16/20G06N5/04G06N20/00H04L41/16
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,307,949
App. No.
15/813,953
Filed
Nov 15, 2017
Granted
Apr 19, 2022
Kind
B2
Examiner
TRUONG, LOAN
Art Unit
2114
USPC
714/15
Abstract

A master processor may retrieve historical and real time machine and human data related to computer system health. The master processor may utilize machine learning and artificial intelligence to predict potential computer malfunctions. The master processor may output notifications regarding the potential computer malfunctions in order to prevent the computer malfunctions from occurring.

Claims (43)

1. A method, comprising:

retrieving, by a processor, machine data from a machine data source;

retrieving, by the processor, incident report data describing one or more computer malfunctions from an incident report data source;

retrieving, by the processor, a change record representing a scheduled change to a configuration of a computer system, the change record specifying when the scheduled change is scheduled to occur, a number of phases in which the scheduled change will be implemented, and one or more environments in which the scheduled change has been tested;

retrieving, by the processor, an algorithm from a model library;

determining, by the processor using the algorithm and based at least in part on the scheduled change described by the change record, the machine data, and the incident report data, a likelihood that the change record will cause a computer malfunction in the computer system;

determining, by the processor using the algorithm and based at least in part on the machine data and the incident report data, an estimated cost associated with the computer malfunction; and

generating, by the processor, a change record evaluation, the change record evaluation comprising the likelihood that the change record will cause the computer malfunction and the estimated cost associated with the computer malfunction.

2. The method of claim 1 , wherein the machine data comprises historical data and real time data.

3. The method of claim 1 , further comprising correlating, by the processor and using the machine data, a first spike in computing resource consumption by a first application with a second spike in computing resource consumption by a second application, the first spike corresponding to a first increase of at least three standard deviations over average variability for the computing resource consumption by the first application, and the second spike corresponding to a second increase of at least three standard deviations over average variability for the computing resource consumption by the second application.

4. The method of claim 1 , wherein the change record comprises a software upgrade and a time of upgrade.

5. The method of claim 1 , further comprising recalibrating, by the processor, the algorithm based at least in part on feedback from a subject matter expert.

6. The method of claim 5 , wherein the problem record identifies a previous problem, and wherein the incident record identifies a current problem.

7. The method of claim 1 , further comprising selecting the algorithm from a plurality of algorithms stored in the model library based at least in part on an accuracy of the algorithm.

8. A system comprising:

a processor; and

a tangible, non-transitory memory configured to communicate with the processor, the tangible, non-transitory memory having instructions stored thereon that, in response to execution by the processor, cause the processor to perform operations comprising:

retrieving, by the processor, machine data from a machine data source;

retrieving, by the processor, incident report data describing one or more computer malfunctions from an incident report data source;

retrieving, by the processor, a change record representing a scheduled change to a configuration of a computer system, the change record specifying when the scheduled change is scheduled to occur, a number of phases in which the scheduled change will be implemented, and one or more environments in which the scheduled change has been tested;

retrieving, by the processor, an algorithm from a model library;

determining, by the processor using the algorithm and based at least in part on the scheduled change described by the change record, the machine data, and the incident report data, a likelihood that the change record will cause a computer malfunction in the computer system;

determining, by the processor using the algorithm and based at least in part on the machine data and the incident report data, an estimated cost associated with the computer malfunction; and

generating, by the processor, a change record evaluation, the change record evaluation comprising the likelihood that the change record will cause the computer malfunction and the estimated cost associated with the computer malfunction.

9. The system of claim 8 , wherein the machine data comprises historical data and real time data.

10. The system of claim 8 , the operations further comprising correlating, by the processor and using the machine data, a first spike in computing resource consumption by a first application with a second spike in computing resource consumption by a second application, the first spike corresponding to a first increase of at least three standard deviations over average variability for the computing resource consumption by the first application, and the second spike corresponding to a second increase of at least three standard deviations over average variability for the computing resource consumption by the second application.

11. The system of claim 8 , wherein the change record comprises a software upgrade and a time of upgrade.

12. The system of claim 8 , the operations further comprising recalibrating, by the processor, the algorithm based at least in part on feedback from a subject matter expert.

13. The system of claim 12 , wherein the problem record identifies a previous problem, and wherein the incident record identifies a current problem.

14. The system of claim 8 , the operations further comprising selecting the algorithm from a plurality of algorithms stored in the model library based at least in part on an accuracy of the algorithm.

15. An article of manufacture including a non-transitory, tangible computer readable storage medium having instructions stored thereon that, in response to execution by a computer based system, cause the computer based system to perform operations comprising:

retrieving, by a processor, machine data from a machine data source;

retrieving, by the processor, incident report data describing one or more computer malfunctions from an incident report data source;

retrieving, by the processor, a change record representing a scheduled change to a configuration of a computer system, the change record specifying when the scheduled change is scheduled to occur, a number of phases in which the scheduled change will be implemented, and one or more environments in which the scheduled change has been tested;

retrieving, by the processor, an algorithm from a model library;

determining, by the processor using the algorithm and based at least in part on the scheduled change described by the change record, the machine data, and the incident report data, a likelihood that the change record will cause a computer malfunction in the computer system;

determining, by the processor using the algorithm and based at least in part on the machine data and the incident report data, an estimated cost associated with the computer malfunction; and

generating, by the processor, a change record evaluation, the change record evaluation comprising the likelihood that the change record will cause the computer malfunction and the estimated cost associated with the computer malfunction.

16. The article of manufacture of claim 15 , wherein the machine data comprises historical data and real time data.

17. The article of manufacture of claim 15 , the operations further comprising correlating, by the processor and using the machine data, a first spike in computing resource consumption by a first application with a second spike in computing resource consumption by a second application, the first spike corresponding to a first increase of at least three standard deviations over average variability for the computing resource consumption by the first application, and the second spike corresponding to a second increase of at least three standard deviations over average variability for the computing resource consumption by the second application.

18. The article of manufacture of claim 15 , wherein the change record comprises a software upgrade and a time of upgrade.

19. The article of manufacture of claim 15 , the operations further comprising recalibrating, by the processor, the algorithm based at least in part on feedback from a subject matter expert.

20. The article of manufacture of claim 19 , wherein the problem record identifies a previous problem, and wherein the incident record identifies a current problem.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Nov 15, 2017
From: ALMASAN, RARES IOAN; FREED, JEFFERY; WANG, KAI
To: AMERICAN EXPRESS TRAVEL RELATED SERVICES COMPANY, INC.
Reel/Frame 044138/0160 →
Continuity (1)
Related Publication 20190149426A1 · May 16, 2019
Cited By (1)
US 12,699,609