IP Library › Granted Patent US 10,965,541
Granted Patent B2
US 10,965,541 · App. 15/977,419 · Granted Mar 30, 2021

Method and system to proactively determine potential outages in an information technology environment

Inventors: Muraleedharan Vijayakumar (Chennai, IN); Veeramanikandan Pandiaraj (Chennai, IN); Mahesh Marimuthu (Chennai, IN); Govindaraj Muniyandi (Chennai, IN)
Assignee: GAVS Technologies Pvt. Ltd.
H04L41/147G06F11/008H04L41/06H04L41/14H04L41/22H04L41/5016H04L41/5074H04L43/0876G06F16/345H04L43/16
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 10,965,541
App. No.
15/977,419
Granted
Mar 30, 2021
Kind
B2
Abstract

A method and a system for determining and preventing outages in an IT network by predicting status, utilization, performance, or a combination thereof for IT resources is disclosed. The method includes extracting and classifying data for one or more parameters associated with a plurality of nodes. A set of historical metrics and real-time metrics are used for predicting status score, utilization score, and performance score of IT infrastructure resources. The predictions are compared with a predetermined threshold limit for identifying potential outage in the network. A summary indicating the predictions are displayed to an administrator for preventing and mitigating the potential downtime.

Claims (51)

1. A computer-implemented method for determining potential outages in an information technology (IT) environment, the computer-implemented method comprising:

extracting, from one or more data sources, data for one or more parameters associated with a plurality of nodes in the IT environment, the data comprising at least utilization metrics, performance metrics and a time identifier for each of the utilization and performance metrics;

classifying the data as historical data or current data based on the time identifier;

predicting a status score, an utilization score, and a performance score for the plurality of nodes from the classified data, wherein predicting the status score, utilization score, and performance score comprises:

extracting a training dataset from the historical data,

training a prediction model based on the training dataset using a machine learning engine,

providing the current data to the trained prediction model as a test dataset, and

obtaining predictions on the status score, utilization score, and performance score from the prediction model;

determining a potential outage in the IT environment from the predicted scores;

automatically identifying one or more processes with high utilization and less importance based on the determined potential outages and historical data;

displaying a summary, in one or more devices, indicating the predictions and the potential outage, wherein the summary comprises at least trends and statistics associated with the predicted scores;

determining that the predicted scores or potential outage exceed a threshold limit;

triggering automated workflows to kill the one or more identified processes to optimize utilization based on determining that the predicted scores or potential outage exceed the threshold limit;

sending alerts to the one or more devices based on the predicted scores or potential outage exceed the threshold limit;

assigning tickets to one or more operators based on the determined potential outages.

2. A system for determining potential outages in an information technology (IT) environment, the system comprising:

a user interface;

one or more hardware processing units;

a hardware memory unit coupled to the one or more hardware processing units, wherein the hardware memory unit comprises:

a data extraction module configured to extract, from or more data sources, data for one or more parameters associated with a plurality of nodes in the IT environment, the data comprising at least utilization metrics, performance metrics and a time identifier for each of the utilization and performance metrics;

a data classifier module configured to classify the data as historical data or current data based on the time identifier;

a prediction module configured to predict a status score, an utilization score, and a performance score for the plurality of nodes based on the classified data, and determine potential outage in the IT environment from predicted scores, wherein predicting the status score, utilization score, and performance score comprises:

extracting a training dataset from the historical data,

training a prediction model based on the training dataset using a machine learning engine,

providing the current data to the trained prediction model as a test dataset, and

obtaining predictions on the status score, utilization score, and performance score from the prediction model;

identify one or more processes with high utilization and less importance automatically based on the determined potential outages and historical data;

determine that the predicted scores or potential outage exceed a threshold limit; and

trigger automated workflows to kill the one or more identified processes to optimize utilization based on the determined predicted scores or potential outage exceed the threshold limit;

a display module configured to display a summary, in one or more devices, indicating the predictions and the potential outage, wherein the summary comprises at least trends and statistics associated with the predicted scores;

an alert module configured to send alerts to the one or more devices based on the predicted scores or potential outage exceed the threshold limit;

a ticketing module configured to assign tickets to one or more operators based on the determined potential outages.

3. The system of claim 2 , wherein the hardware memory unit further comprises a summary generation module configured to generate the summary comprising at least the trends and the statistics associated with the predicted scores.

4. The system of claim 2 , wherein the data extraction module communicates with a plurality of agents and counters to extract the data from the one or more data sources.

5. The system of claim 4 , wherein the data sources comprise monitoring tools installed on servers, desk tools installed on user devices, or database in the IT environment.

6. A computer program product having non-volatile memory therein, carrying computer executable instructions stored therein for determining potential outages in an Information Technology (IT) environment, the computer executable instructions comprising:

extracting, from one or more data sources, data for one or more parameters associated with a plurality of nodes in the IT environment, the data comprising at least utilization metrics, performance metrics and a time identifier for each of the utilization and performance metrics;

classifying the data as historical data or current data based on the time identifier;

predicting a status score, an utilization score, and a performance score for the plurality of nodes from the classified data, wherein predicting the status score, utilization score, and performance score comprises:

extracting a training dataset from the historical data,

training a prediction model based on the training dataset using a machine learning engine,

providing the current data to the trained prediction model as a test dataset, and

obtaining predictions on the status score, utilization score, and

performance score from the prediction model;

determining a potential outage in the IT environment from the predicted scores;

automatically identifying one or more processes with high utilization and less importance based on the determined potential outages and historical data;

displaying a summary, in one or more devices, indicating the predictions and the potential outage, wherein the summary comprises at least trends and statistics associated with the predicted scores;

determining that the predicted scores or potential outage exceed a threshold limit;

triggering automated workflows to kill the one or more identified processes to optimize utilization based on determining that the predicted scores or potential outage exceed the threshold limit;

sending alerts to the one or more devices based on the predicted scores or potential outage exceed the threshold limit;

assigning tickets to one or more operators based on the determined potential outages.

Assignments (1)
CHANGE OF NAME Recorded Jul 24, 2026
From: GAVS TECHNOLOGIES PRIVATE LIMITED
To: NEUREALM PRIVATE LIMITED
Reel/Frame 076048/0886 →
Priority Claims (1)
IN 201841006251 · Feb 19, 2018 · national
Continuity (1)
Related Publication 20190044825A1 · Feb 7, 2019
Cited By (1)
US 12,542,723