IP Library Granted Patent US 8,645,769
Granted Patent B2
US 8,645,769 · App. 13/133,719 · Granted Feb 4, 2014

Operation management apparatus, operation management method, and program storage medium

Inventor: Hideo Hasegawa (Tokyo, JP)
Assignee: NEC Corporation
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 8,645,769
App. No.
13/133,719
Granted
Feb 4, 2014
Kind
B2
Abstract

A time in which an administrator identifies a cause of a failure when the abnormality is detected in a plurality of servers is shortened. An operation management apparatus includes a failure detection unit 26 and a spread determination unit 27 . The failure detection unit 26 obtains measured values of a plurality of performance metrics with respect to each of a plurality of target apparatuses to be monitored that are connected to a common apparatus and detects an abnormal item which is the performance metric whose measured value is abnormal with respect to each of the plurality of target apparatuses to be monitored. The spread determination unit 27 outputs the remaining abnormal item excluding the abnormal item detected in all the plurality of target apparatuses to be monitored from the abnormal item of each of the plurality of target apparatuses to be monitored.

Claims (34)

1. An operation management apparatus comprising:

a failure detection unit which obtains measured values of a plurality of performance metrics with respect to each of a plurality of target apparatuses to be monitored that are connected to a common apparatus and detects an abnormal item which is a performance metric whose measured value is abnormal with respect to each of said plurality of target apparatuses to be monitored;

a spread determination unit which outputs remaining abnormal item excluding said abnormal item detected in all said plurality of target apparatuses to be monitored from said abnormal item of each of said plurality of target apparatuses to be monitored; and

a memory which stores said obtained measured values of said plurality of performance metrics.

2. The operation management apparatus according to claim 1 further comprising a correlation model storage unit which stores a transform function for each two different performance metrics among said plurality of performance metrics with respect to each of said plurality of target apparatuses to be monitored, said transform function indicating a correlation between said two performance metrics,

wherein said failure detection unit detects said two performance metrics as said abnormal item when a difference between a value obtained by inputting a measured value of one of said two performance metrics among said plurality of performance metrics in said transform function corresponding to said two performance metrics and a measured value of the other is equal to or greater than a predetermined value.

3. The operation management apparatus according to claim 2 further comprising an abnormality score calculation unit which outputs a proportion of the number of said abnormal items outputted by said spread determination unit in the number of said transform functions of said target apparatus to be monitored as an abnormality score with respect to each of said plurality of target apparatuses to be monitored.

4. The operation management apparatus according to claim 3 ,

wherein said memory stores said measured values of said plurality of performance metrics in time series that are measured in each of said plurality of target apparatuses to be monitored; and

further comprising a correlation model generation unit which generates said transform function based on said measured values of said plurality of performance metrics for a predetermined period that are stored in said memory and saves said generated transform function in said correlation model storage unit.

5. An operation management method comprising:

obtaining measured values of a plurality of performance metrics with respect to each of a plurality of target apparatuses to be monitored that are connected to a common apparatus;

detecting an abnormal item which is a performance metric whose measured value is abnormal with respect to each of said plurality of target apparatuses to be monitored; and

outputting remaining abnormal item excluding said abnormal item detected in all said plurality of target apparatuses to be monitored from said abnormal item of each of said plurality of target apparatuses to be monitored.

6. The operation management method according to claim 5 further comprising storing a transform function for each two different performance metrics among said plurality of performance metrics with respect to each of said plurality of target apparatuses to be monitored, said transform function indicating a correlation between said two performance metrics,

wherein said detecting an abnormal item detects said two performance metrics as said abnormal item when a difference between a value obtained by inputting a measured value of one of said two performance metrics among said plurality of performance metrics in said transform function corresponding to said two performance metrics and a measured value of the other is equal to or greater than a predetermined value.

7. The operation management method according to claim 6 further comprising outputting a proportion of the number of said abnormal items in the number of said transform functions of said target apparatus to be monitored as an abnormality score with respect to each of said plurality of target apparatuses to be monitored.

8. The operation management method according to claim 7 further comprising:

storing said measured values of said plurality of performance metrics in time series that are measured in each of said plurality of target apparatuses to be monitored; and

generating said transform function based on said measured values of said plurality of performance metrics for a predetermined period.

9. A non-transitory computer readable medium recording thereon an operation management program, causing computer to perform a method comprising:

obtaining measured values of a plurality of performance metrics with respect to each of a plurality of target apparatuses to be monitored that are connected to a common apparatus;

detecting an abnormal item which is a performance metric whose measured value is abnormal with respect to each of said plurality of target apparatuses to be monitored; and

outputting remaining abnormal item excluding said abnormal item detected in all said plurality of target apparatuses to be monitored from said abnormal item of each of said plurality of target apparatuses to be monitored.

10. The non-transitory computer readable medium according to claim 9 , recording thereon said operation management program, further comprising storing a transform function for each two different performance metrics among said plurality of performance metrics with respect to each of said plurality of target apparatuses to be monitored, said transform function indicating a correlation between said two performance metrics,

wherein said detecting an abnormal item detects said two performance metrics as said abnormal item when a difference between a value obtained by inputting a measured value of one of said two performance metrics among said plurality of performance metrics in said transform function corresponding to said two performance metrics and a measured value of the other is equal to or greater than a predetermined value.

11. The non-transitory computer readable medium according to claim 10 , recording thereon said operation management program, further comprising outputting a proportion of the number of said abnormal items in the number of said transform functions of said target apparatus to be monitored as an abnormality score with respect to each of said plurality of target apparatuses to be monitored.

12. The non-transitory computer readable medium according to claim 11 , recording thereon said operation management program, further comprising:

storing said measured values of said plurality of performance metrics in time series that are measured in each of said plurality of target apparatuses to be monitored; and

generating said transform function based on said measured values of said plurality of performance metrics for a predetermined period.

13. An operation management apparatus comprising:

failure detection means for obtaining measured values of a plurality of performance metrics with respect to each of a plurality of target apparatuses to be monitored that are connected to a common apparatus and detecting an abnormal item which is a performance metric whose measured value is abnormal with respect to each of said plurality of target apparatuses to be monitored;

spread determination means for outputting remaining abnormal item excluding said abnormal item detected in all said plurality of target apparatuses to be monitored from said abnormal item of each of said plurality of target apparatuses to be monitored; and

memory means for storing said obtained measured values of said plurality of performance metrics.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 10, 2011
From: HASEGAWA, HIDEO
To: NEC CORPORATION
Reel/Frame 026423/0732 →
Priority Claims (1)
JP 2010-003008 · Jan 8, 2010 · national
Continuity (1)
Related Publication 20120278663A1 · Nov 1, 2012