IP Library › Granted Patent US 12,066,908
Granted Patent B2
US 12,066,908 · App. 17/877,076 · Granted Aug 20, 2024

System and method for predicting and avoiding hardware failures using classification supervised machine learning

Inventors: Deeder M. Aurongzeb (Austin, TX); Malathi Ramakrishnan (Madurai, IN); Parminder Singh Sethi (Punjab, IN)
Assignee: DELL PRODUCTS LP
G06F11/2033G06F9/5044G06F9/5055
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,066,908
App. No.
17/877,076
Granted
Aug 20, 2024
Kind
B2
Abstract

A hardware failure prediction and avoidance system executing on a unified endpoint management platform information handling system comprising a network interface device (NID) to receive failed operational telemetries for client devices including power and software analytics, and event viewer error logs, and an error associated with a hardware type, a processor to execute a classification supervised learning algorithm on the failed operational telemetries to determine that a hardware or software configuration will likely co-occur with a future error of the hardware type, identify the hardware or software configuration as problematic, the processor to identify the problematic system configuration in a current operational telemetry by correlating the current operational telemetry with a portion of the failed operational telemetries, and the NID to transmit a recommendation to the first client device to adjust the problematic system configuration in order to avoid occurrence of the error at the first client device.

Claims (40)

1. A hardware failure prediction and avoidance system executing on a unified endpoint management (UEM) platform information handling system comprising:

a network interface device to receive failed operational telemetries for a plurality of client information handling systems including power analytics, software application analytics, and event viewer error logs, wherein each failed operational telemetry includes an error associated with a labeled hardware type;

a processor executing code instructions of the hardware failure prediction and avoidance system to:

execute a classification supervised learning algorithm on the failed operational telemetries to determine a failure probability, meeting a preset failure probability threshold value, that an adjustable hardware or software configuration described by future operational telemetries will co-occur with the error of the labeled hardware type;

identify the adjustable hardware or software configuration as a problematic system configuration;

the network interface device to receive a current operational telemetry for a first client information handling system including current power analytics, current software application analytics, and current event viewer error logs;

the processor to identify the problematic system configuration by correlating the current operational telemetry with a portion of the failed operational telemetries such that the failure probability meets the preset failure probability threshold value; and

the network interface device to transmit a recommendation to the first client information handling system to adjust or terminate the problematic system configuration in order to avoid occurrence of the error identified in the failed operational telemetries at the first client information handling system.

2. The information handling system of claim 1 , wherein the adjustable hardware or software configuration includes hardware policy settings.

3. The information handling system of claim 1 , wherein the adjustable hardware or software configuration includes a performance mode for a hardware component.

4. The information handling system of claim 1 , wherein the adjustable hardware or software configuration includes a power conservation setting for a hardware component.

5. The information handling system of claim 1 , wherein the adjustable hardware or software configuration includes a limitation of hardware component resources consumed during execution of background software applications.

6. The information handling system of claim 1 , wherein the labeled hardware type is a type of processing unit.

7. The information handling system of claim 1 , wherein the hardware type is a type of network interface device.

8. A method of hardware failure prediction and avoidance comprising:

receiving, via a network interface device, failed operational telemetries for a plurality of client information handling systems including power analytics, software application analytics, and event viewer error logs, wherein each failed operational telemetry includes an error associated with a labeled hardware type;

applying a classification supervised learning algorithm to the failed operational telemetries, via a processor, to determine a failure probability, meeting a preset failure probability threshold value, that an adjustable hardware or software configuration described by future operational telemetries will co-occur with the error of the labeled hardware type;

identifying the adjustable hardware or software configuration as a problematic system configuration;

receiving a current operational telemetry for a first client information handling system including current power analytics, current software application analytics, and current event viewer error logs;

identifying the problematic system configuration within the current operational telemetry as correlated with at least a portion of the failed operational telemetry such that the failure probability of the adjustable hardware or software configuration that is identified as the problematic system configuration meets the preset failure probability threshold value; and

transmitting a recommendation to the first client information handling system to adjust or terminate the problematic system configuration in order to avoid occurrence of the error identified in the failed operational telemetries at the first client information handling system.

9. The method of claim 8 , wherein the hardware type is a type of memory device.

10. The method of claim 8 , wherein the hardware type is a type of cooling unit.

11. The method of claim 8 , wherein the adjustable hardware or software configuration includes background software application usage.

12. The method of claim 8 , wherein the adjustable hardware or software configuration includes software or firmware update settings.

13. The method of claim 8 , wherein the adjustable hardware or software configuration includes a version of software or firmware installed.

14. The method of claim 8 , wherein the adjustable hardware or software configuration includes repeated initialization of a process.

15. A hardware failure prediction and avoidance system executing on a unified endpoint management (UEM) platform information handling system comprising:

a network interface device to receive failed operational telemetries for a plurality of client information handling systems including power analytics, software application analytics, and event viewer error logs, wherein each failed operational telemetry includes an error associated with a labeled hardware type;

a processor executing code instructions of the hardware failure prediction and avoidance system to:

execute a classification supervised learning algorithm on the failed operational telemetries to determine a failure probability, meeting a preset failure probability threshold value, that a plurality of adjustable hardware or software configurations described by future operational telemetries will co-occur with the error of the labeled hardware type;

identify the plurality of adjustable hardware or software configurations as a problematic system configuration combination;

the network interface device to receive a current operational telemetry for a first client information handling system including current power analytics, current software application analytics, and current event viewer error logs;

the processor to identify the problematic system configuration combination corresponding to at least a portion of the failed operational telemetries within the current operational telemetry such that the failure probability of the identified problematic system configuration combination meets the preset failure probability threshold value; and

the network interface device to transmit a recommendation to the first client information handling system to adjust or terminate one or more of the plurality of adjustable hardware or software configurations within the problematic system configuration combination in order to avoid occurrence of the error identified in the failed operational telemetries at the first client information handling system.

16. The information handling system of claim 15 , wherein the plurality of adjustable hardware or software configurations within the problematic system configuration includes consumption of power by a fan above a preset maximum fan power draw threshold value, and usage of a specific software application.

17. The information handling system of claim 15 , wherein the plurality of adjustable hardware or software configurations within the problematic system configuration includes consumption of memory resources above a preset maximum memory resource threshold value, and usage of a specific software application.

18. The information handling system of claim 15 , wherein the plurality of adjustable hardware or software configurations within the problematic system configuration includes consumption of processor resources above a preset maximum processor resource threshold value, and usage of a specific software application.

19. The information handling system of claim 15 , wherein the plurality of adjustable hardware or software configurations within the problematic system configuration includes consumption of memory resources above a preset maximum memory resource threshold value, and repeated initialization of a specific software application.

20. The information handling system of claim 15 , wherein the plurality of adjustable hardware or software configurations within the problematic system configuration includes consumption of processor resources above a preset maximum processor resource threshold value, and execution of a background software application.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jul 29, 2022
From: AURONGZEB, DEEDER M.; RAMAKRISHNAN, MALATHI; SETHI, PARMINDER SINGH
To: DELL PRODUCTS, LP
Reel/Frame 060672/0253 →
Continuity (1)
Related Publication 20240036999A1 · Feb 1, 2024