IP Library › Granted Patent US 12,724,659
Granted Patent B2
US 12,724,659 · App. 18/417,429 · Granted Sep 1, 2026

Interactive data processing system failure management using hidden knowledge from predictive models

Inventors: Deepaganesh Paulraj (Bangalore, IN); Min Gong (Shanghai, CN); Ashok Narayanan Potti (Bangalore, IN); Dale Wang (Hayward, CA)
Assignee: Dell Products L.P.
G06F11/079G06F11/0793G06F11/3476
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,724,659
App. No.
18/417,429
Filed
Jan 19, 2024
Granted
Sep 1, 2026
Kind
B2
Art Unit
2113
USPC
714/37
Abstract

Methods and systems for managing data processing systems are disclosed. A data processing system may include and depend on the operation of hardware and/or software components. Inference models may be implemented to predict future system infrastructure outcomes (e.g., component failures) using information recorded in logs that reflect the operation of the components. However, the models may be complex “black boxes” and may generate critical outcome predictions for downstream consumers without explanations of how the predictions are determined, resulting in downstream consumers having low confidence in the predictions. Therefore, hidden knowledge (e.g., structured knowledge attributes) of the models may be extracted and/or used to understand the underlying processes that the models use to predict the system infrastructure outcomes. The hidden knowledge may be provided for interactively managing data processing system(s) failures in order to increase the likelihood of preventing and/or mitigating future data processing system failures.

Claims (69)

1 . A method for managing failures of data processing systems, comprising and by a data processing system manager configured to manage the data processing systems:

obtaining a data request, from a requestor, for data stored in a structured knowledge repository, the data comprising structured knowledge attributes extracted from an architecture of a trained machine learning model, wherein the architecture of the trained machine learning model is hidden;

determining, based on the data request, a type of components of the data processing system;

filtering the structured knowledge attributes based on the type of components to obtain filtered structured knowledge attributes;

generating one or more customized user response prompts using the data and generic response prompts stored in a sample prompt repository;

obtaining a response to the data request using the one or more customized user response prompts;

providing the response to the requestor, through an interactive user interface through which the data request was received, to service the data request, wherein the response comprises a failure prediction and a portion of the filtered structured knowledge attributes that provide visibility for the requestor into understanding how and why the trained machine learning model generated the failure prediction in a manner that the trained machine learning model generated the failure prediction without the requestor having direct accessibility to the architecture, the failure prediction being one of inferences generated by the machine learning model, wherein the failure prediction is associated with the type of components; and

troubleshooting a component having the type of components according to the failure prediction.

2 . The method of claim 1 , wherein the structured knowledge attributes are usable to manage an indication of failure for a data processing system of the data processing systems.

3 . The method of claim 2 , wherein the one or more customized user response prompts are generated using few shot learning techniques.

4 . The method of claim 3 , further comprising by the data processing system manager:

refining the data request to obtain a refined data request, wherein the refining comprises:

obtaining a user intention from the data request; and

refining the data request based on the user intention and the structured knowledge attributes stored in the structured knowledge repository,

wherein the one or more customized user response prompts is further generated using the refined data request.

5 . The method of claim 4 , further comprising by the data processing system manager:

obtaining user preference data from a local domain context repository, wherein the user preference data is associated with the requestor,

wherein the one or more customized user response prompts is further generated using the user preference data.

6 . The method of claim 2 , further comprising by the data processing system manager:

prior to generating the response:

identifying an occurrence of the indication of failure for the data processing system; and

based on the occurrence, using an inference model to obtain an indication of a root cause for the failure, the structured knowledge repository being based, at least in part, on the inference model and logs on which the inference model is based, the inference model being the trained machine learning model that generates the failure prediction, and the indication of the root cause being specified in the failure prediction.

7 . The method of claim 6 , further comprising by the data processing system manager:

after providing the response:

assessing a likelihood of the root cause being accurate; and

in an instance of the assessing where the likelihood meets a threshold:

identifying at least one remediation action based on the root cause; and

performing the at least one remediation action to obtain an updated data processing system to attempt to remediate the failure.

8 . A non-transitory machine-readable medium having instructions stored therein, which when executed by a processor of a data processing system manager configured to manage data processing systems, cause the data processing system manager to perform operations for managing failures of the data processing systems, the operations comprising:

obtaining a data request, from a requestor, for data stored in a structured knowledge repository;

generating one or more customized user response prompts using the data and generic response prompts stored in a sample prompt repository, the data comprising structured knowledge attributes extracted from an architecture of a trained machine learning model, wherein the architecture of the trained machine learning model is hidden,

determining, based on the data request, a type of components of the data processing systems;

filtering the structured knowledge attributes based on the type of components to obtain filtered structured knowledge attributes;

obtaining a response to the data request using the one or more customized user response prompts;

providing the response to the requestor, through an interactive user interface through which the data request was received, to service the data request, wherein the response comprises a failure prediction and a portion of the filtered structured knowledge attributes that provide visibility for the requestor into understanding how and why the trained machine learning model generated the failure prediction in a manner that the trained machine learning model generated the failure prediction without the requestor having direct accessibility to the architecture, the failure prediction being one of inferences generated by the machine learning model, wherein the failure prediction is associated with the type of components; and

troubleshoot a component having the type of components according to the failure prediction.

9 . The non-transitory machine-readable medium of claim 8 , wherein the structured knowledge attributes are usable to manage an indication of failure for a data processing system of the data processing systems.

10 . The non-transitory machine-readable medium of claim 9 , wherein the one or more customized user response prompts are generated using few shot learning techniques.

11 . The non-transitory machine-readable medium of claim 10 , wherein the operations further comprise:

refining the data request to obtain a refined data request, wherein the refining comprises:

obtaining a user intention from the data request; and

refining the data request based on the user intention and the structured knowledge attributes stored in the structured knowledge repository,

wherein the one or more customized user response prompts is further generated using the refined data request.

12 . The non-transitory machine-readable medium of claim 11 , wherein the operations further comprise:

obtaining user preference data from a local domain context repository, wherein the user preference data is associated with the requestor,

wherein the one or more customized user response prompts is further generated using the user preference data.

13 . A data processing system manager, comprising:

a processor; and

a memory coupled to the processor to store instructions, which when executed by the processor, cause the data processing system manager to perform operations for managing failures of data processing systems, the operations comprising:

obtaining a data request, from a requestor, for data stored in a structured knowledge repository, the data comprising structured knowledge attributes extracted from an architecture of a trained machine learning model, wherein the architecture of the trained machine learning model is hidden;

determining, based on the data request, a type of components of the data processing system;

filtering the structured knowledge attributes based on the type of components to obtain filtered structured knowledge attributes;

generating one or more customized user response prompts using the data and generic response prompts stored in a sample prompt repository;

obtaining a response to the data request using the one or more customized user response prompts;

providing the response to the requestor, through an interactive user interface through which the data request was received, to service the data request, wherein the response comprises a failure prediction and a portion of the filtered structured knowledge attributes that provide visibility for the requestor into understanding how and why the trained machine learning model generated the failure prediction in a manner that the trained machine learning model generated the failure prediction without the requestor having direct accessibility to the architecture, the failure prediction being one of inferences generated by the machine learning model, wherein the failure prediction is associated with the type of components; and

troubleshoot a component having the type of components according to the failure prediction.

14 . The data processing system manager of claim 13 , wherein the structured knowledge attributes are usable to manage an indication failure for a data processing system of the data processing systems.

15 . The data processing system manager of claim 14 , wherein the one or more customized user response prompts are generated using few shot learning techniques.

16 . The data processing system manager of claim 15 , wherein the operations further comprise:

refining the data request to obtain a refined data request, wherein the refining comprises:

obtaining a user intention from the data request; and

refining the data request based on the user intention and the structured knowledge attributes stored in the structured knowledge repository,

wherein the one or more customized user response prompts is further generated using the refined data request.

17 . The method of claim 1 , further comprising and by the data processing system manager prior to obtaining the data request:

extracting the structured knowledge attributes from the architecture of the trained machine learning model, the trained machine learning model being hosted by the data processing system manager; and

storing the structured knowledge attributes extracted from the architecture of the trained machine learning model into the structured knowledge repository.

18 . The method of claim 1 , wherein the architecture of the trained machine learning model comprises at least information regarding one or more other trained machine learning models that were used to train the trained machine learning model and regarding respective other architectures of each of the one or more other trained machine learning models.

19 . The method of claim 1 , wherein the one or more structured knowledge attributes comprise parameters on which the trained machine learning model is trained to generate the inferences.

20 . The method of claim 19 , wherein the one or more structured knowledge attributes further comprise relationships between components of the trained machine learning model, the components comprising at least input features of data ingested into the trained machine learning model, the inferences generated by the trained machine learning model, and one or more rules followed by the trained machine learning model to generate the inferences.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 14, 2024
From: PAULRAJ, DEEPAGANESH; GONG, MIN; POTTI, ASHOK NARAYANAN; WANG, DALE
To: DELL PRODUCTS L.P.
Reel/Frame 066460/0659 →
Continuity (1)
Related Publication 20250238303A1 · Jul 24, 2025
References Cited (89)
US 7792774B2 · Friedlander · 2010 [cited by applicant]
US 8538897B2 · Han et al. · 2013 [cited by applicant]
US 10235227B2 · Purushothaman · 2019 [cited by applicant]
US 10572329B2 · Harutyunyan et al. · 2020 [cited by applicant]
US 10616314B1 · Plenderleith et al. · 2020 [cited by applicant]
US 10740793B1 · Sussman et al. · 2020 [cited by applicant]
US 10776196B2 · Ohana et al. · 2020 [cited by applicant]
US 10853867B1 · Bulusu et al. · 2020 [cited by applicant]
US 11176015B2 · Rallapalli · 2021 [cited by applicant]
US 11461163B1 · Aggarwal · 2022 [cited by applicant]
US 11513930B2 · Chan et al. · 2022 [cited by applicant]
US 11514347B2 · Dinh · 2022 [cited by applicant]
US 11714721B2 · Ehsan · 2023 [cited by applicant]
US 11720940B2 · Lakshminarayan et al. · 2023 [cited by applicant]
US 11734102B1 · Wang et al. · 2023 [cited by applicant]
US 11748185B2 · Xu et al. · 2023 [cited by applicant]
US 11789846B2 · Loeb · 2023 [cited by applicant]
US 11803440B2 · Harutyunyan · 2023 [cited by applicant]
US 11809271B1 · Wang · 2023 [cited by applicant]
US 11853187B1 · Roche · 2023 [cited by applicant]
US 11909836B2 · Wulf et al. · 2024 [cited by applicant]
US 12061970B1 · Lo et al. · 2024 [cited by applicant]
US 20050033761A1 · Guttman · 2005 [cited by applicant]
US 20060168195A1 · Maturana · 2006 [cited by applicant]
US 20090113248A1 · Bock et al. · 2009 [cited by applicant]
US 20090216910A1 · Duchesneau · 2009 [cited by applicant]
US 20100257058A1 · Karidi et al. · 2010 [cited by applicant]
US 20100318856A1 · Yoshida · 2010 [cited by applicant]
US 20130041748A1 · Hsiao et al. · 2013 [cited by applicant]
US 20130198240A1 · Ameri-Yahia et al. · 2013 [cited by applicant]
US 20140310222A1 · Davlos · 2014 [cited by examiner]
US 20150227838A1 · Wang et al. · 2015 [cited by applicant]
US 20150288557A1 · Gates et al. · 2015 [cited by applicant]
US 20180205645A1 · Bays · 2018 [cited by applicant]
US 20190095313A1 · Xu et al. · 2019 [cited by applicant]
US 20190129785A1 · Liu et al. · 2019 [cited by applicant]
US 20190244122A1 · Li · 2019 [cited by examiner]
US 20200026590A1 · Lopez et al. · 2020 [cited by applicant]
US 20200174870A1 · Xu · 2020 [cited by applicant]
US 20200401397A1 · Thomas · 2020 [cited by applicant]
US 20210027205A1 · Sevakula et al. · 2021 [cited by applicant]
US 20210133607A1 · Stubbs · 2021 [cited by applicant]
US 20210241141A1 · Dugger et al. · 2021 [cited by applicant]
US 20210241152A1 · Fong · 2021 [cited by applicant]
US 20210287109A1 · Cmielowski et al. · 2021 [cited by applicant]
US 20220050733A1 · Selvaraju · 2022 [cited by applicant]
US 20220100187A1 · Isik et al. · 2022 [cited by applicant]
US 20220171991A1 · Das · 2022 [cited by applicant]
US 20220188181A1 · Saha · 2022 [cited by applicant]
US 20220237101A1 · Singh · 2022 [cited by applicant]
US 20220283890A1 · Chopra et al. · 2022 [cited by applicant]
US 20220358005A1 · Saha et al. · 2022 [cited by applicant]
US 20220382611A1 · Kapish · 2022 [cited by applicant]
US 20220391300A1 · Trapani · 2022 [cited by examiner]
US 20220417078A1 · Matsuo et al. · 2022 [cited by applicant]
US 20230016199A1 · Jividen et al. · 2023 [cited by applicant]
US 20230094373A1 · Muralidharan · 2023 [cited by applicant]
US 20230099001A1 · Harutyunyan · 2023 [cited by applicant]
US 20230214285A1 · Arumugam Maharaja · 2023 [cited by applicant]
US 20230325468A1 · Srinivasan · 2023 [cited by applicant]
US 20240020191A1 · Harutyunyan · 2024 [cited by applicant]
US 20240028955A1 · Harutyunyan et al. · 2024 [cited by applicant]
US 20240152442A1 · Sethi · 2024 [cited by applicant]
US 20240168835A1 · Wang et al. · 2024 [cited by applicant]
US 20240177026A1 · Ezrielev · 2024 [cited by applicant]
US 20240311224A1 · Paulraj · 2024 [cited by applicant]
US 20240311631A1 · Bycroft · 2024 [cited by applicant]
US 20240362097A1 · Paulraj · 2024 [cited by applicant]
US 20240411752A1 · Prabhakar · 2024 [cited by applicant]
US 20240428128A1 · Kim · 2024 [cited by applicant]
US 20250061042A1 · Krishnan · 2025 [cited by applicant]
US 20250086211A1 · Bolcer · 2025 [cited by examiner]
US 20250139374A1 · Cui · 2025 [cited by applicant]
US 20250147754A1 · Kholodkov · 2025 [cited by applicant]
US 20250147832A1 · Agrawal · 2025 [cited by applicant]
US 20250190763A1 · Banuelos · 2025 [cited by applicant]
US 20250238303A1 · Paulraj · 2025 [cited by applicant]
US 20250238306A1 · Paulraj · 2025 [cited by examiner]
US 20250378766A1 · Naufel · 2025 [cited by applicant]
CN 108280168A · 2018 [cited by applicant]
CN 111476371A · 2020 [cited by applicant]
CN 112541806A · 2021 [cited by applicant]
EP 4235505A1 · 2023 [cited by applicant]
Zhao, Wayne Xin, et al., “A Survey of Large Language Models,” arXiv preprint arXiv:2303.18223 (2023) (97 Pages). [cited by applicant]
Kaddour, Jean, et al., “Challenges and Applications of Large Language Models,” arXiv preprint arXiv:2307.10169 (2023) (72 Pages). [cited by applicant]
Naveed, Humza, et al., “A Comprehensive Overview of Large Language Models,” arXiv preprint arXiv:2307.06435 (2023) (35 Pages). [cited by applicant]
Boffa, Matteo, et al., “LogPrécis: Unleashing Language Models for Automated Shell Log Analysi,” arXiv preprint arXiv:2307.08309 (2023) (17 Pages). [cited by applicant]
Chen, Yinfang, et al., “Empowering Practical Root Cause Analysis by Large Language Models for Cloud Incidents,” arXiv preprint arXiv:2305.15778 (2023) (15 Pages). [cited by applicant]
Lee, Yukyung, et al., “LAnoBERT : System Log Anomaly Detection based on BERT Masked Language Model,” Applied Soft Computing 146 (2023): 110689 (18 Pages). [cited by applicant]