IP Library Granted Patent US 11,544,149
Granted Patent B2
US 11,544,149 · App. 17/142,995 · Granted Jan 3, 2023

System and method for improved fault tolerance in a network cloud environment

Inventor: Parthasarathy Srinivasan (American Fork, UT)
Assignee: ORACLE INTERNATIONAL CORPORATION
G06F11/1423G06F9/5072G06F9/5077G06F11/1484G06F11/3006
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,544,149
App. No.
17/142,995
Granted
Jan 3, 2023
Kind
B2
Abstract

Described herein are systems and methods for fault tolerance in a network cloud environment. In accordance with various embodiments, the present disclosure provides an improved fault tolerance solution, and improvement in the fault tolerance of systems, by way of failure prediction, or prediction of when an underlying infrastructure will fail, and using the predictions to counteract the failure by spinning up or otherwise providing new component pieces to compensate for the failure.

Claims (38)

1. A system for fault tolerance in a network cloud environment, comprising:

a network cloud environment having deployed thereon an application utilizing a plurality of sockets of the network cloud environment;

a fault tolerance module of the network cloud, the fault tolerance module comprising a monitoring component and a pattern recognition component;

wherein the monitoring component monitors and records a plurality of faults at the plurality of sockets, each fault being time-stamped;

wherein the monitoring component outputs a representation of the plurality of faults to the pattern recognition module;

wherein the pattern recognition module builds, based upon the representation of the plurality of faults, a failure model;

wherein the pattern recognition module builds a number of new sockets for use by the application, the number of new sockets being based upon the failure model; and

wherein the fault tolerance module further comprises a socket pool maintainer module, wherein the socket pool maintainer module opens a set of the built new sockets based upon a difference between a set of plurality of sockets currently available and a prediction of a number of sockets needed based upon the failure model.

2. The system of claim 1 , wherein the fault tolerance module further comprises a read socket consumer module, wherein the ready socket consumer module provides the opened set of the built new sockets to the application.

3. The system of claim 2 , wherein the socket pool maintainer module runs in a while loop while the application is running.

4. The system of claim 3 , wherein the pattern recognition module runs on a separate virtual machine from the rest of the monitoring component.

5. The system of claim 4 , wherein the application comprises a distributed data grid application.

6. The system of claim 4 , wherein the application comprises an application server.

7. A method for fault tolerance in a network cloud environment, comprising:

providing a network cloud environment having deployed thereon an application utilizing a plurality of sockets of the network cloud environment;

providing a fault tolerance module of the network cloud, the fault tolerance module comprising a monitoring component and a pattern recognition component;

monitoring and recording, by the monitoring component, a plurality of faults at the plurality of sockets, each fault being time-stamped;

outputting, by the monitoring component, a representation of the plurality of faults to the pattern recognition module;

building, by the pattern recognition module and based upon the representation of the plurality of faults, a failure model; and

building, by the pattern recognition module, a number of new sockets for use by the application, the number of new sockets being based upon the failure model;

wherein the fault tolerance module further comprises a socket pool maintainer module, wherein the socket pool maintainer module opens a set of the built new sockets based upon a difference between a set of plurality of sockets currently available and a prediction of a number of sockets needed based upon the failure model.

8. The method of claim 7 , wherein the fault tolerance module further comprises a read socket consumer module, wherein the ready socket consumer module provides the opened set of the built new sockets to the application.

9. The method of claim 8 , wherein the socket pool maintainer module runs in a while loop while the application is running.

10. The method of claim 9 , wherein the pattern recognition module runs on a separate virtual machine from the rest of the monitoring component.

11. The method of claim 10 , wherein the application comprises a distributed data grid application.

12. The method of claim 10 , wherein the application comprises an application server.

13. A non-transitory computer readable storage medium having instructions thereon for fault tolerance in a network cloud environment, which when read an executed by a computer cause the computer to perform steps comprising:

providing a network cloud environment having deployed thereon an application utilizing a plurality of sockets of the network cloud environment;

providing a fault tolerance module of the network cloud, the fault tolerance module comprising a monitoring component and a pattern recognition component;

monitoring and recording, by the monitoring component, a plurality of faults at the plurality of sockets, each fault being time-stamped;

outputting, by the monitoring component, a representation of the plurality of faults to the pattern recognition module;

building, by the pattern recognition module and based upon the representation of the plurality of faults, a failure model; and

building, by the pattern recognition module, a number of new sockets for use by the application, the number of new sockets being based upon the failure model;

wherein the fault tolerance module further comprises a socket pool maintainer module, wherein the socket pool maintainer module opens a set of the built new sockets based upon a difference between a set of plurality of sockets currently available and a prediction of a number of sockets needed based upon the failure model.

14. The non-transitory computer readable storage medium of claim 13 wherein the fault tolerance module further comprises a read socket consumer module, wherein the ready socket consumer module provides the opened set of the built new sockets to the application.

15. The non-transitory computer readable storage medium of claim 14 , wherein the socket pool maintainer module runs in a while loop while the application is running.

16. The non-transitory computer readable storage medium of claim 15 , wherein the pattern recognition module runs on a separate virtual machine from the rest of the monitoring component.

17. The non-transitory computer readable storage medium of claim 16 , wherein the application comprises a distributed data grid application.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 2, 2021
From: SRINIVASAN, PARTHASARATHY
To: ORACLE INTERNATIONAL CORPORATION
Reel/Frame 055118/0991 →
Continuity (3)
Provisional Application 63000097 · Mar 26, 2020
Provisional Application 62957976 · Jan 7, 2020
Related Publication 20210208971A1 · Jul 8, 2021
Cited By (1)
US 12,250,267