IP Library Granted Patent US 7,302,607
Granted Patent B2
US 7,302,607 · App. 10/652,509 · Granted Nov 27, 2007

Two node virtual shared disk cluster recovery

View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 7,302,607
App. No.
10/652,509
Granted
Nov 27, 2007
Kind
B2
Abstract

A method for recovery in a two-node data processing system is provided wherein each node is a primary server for a first nonvolatile storage device and for which there is provided shared access to a second nonvolatile storage device for which the other node is a primary server and wherein each node also includes a direct connection to the shared nonvolatile storage device for which the other node is the primary server. Upon notification of failure, the method operates by first confirming continued access by each node to the nonvolatile storage device for which it is the primary server and then by attempting to access the shared nonvolatile storage device via the direct connection and by waiting for a time sufficient for the same process to be carried out by the other node. If access to the shared nonvolatile storage device is successful, the node takes control of both nonvolatile storage devices. If the access is not successful a comparison of node numbers is carried out to decide the issue of control. Whenever a node determines that it does not have access to the storage device for which it is the primary server, it shuts down recovery at the node.

Claims (11)

1. A method for recovery in a two-node data processing system wherein each node is a primary server for a first disk drive and for which there is provided shared access to a second disk for which the other node is a primary server and wherein each node also includes a direct connection to the shared disk for which the other node is the primary server, said method comprising the steps of:

receiving, at a first node, notification of communication failure with said second node;

determining if said first node has access to the disk for which said first node is the primary server;

shutting down recovery at said first node if said access is not present, but if it is present, accessing the disk for which said second node is the primary server via said hardware connection and waiting for a period of time sufficient to assure that recovery processes at said other node have completed past the same point as said first node;

determining if said first node still has access to the disk for which it is the primary server and if said first node still has access, taking over control of the second node's disk; and

if said first node doesn't have said access per said immediately preceding determining step, comparing node numbers to decide which node controls the other node's disks.

2. The method of claim 1 in which said comparing also determines that the node that does not control the other node's disk shuts down its recovery process.

3. The method of claim 1 in which said nodes communicate via communication adapters at each node wherein said adapters connect the nodes through a switch.

4. The method of claim 3 in which said notification is transmitted through an adapter.

5. The method of claim 1 in which said period of time is greater than about:

(process swap time+CPU time slice+time taken to break reservation) *2.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 28, 2021
From: INTERNATIONAL BUSINESS MACHINES CORPORATION
To: MAPLEBEAR INC.
Reel/Frame 055155/0943 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Dec 31, 2003
From: GUNDA, KALYAN C.; HERR, BRIAN D.
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 014229/0870 →