IP Library Granted Patent US 11,438,249
Granted Patent B2
US 11,438,249 · App. 17/224,874 · Granted Sep 6, 2022

Cluster management method, apparatus and system

Inventor: Lin Cheng (Hangzhou, CN)
Assignee: Alibaba Group Holding Limited
H04L43/0817H04L41/0627H04L41/0856H04L41/0863H04L41/22H04L43/0823
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 11,438,249
App. No.
17/224,874
Granted
Sep 6, 2022
Kind
B2
Abstract

A method including acquiring an operation request issued by a distributed consistency system in a cluster, and determining deciding information for processing the operation request and sending the deciding information to an operation and maintenance control platform, wherein the deciding information is determined based on data consistency and availability of the distributed consistency system. The present disclosure solves the technical problems of data loss or service interruption in the distributed consistency system, which are caused by the higher error rate of manual operations used in cluster management methods for distributed consistency systems.

Claims (68)

1. A method comprising:

acquiring an operation request issued by a distributed consistency system in a cluster;

determining deciding information for processing the operation request, the determining the deciding information for processing the operation request including:

determining an operation type corresponding to the operation request; and

determining the deciding information for executing an operation corresponding to the operation type, the determining the deciding information for executing the operation corresponding to the operation type including:

before selecting a host from the cluster for upgrading, determining that a service of a previous host before the host is upgraded meets an availability condition;

selecting the host from the cluster for upgrading in response to determining that the operation type is an upgrading service in the distributed consistency system; and

determining to upgrade the selected host in response to determining that the service meets the availability condition; and

sending the deciding information to an operation and maintenance control platform.

2. The method of claim 1 , wherein the deciding information for processing the operation request is determined based on data consistency and availability of the distributed consistency system.

3. The method according to claim 1 , wherein the deciding information for executing the operation corresponding to the operation type is determined based on the availability condition and a security condition of the distributed consistency system.

4. The method according to claim 1 , further comprising:

in response to determining that another operation type corresponding to another operation request is upgrading configuration information of the distributed consistency system, determining to permit an execution for an operation corresponding to the another operation request.

5. The method according to claim 1 ,

further comprising:

acquiring a serial number of a newly added host in response to determining that another operation type corresponding to another operation request is replacing a host deployed by the distributed consistency system; and

establishing an association between the serial number of the newly added host and serial numbers of hosts not having been replaced in the distributed consistency system to form a new distributed consistency system.

6. The method according to claim 1 , further comprising:

stopping an external service of the distributed consistency system in response to determining that another operation type corresponding to another operation request is replacing a disk used by a designated host in the distributed consistency system;

acquiring log information and snapshot data from another host in the distributed consistency system; and

resuming the external service after storing the log information and the snapshot data to a newly added disk.

7. An apparatus comprising:

one or more processors; and

one or more computer readable media storing thereon computer-readable instructions that, when executed by the one or more processors, cause the one or more processors to perform acts comprising:

acquiring an operation request issued by a distributed consistency system in a cluster;

determining deciding information for processing the operation request, the determining the deciding information for processing the operation request including:

determining an operation type corresponding to the operation request; and

determining the deciding information for executing an operation corresponding to the operation type, the determining the deciding information for executing the operation corresponding to the operation type including:

before selecting a host from the cluster for upgrading, determining that a service of a previous host before the host is upgraded meets an availability condition;

selecting the host from the cluster for upgrading in response to determining that the operation type is an upgrading service in the distributed consistency system; and

determining to upgrade the selected host in response to determining that the service meets the availability condition; and

sending the deciding information to an operation and maintenance control platform.

8. The apparatus according to claim 7 , wherein the deciding information is determined based on data consistency and availability of the distributed consistency system.

9. The apparatus according to claim 7 , wherein the deciding information for executing the operation corresponding to the operation type is determined based on the availability condition and a security condition of the distributed consistency system.

10. The apparatus according to claim 7 , wherein the acts further comprise:

in response to determining that another operation type corresponding to another operation request is upgrading configuration information of the distributed consistency system, determining to permit an execution for an operation corresponding to the another operation request.

11. The apparatus according to claim 7 , wherein the acts further comprise:

acquiring a serial number of a newly added host in response to determining that another operation type corresponding to another operation request is replacing a host deployed by the distributed consistency system; and

establishing an association between the serial number of the newly added host and serial numbers of hosts not having been replaced in the distributed consistency system to form a new distributed consistency system.

12. The apparatus according to claim 7 , wherein the acts further comprise:

stopping an external service of the distributed consistency system in response to determining that another operation type corresponding to another operation request is replacing a disk used by a designated host in the distributed consistency system;

acquiring log information and snapshot data from another host in the distributed consistency system; and

resuming the external service after storing the log information and the snapshot data to a newly added disk.

13. One or more memories storing thereon computer-readable instructions that, when executed by one or more processors, cause the one or more processors to perform acts comprising:

issuing an operation request to a distributed consistency system in a cluster and display an operating state of the distributed consistency system; and

determining deciding information for processing the operation request and sends the deciding information to an operation and maintenance control platform, wherein the deciding information for processing the operation request is determined based on data consistency and availability of the distributed consistency system, the determining the deciding information for processing the operation request including:

determining an operation type corresponding to the operation request; and

determining the deciding information for executing an operation corresponding to the operation type, wherein the deciding information for executing the operation corresponding to the operation type is determined based on an availability condition and a security condition of the distributed consistency system, the determining the deciding information for executing the operation corresponding to the operation type including:

before selecting a host from the cluster for upgrading, determining that a service of a previous host before the host is upgraded meets the availability condition,

selecting the host from the cluster for upgrading in response to determining that the operation type is an upgrading service in the distributed consistency system, and

determining to upgrade the selected host in response to determining that the service meets the availability condition.

14. The one or more memories according to claim 13 , wherein the acts further comprise:

acquiring monitoring data in a respective host of the cluster; and

collecting the monitoring data.

15. The one or more memories according to claim 13 , wherein the acts further comprise displaying the monitoring data.

16. The one or more memories according to claim 13 , wherein the acts further comprise:

in response to determining that another operation type corresponding to another operation request is upgrading configuration information of the distributed consistency system, determining to permit an execution of an operation corresponding to the another operation request.

17. The one or more memories according to claim 13 , wherein the acts further comprise:

acquiring a serial number of a newly added host in response to determining that another operation type corresponding to another operation request is replacing a host deployed by the distributed consistency system; and

establishing an association between the serial number of the newly added host and serial numbers of hosts not having been replaced in the distributed consistency system to form a new distributed consistency system.

18. The one or more memories according to claim 13 , wherein the acts further comprise:

stopping an external service of the distributed consistency system in response to determining that another operation type corresponding to another operation request is replacing a disk used by a designated host in the distributed consistency system;

acquiring log information and snapshot data from another host in the distributed consistency system; and

resuming the external service after storing the log information and the snapshot data to a newly added disk.

19. The one or more memories according to claim 14 , wherein the monitoring data relates to coordinating and processing user requests in the respective host of the cluster.

20. The one or more memories according to claim 14 , wherein the acts further comprise:

generating alarm information according to the monitoring data; and

sending the alarm information to a user-side device.

Assignments (2)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Apr 29, 2026
From: ALIBABA GROUP HOLDING LIMITED
To: CLOUD INTELLIGENCE ASSETS HOLDING (SINGAPORE) PRIVATE LIMITED
Reel/Frame 075499/0384 →
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded May 19, 2021
From: CHENG, LIN
To: ALIBABA GROUP HOLDING LIMITED
Reel/Frame 056290/0118 →
Priority Claims (1)
CN 201811168317.0 · Oct 8, 2018 · national
Continuity (2)
Continuation PCTCN2019108367 · Sep 27, 2019
Related Publication 20210226871A1 · Jul 22, 2021