System for pause and play dynamic distributed computing
View Patent ↗A distributed computing system includes a memory storing execution state collected prior to an operator pause state. A controller modifies, during the operator pause state, characteristics of the distributed computing system. The controller invokes execution of the operator after the pause state such that the operator accesses the execution state to complete an operation.
1. A distributed computing system, comprising:
a memory storing execution state collected prior to an operator pause state invoked in response to a violation of a service level agreement;
a controller to modify in response to the violation of the service level agreement, during the operator pause state, characteristics of the distributed computing system, wherein the controller invokes execution of the operator after the pause state such that the operator accesses the execution state to complete an operation.
2. The distributed computing system of claim 1 wherein the controller collects data from a plurality of nodes in the distributed computing system characterizing violations of service level agreements, wherein each service level agreement specifies an operating condition threshold within the distributed computing system.
3. The distributed computing system of claim 2 wherein the controller initiates protocols to cure violations of service level agreements.
4. The distributed computing system of claim 3 wherein the protocols include a key processing flood restart protocol.
5. The distributed computing system of claim 3 wherein the protocols include a key processing flood repair protocol.
6. The distributed computing system of claim 5 wherein the controller splits a key.
7. The distributed computing system of claim 3 wherein the protocols include a key processing flood relocation protocol.
8. The distributed computing system of claim 7 wherein the controller moves a key to a different partition.
9. The distributed computing system of claim 7 wherein the controller moves a key to a new computing node.
10. The distributed computing system of claim 3 wherein the protocols include a partition processing flood key restart protocol.
11. The distributed computing system of claim 3 wherein the protocols include a partition processing flood repair protocol.
12. The distributed computing system of claim 11 wherein the controller splits a data partition.
13. The distributed computing system of claim 3 wherein the protocols include a partition processing flood relocation protocol.
14. The distributed computing system of claim 13 wherein the controller moves a data partition.
15. The distributed computing system of claim 13 wherein the controller moves a data partition to a new computing node.
16. The distributed computing system of claim 2 wherein the controller reconfigures a volume of stored data.
17. The distributed computing system of claim 2 wherein the controller deploys additional computing nodes.