IP Library › Granted Patent US 12,008,120
Granted Patent B2
US 12,008,120 · App. 17/339,826 · Granted Jun 11, 2024

Data distribution and security in a multilayer storage infrastructure

Inventors: Craig M. Trim (Ventura, CA); Shikhar Kwatra (San Jose, CA); Bennet Prabhu (Tiruvallur, IN); Jayabalan Arumugam (Chennai, IN)
Assignee: International Business Machines Corporation
G06F21/6209G06F11/1004G06F18/23G06F21/602G06F40/205G06N20/20
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,008,120
App. No.
17/339,826
Granted
Jun 11, 2024
Kind
B2
Abstract

Techniques are described relating to data distribution and security in a multilayer storage infrastructure. An associated computer-implemented method includes receiving file data associated with a user for storage in a managed services domain, applying an ensemble learning model to devise a data distribution technique for the file data based upon contextual information associated with the user, and encrypting the file data. The method further includes, based upon the data distribution technique, dividing the file data to store among a cloud computing layer, a fog computing layer, and a local computing layer by performing a hash transformation and applying at least one cyclic error correcting code. In an embodiment, the method further includes receiving a data access request associated with the file data, authenticating the data access request, and restoring the file data via decryption.

Claims (62)

1. A computer-implemented method comprising:

receiving file data associated with a user for storage in a managed services domain;

applying an ensemble learning model to devise a data distribution technique for the file data based upon contextual information associated with the user, wherein applying the ensemble learning model comprises:

deriving topical context data by applying natural language processing (NLP) to the contextual information; and

deriving a plurality of file data portions by applying NLP to the file data based upon the topical context data;

encrypting the file data; and

based upon the data distribution technique, dividing the file data to store among a cloud computing layer, a fog computing layer, and a local computing layer by performing a hash transformation and applying at least one cyclic error correcting code.

2. The computer-implemented method of claim 1 , further comprising:

receiving a data access request associated with the file data;

authenticating the data access request; and

restoring the file data via decryption.

3. The computer-implemented method of claim 2 , wherein authenticating the data access request comprises:

parsing identification and a password from the data access request;

fetching a stored hash value pre-associated with the user based upon the parsed identification;

hashing the password by applying to the parsed password one or more cryptographic hash functions corresponding to the stored hash value; and

responsive to determining that a result of the password hash matches the stored hash value, completing authentication of the data access request.

4. The computer-implemented method of claim 1 , wherein applying the ensemble learning model further comprises:

applying at least one multi-class classification technique to the plurality of file data portions.

5. The computer-implemented method of claim 4 , wherein applying the ensemble learning model further comprises:

determining a distribution plan for the plurality of file data portions based upon application of the at least one multi-class classification technique.

6. The computer-implemented method of claim 1 , wherein deriving the topical context data comprises:

creating a plurality of encoded feature vectors by applying at least one NLP model to raw data associated with the user; and

obtaining numerical topical output by applying at least one clustering algorithm to the plurality of encoded feature vectors.

7. The computer-implemented method of claim 6 , wherein creating the plurality of encoded feature vectors comprises:

deriving at least one data representation based upon application types or data types accessed by the user.

8. The computer-implemented method of claim 6 , wherein creating the plurality of encoded feature vectors comprises:

deriving at least one data representation pertaining to frequency of application access or data access by the user; and

deriving at least one data representation pertaining to computing resource usage or storage patterns related to application access or data access by the user.

9. The computer-implemented method of claim 6 , wherein a respective feature of one of the plurality of encoded feature vectors is associated with a respective topic based upon the numerical topical output obtained for the respective feature.

10. The computer-implemented method of claim 1 , wherein deriving the plurality of file data portions comprises:

identifying topical patterns among datapoints within the file data by applying at least one NLP model in view of the topical context data; and

allocating the datapoints within the file data to the plurality of file data portions by applying at least one clustering algorithm based upon the identified topical patterns.

11. The computer-implemented method of claim 1 , wherein dividing the file data comprises:

storing separate parts of the file data among the cloud computing layer, the fog computing layer, and the local computing layer.

12. The computer-implemented method of claim 1 , wherein dividing the file data comprises:

storing redundant dependencies among each of the cloud computing layer, the fog computing layer, and the local computing layer.

13. The computer-implemented method of claim 1 , wherein the contextual information includes user frequency of data user or user frequency of application use.

14. The computer-implemented method of claim 1 , wherein the contextual information includes user system configuration or user file storage pattern.

15. A computer program product comprising a computer readable storage medium having program instructions embodied therewith, the program instructions executable by a computing device to cause the computing device to:

receive file data associated with a user for storage in a managed services domain;

apply an ensemble learning model to devise a data distribution technique for the file data based upon contextual information associated with the user, wherein applying the ensemble learning model comprises:

deriving topical context data by applying natural language processing (NLP) to the contextual information; and

deriving a plurality of file data portions by applying NLP to the file data based upon the topical context data;

encrypt the file data; and

based upon the data distribution technique, divide the file data to store among a cloud computing layer, a fog computing layer, and a local computing layer by performing a hash transformation and applying at least one cyclic error correcting code.

16. The computer program product of claim 15 , wherein the program instructions further cause the computing device to:

receive a data access request associated with the file data;

authenticate the data access request; and

restore the file data via decryption.

17. A system comprising:

at least one processor; and

a memory storing an application program, which, when executed on the at least one processor, performs an operation comprising:

receiving file data associated with a user for storage in a managed services domain;

applying an ensemble learning model to devise a data distribution technique for the file data based upon contextual information associated with the user, wherein applying the ensemble learning model comprises:

deriving topical context data by applying natural language processing (NLP) to the contextual information; and

deriving a plurality of file data portions by applying NLP to the file data based upon the topical context data;

encrypting the file data; and

based upon the data distribution technique, dividing the file data to store among a cloud computing layer, a fog computing layer, and a local computing layer by performing a hash transformation and applying at least one cyclic error correcting code.

18. The system of claim 17 , wherein the operation further comprises:

receiving a data access request associated with the file data;

authenticating the data access request; and

restoring the file data via decryption.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jun 4, 2021
From: TRIM, CRAIG M.; KWATRA, SHIKHAR; PRABHU, BENNET; ARUMUGAM, JAYABALAN
To: INTERNATIONAL BUSINESS MACHINES CORPORATION
Reel/Frame 056446/0634 →
Continuity (1)
Related Publication 20220391519A1 · Dec 8, 2022