IP Library Granted Patent US 12,468,982
Granted Patent B2
US 12,468,982 · App. 17/671,092 · Granted Nov 11, 2025

Adaptive and evolutionary federated learning system

Inventors: Bhushan Gurmukhdas Jagyasi (Maharashtra, IN); Siva Rama Sarma Theerthala (Secunderaba, IN); Saurabh Pashine (Madhya Pradesh, IN); Soumit Bhowmick (Kolkata, IN); Gopali Raval Contractor (Maharashtra, IN)
Assignee: Accenture Global Solutions Limited
G06N20/00G06N3/126
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,468,982
App. No.
17/671,092
Granted
Nov 11, 2025
Kind
B2
Abstract

This application discloses a system and method for federated collaborative machine learning model development using local training datasets that are not shared. An adaptive and evolutionary approach is used to select local training nodes that are most fit from one training round to the next training round to optimize an overall cost and performance function for the federated learning, to cross-over model architecture between local training nodes, and to perform model architecture mutation within local training nodes. The local training nodes are further clustered to account for the inhomogeneity in the local datasets. Such adaptive, evolutionary, and collaborative federated learning thus provides cost-effective and high-performance model development.

Claims (48)

1 . A system comprising:

memory circuitry for storing computer instructions;

a network interface circuitry; and

processor circuitry in communication with the network interface circuitry and the memory circuitry, the processor circuitry configured to execute the computer instructions to:

receive sharable data from a plurality of local computation nodes;

cluster the plurality of local computation nodes into a plurality of clusters based on a set of clustering features extracted from the sharable data;

select a subset of local computation nodes from the plurality of local computation nodes as representatives of the plurality of clusters to participate in a collaborative machine learning; and

iteratively provision the collaborative machine learning by the subset of the local computation nodes until a termination condition is met by:

receipt, from the subset of local computation node, sets of model hyper parameters and sets of model metrics associated with machine learning models trained at the subset of local computation nodes using non-sharable datasets of the subset of local computation nodes;

performance of at least one model architectural hyper parameter cross-over of the machine learning models among the subset of local computation nodes to update the sets of model hyper parameters for the subset of local computing nodes, wherein the performance of the at least one model architectural hyper parameter cross-over of the machine learning models among the subset of local computation nodes is limited to intra-cluster cross-over;

elimination of selected local computation nodes of the subset of local computation nodes to obtain a remaining subset of local computation nodes using a multi-dimensional cost/performance function; and

instruction of the remaining subset of local computation nodes to perform a next round of training using the non-sharable datasets based on the updated sets of model hyper parameters.

2 . The system of claim 1 , wherein the non-sharable data from the plurality of local computation nodes comprises a historical time series.

3 . The system of claim 2 , wherein the set of clustering features comprise at least one of a mean, a trough, a variance, a trend, a peak, or a seasonality extracted from the historical time series by the plurality of local computation nodes.

4 . The system of claim 1 , wherein to select the subset of local computation nodes as the representatives of the plurality of clusters to participate in the collaborative machine learning comprises selection, by the processor circuitry, of local computation nodes at or near centroids of the plurality of clusters in a clustering space formed by the set of clustering features.

5 . The system of claim 1 , wherein to iteratively provision the collaborative machine learning further comprises performance, by the processor circuitry, of mutation of the set-sets of model hyper parameters with respect to at least one of the subset of local computation nodes.

6 . The system of claim 5 , wherein the mutation comprises modification, by the processor circuitry, of at least one of the sets of model hyper parameters.

7 . The system of claim 1 , wherein the multi-dimensional cost/performance function comprises at least one of a cost component, a performance component, or a local dataset quality component.

8 . The system of claim 7 , wherein the cost component comprises at least one of a communication cost between the subset of local computing nodes and the system or a computation cost at the subset of local computing nodes.

9 . The system of claim 8 , wherein the computation cost is determined by a complexity of the machine learning models trained at the subset of local computation nodes.

10 . The system of claim 9 , wherein the complexity of the machine learning models is computed based at least one of a number of model layers, a number of model elements in each model layer of the machine learning models trained at the subset of local computation nodes.

11 . The system of claim 7 , wherein the performance component comprises a mean absolute percentage error (MAPE) of the machine learning models.

12 . The system of claim 7 , wherein the local dataset quality component comprises at least one of a time take-taken or a learning rate for training the machine learning models at the subset of local computation nodes.

13 . The system of claim 1 , wherein the sets of model hyper parameters comprise parameters representing architectures of the machine learning models.

14 . The system of claim 1 , wherein to cluster the plurality of local computation nodes into the plurality of clusters, the processor circuitry is configured to:

identify a feature space with orthogonal features;

convert the sharable data into the orthogonal feature space; and

establish clustering dimensions using the orthogonal feature space to delineate the local computation nodes into the plurality of clusters.

15 . A method for adaptive federated machine learning performed by a computer server, comprising:

receiving sharable data from a plurality of local computation nodes;

clustering the plurality of local computation nodes into a plurality of clusters based on a set of clustering features extracted from the sharable data;

selecting a subset of local computation nodes from the plurality of local computation nodes as representatives of the plurality of clusters to participate in a collaborative machine learning; and

iteratively provisioning the collaborative machine learning until a termination condition is met by:

receiving, from the subset of local computation node, sets of model hyper parameters and sets of model metrics associated with machine learning models trained by the subset of local computation nodes using non-sharable datasets of the subset of local computation nodes;

performing at least one model architectural hyper parameter cross-over of the machine learning models among the subset of local computation nodes to update the sets of model hyper parameters for the subset of local computing nodes, wherein performing the at least one model architectural hyper parameter cross-over of the machine learning models among the subset of local computation nodes is limited to intra-cluster cross-over;

performing an elimination of selected local computation nodes of the subset of local computation nodes to obtain a remaining subset of local computation nodes using a multi-dimensional cost/performance function; and

instructing the remaining subset of local computation nodes to perform a next round of training using the non-sharable datasets based on the updated sets of model hyper parameters.

16 . The method of claim 15 , wherein selecting the subset of local computation nodes as the representatives of the plurality of clusters to participate in the collaborative machine learning comprises selecting local computation nodes at or near centroids of the plurality of clusters in a clustering space formed by the set of clustering features.

17 . The method of claim 15 , wherein iteratively provisioning the collaborative machine learning further comprises performing mutation of the sets of model hyper parameters with respect to at least one of the subset of local computation nodes.

18 . The method of claim 15 , wherein the multi-dimensional cost/performance function comprises at least one of a cost component, a performance component, or a local dataset quality component.

19 . The method of claim 18 , wherein:

the cost component comprises at least one of a communication cost between the subset of local computing nodes and a system or a computation cost at the subset of local computing nodes;

the performance component comprises a mean absolute percentage error (MAPE) of the machine learning models; and

the local dataset quality component comprises at least one of a time taken or a learning rate for training the machine learning models at the subset of local computation nodes.

20 . The method of claim 15 , wherein clustering the plurality of local computation nodes into the plurality of clusters comprises:

identifying a feature space with orthogonal features;

converting the sharable data into the orthogonal feature space; and

establishing clustering dimensions using the orthogonal feature space to delineate the local computation nodes into the plurality of clusters.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Feb 14, 2022
From: JAGYASI, BHUSHAN GURMUKHDAS; THEERTHALA, SIVA RAMA SARMA; PASHINE, SAURABH; BHOWMICK, SOUMIT; CONTRACTOR, GOPALI RAVAL
To: ACCENTURE GLOBAL SOLUTIONS LIMITED
Reel/Frame 059008/0127 →
Continuity (1)
Related Publication 20230259812A1 · Aug 17, 2023
References Cited (29)
US 20070208677A1 · Goldberg · 2007 [cited by examiner]
US 20210042628A1 · Zhou et al. · 2021 [cited by applicant]
US 20210067339A1 · Schiatti et al. · 2021 [cited by applicant]
US 20210374617A1 · Chu · 2021 [cited by examiner]
US 20220344049A1 · Hall · 2022 [cited by examiner]
US 20230068386A1 · Akdeniz · 2023 [cited by examiner]
US 20230177349A1 · Balakrishnan · 2023 [cited by examiner]
WO WO2021247448A1 · 2021 [cited by applicant]
Al-Saedi et al., “Reducing Communication Overhead of Federated Learning through Clustering Analysis”, 2021 IEEE Symposium on Computers and Communications (ISCC), Sept 5, 2021, pp. 1-7. (Year: 2021). [cited by examiner]
Zhu et al., “Multi-objective Evolutionary Federated Learning”, ARXIV ID: 1812.07478, published on Dec. 18, 2018, pp. 1-12. (Year: 2018). [cited by examiner]
Liu et al., “Deep Anomaly Detection for Time-series Data in Industrial IoT: A Communication-Efficient On-device Federated Learning Approach”, ARXIV ID: 2007.09712, published on Jul. 19, 2020, pp. 1-11. (Year: 2020). [cited by examiner]
Yeo et al., “Crossover-SGD: A gossip-based communication in distributed deep learning for alleviating large mini-batch problem and enhancing scalability”, ARXIV ID: 2012.15198, published on Dec. 20, 2020, pp. 1-14. (Yea… [cited by examiner]
Hadjiivanov et al., “Epigenetic evolution of deep convolutional models ”, ARXIV ID: 2104.05411, Apr. 12, 2021, pp. 1-9. (Year: 2021). [cited by examiner]
Li et al., “Evolutionary-based Federated Ensemble Learning on Face Recognition”, 2021 IEEE 4th Advanced Information Management, Communicates, Electronic and Automation Control Conference (IMCEC), vol. 4, Jun. 2021, pp. … [cited by examiner]
Zhu et al., “Real-Time Federated Evolutionary Neural Architecture Search”, IEEE Transactions on Evolutionary Computation, vol. 26, No. 2, Jul. 26, 2021, pp. 364-378. (Year: 2021). [cited by examiner]
Al-Saedi et al., “Reducing Communication Overhead of Federated Learning through Clustering Analysis”, 2021 IEEE Symposium on Computers and Communications (ISCC), Sep. 5, 2021, pp. 1-7. (Year: 2021). [cited by examiner]
He et al., “Short-Term Residential Load Forecasting Based on Federated Learning and Load Clustering”, 2021 IEEE International Conference on Communications, Control, and Computing Technologies for Smart Grids (SmartGridC… [cited by examiner]
Agrawal et al., “Genetic CFL: Hyperparameter Optimization in Clustered Federated Learning”, Computational Intelligence and Neuroscience, vol. 2021, Article ID 7156420, published on Nov. 18, 2021, pp. 1-10. (Year: 2021). [cited by examiner]
Nie et al., “Cross-Domain Recommendation via User-Clustering and Multidimensional Information Fusion”, IEEE Transactions on Multimedia, vol. 25, date of publication: Dec. 9, 2021, pp. 868-880. (Year: 2021). [cited by examiner]
Agrawal, Shaashwat et al., “Genetic CFL: Hyperparameter Optimization in Clustered Federated Learning”, Hindawi Computational Intelligence and Neuroscience, vol. 2021, Article ID 7156420; Nov. 18, 2021; 10 pages. [cited by applicant]
Beutel, Daniel J ., et al. “Flower: A friendly federated learning framework.” arXiv preprint arXiv:2007.14390 (2020). [cited by applicant]
Bonawitz, Keith, et al. “Towards federated learning at scale: System design.” arXiv preprint arXiv:1902.01046 (2019). [cited by applicant]
Kholod, Ivan, et al. “Open-Source Federated Learning Frameworks for IoT: A Comparative Review and Analysis.” Sensors 21.1 (2021): 167. [cited by applicant]
Li, Qinbin, et al. “A survey on federated learning systems: Vision, hype and reality for data privacy and protection.” arXiV preprint arXiv:1907.09693 (2019). [cited by applicant]
Lo, Sin Kit, et al. “Architectural Pattenls for the Design of Federated Learning Systems.” arXiV preprint arXiv:2101.02373 (2021). [cited by applicant]
Lo, Sin Kit, et al. “A systematic literature review on federated machine learning: From a software engineering perspective.” [cited by applicant]
Wang, Guan, Charlie Xiaoqian Dang, and Ziye Zhou. “Measure contribution of participants in federated learning.” 2019 IEEE International Conference on Big Data (Big Data). IEEE, 2019. [cited by applicant]
Zhu, Hangyu et al.; “Mum-objective Evolutionary Federated Learning”; arXiv:1812.07478V2 [cs.LG]; Jun. 8, 2019; 13 pages. [cited by applicant]
Extended European Search Report issued on EP23151496.9 on Jun. 26, 2023, 8 pages. [cited by applicant]
Cited By (1)
US 12,664,466