IP Library › Granted Patent US 12,561,611
Granted Patent B2
US 12,561,611 · App. 17/995,300 · Granted Feb 24, 2026

Method for efficient distributed machine learning hyperparameter search

Inventors: Hasan Farooq (Santa Clara, CA); Julien Forgeat (San Jose, CA); Meral Shirazipour (Santa Clara, CA)
Assignee: Telefonaktiebolaget LM Ericsson (publ)
G06N20/00
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,561,611
App. No.
17/995,300
Granted
Feb 24, 2026
Kind
B2
Abstract

A method of a hyperparameter server improves hyper-parameter search efficiency for devices in a self-organizing network (SON) includes sending configuration for data feature collection to at least one edge device in the self-organizing network, receiving hyper-parameter performance data from the at least one edge device, and training a shared hyperparameter machine learning model using a global training database including the hyperparameter performance data to identify optimal hyperparameters for use by the at least one edge device. A further method of an edge device improves hyperparameter search efficiency for devices in a SON includes receiving configuration for data feature collection from a hyperparameter server, training an edge machine learning model using local training data and selected hyperparameters, and sending performance data to the hyperparameter server obtained from the training of the edge machine learning model.

Claims (40)

1 . A method of a hyperparameter server to improve hyperparameter search efficiency for devices in a self-organizing network associated with a communication network, where hyperparameters have values that are set prior to training a machine learning model and that affect the training of the machine learning model, and where the self-organizing network is a network that can configure, optimize and heal itself through automated processes, the method comprising:

sending configuration for data feature collection to a plurality of edge devices in the self-organizing network, wherein the plurality of edge devices are grouped into a cluster of edge devices comprised of base stations having identified commonalities among the base stations;

receiving hyperparameter performance data from the edge devices that are based training a machine learning model at the edge devices based on the configuration; and

training a shared hyperparameter machine learning model using a global training database including the hyperparameter performance data to identify optimal hyperparameters for use by at least one edge device of the cluster.

2 . The method of claim 1 , further comprising:

adding the hyperparameter performance data to the global training database.

3 . The method of claim 1 , further comprising:

servicing a hyperparameter query from the at least one edge device using the shared hyperparameter machine learning model after updating by the training.

4 . The method of claim 1 , wherein the hyperparameter performance data includes a tuple of probability distribution, hyperparameter settings, and performance metrics.

5 . The method of claim 1 , further comprising:

establishing the cluster with a shared hyperparameter machine learning model and training database for the cluster.

6 . A method of an edge device of a communication network to improve efficiency for a hyperparameter search for devices in a self-organizing network associated with the communication network, where the hyperparameter search is for hyperparameters for a machine learning model for the self-organizing network, where hyperparameters have values that are set prior to training the machine learning model and that affect the training of the machine learning model, and where the self-organizing network is a network that can configure, optimize and heal itself through automated processes using the machine learning model the method comprising:

receiving configuration for data feature collection from a hyperparameter server, wherein the configuration for data feature collection is sent to a plurality of edge devices in the self-organizing network, wherein the plurality of edge devices are grouped into a cluster of edge devices comprised of base stations having identified commonalities among the base stations and wherein the edge device comprises a base station within the cluster;

training an edge machine learning model using local training data and selected hyperparameters from the configuration; and

sending hyperparameter performance data to the hyperparameter server obtained from the training of the edge machine learning model for the hyperparameter server to train a shared hyperparameter machine learning model using a global training database including the hyperparameter performance data to identify optimal hyperparameters for use by the edge device.

7 . The method of claim 6 , further comprising:

collecting data features according to the configuration.

8 . The method of claim 6 , further comprising:

sending a hyperparameter query to the hyperparameter server to obtain optimized hyperparameters from a shared hyperparameter machine learning model.

9 . The method of claim 6 , wherein the hyperparameter performance data includes a tuple of probability distribution, hyperparameter settings, and performance metrics.

10 . A non-transitory machine-readable storage medium that provides instructions that, if executed by a processor, will cause said processor to perform a set of operations for hyperparameter determination for a machine learning model for a self-organizing network associated with a communication network, where hyperparameters have values that are set prior to training the machine learning model and that affect the training of the machine learning model, and where the self-organizing network is a network that can configure, optimize and heal itself through automated processes using the machine learning model, the set of operations comprising:

sending configuration for data feature collection to a plurality of edge devices in the self-organizing network, wherein the plurality of edge devices are grouped into a cluster of edge devices comprised of base stations having identified commonalities among the base stations;

receiving hyperparameter performance data from the edge devices that are based training a machine learning model at the edge devices based on the configuration; and

training a shared hyperparameter machine learning model using a global training database including the hyperparameter performance data to identify optimal hyperparameters for use by at least one edge device of the cluster.

11 . The non-transitory machine-readable storage medium of claim 10 , the set of operations further comprising:

adding the hyperparameter performance data to the global training database.

12 . The non-transitory machine-readable storage medium of claim 10 , the set of operations further comprising:

servicing a hyperparameter query from the at least one edge device using the shared hyperparameter machine learning model after updating by the training.

13 . The non-transitory machine-readable storage medium of claim 10 , wherein the hyperparameter performance data includes a tuple of probability distribution, hyperparameter settings, and performance metrics.

14 . The non-transitory machine-readable storage medium of claim 10 , the set of operations further comprising:

establishing the cluster with a shared hyperparameter machine learning model and training database for the cluster.

15 . A non-transitory machine-readable storage medium that provides instructions that, if executed by a processor associated with an edge device, will cause said processor to perform a set of operations for hyperparameter determination for a machine learning model for a self-organizing network associated with a communication network, where hyperparameters have values that are set prior to training the machine learning model and that affect the training of the machine learning model, and where the self-organizing network is a network that can configure, optimize and heal itself through automated processes using the machine learning model, the set of operations comprising:

receiving configuration for data feature collection from a hyperparameter server, wherein the configuration for data feature collection is sent to a plurality of edge devices in the self-organizing network, wherein the plurality of edge devices are grouped into a cluster of edge devices comprised of base stations having identified commonalities among the base stations and wherein the edge device comprises a base station within the cluster;

training an edge machine learning model using local training data and selected hyperparameters from the configuration; and

sending hyperparameter performance data to the hyperparameter server obtained from the training of the edge machine learning model for the hyperparameter server to train a shared hyperparameter machine learning model using a global training database including the hyperparameter performance data to identify optimal hyperparameters for use by the edge device.

16 . The non-transitory machine-readable storage medium of claim 15 , the set of operations further comprising:

collecting data features according to the configuration.

17 . The non-transitory machine-readable storage medium of claim 15 , the set of operations further comprising:

sending a hyperparameter query to the hyperparameter server to obtain optimized hyperparameters from a shared hyperparameter machine learning model.

18 . The non-transitory machine-readable storage medium of claim 15 , wherein the hyperparameter performance data includes a tuple of probability distribution, hyperparameter settings, and performance metrics.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Sep 30, 2022
From: FAROOQ, HASAN; FORGEAT, JULIEN; SHIRAZIPOUR, MERAL
To: TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)
Reel/Frame 061278/0403 →
Continuity (1)
Related Publication 20230162089A1 · May 25, 2023
References Cited (30)
US 20070100993A1 · Malhotra · 2007 [cited by examiner]
US 20150046515A1 · Pei · 2015 [cited by examiner]
US 20160162779A1 · Marcus · 2016 [cited by examiner]
US 20170109322A1 · McMahan et al. · 2017 [cited by applicant]
US 20170185915A1 · Chawla · 2017 [cited by examiner]
US 20180032908A1 · Nagaraju et al. · 2018 [cited by applicant]
US 20180060738A1 · Achin · 2018 [cited by examiner]
US 20180375720A1 · Yang · 2018 [cited by examiner]
US 20190273509A1 · Elkind · 2019 [cited by examiner]
US 20190273510A1 · Elkind · 2019 [cited by examiner]
US 20190297186A1 · Karani · 2019 [cited by examiner]
US 20200019888A1 · McCourt · 2020 [cited by examiner]
US 20200184376A1 · Parameswaran · 2020 [cited by examiner]
US 20210089937A1 · Zhang · 2021 [cited by examiner]
US 20220156642A1 · Schmidt · 2022 [cited by examiner]
US 20240394443A1 · Li · 2024 [cited by examiner]
A. P. Athreya et al., “Network Self-Organization in the Internet of Things,” 2013, pp. 25-33, IEEE International Conference on Sensing, Communications and Networking (SECON). [cited by applicant]
B. Hughes et al., “Generative Adversarial Learning for Machine Learning empowered Self Organizing 5G Networks,” 2019, pp. 282-286, 2019 International Conference on Computing, Networking and Communications (CNC). [cited by applicant]
C. Svahn et al., “Inter-Frequency Radio Signal Quality Prediction for Handover, Evaluated in 3GPP LTE,” 2019, pp. 1-5, 2019 IEEE 89th Vehicular Technology Conference (VTC2019-Spring). [cited by applicant]
F. B. Mismar et al., “Partially Blind Handovers for mmWave New Radio Aided by Sub-6 GHz LTE Signaling,” 2018, pp. 1-5, 2018 IEEE International Conference on Communications Workshops (ICC Workshops). [cited by applicant]
F. Chen et al., “Federated Meta-Learning with Fast Convergence and Efficient Communication,” Dec. 14, 2019, pp. 1-14, Computer Science and Machine Learning (ARXIV). [cited by applicant]
H. Ryden et al., “Predicting Strongest Cell on Secondary Carrier Using Primary Carrier Data,” 2018, pp. 137-142, 2018 IEEE Wireless Communications and Networking Conference Workshops (WCNCW): 7th International Workshop … [cited by applicant]
International Preliminary Report on Patentability, PCT App. No. PCT/IB2020/053231, May 23, 2022, 19 pages. [cited by applicant]
International Search Report and Written Opinion, PCT App. No. PCT/IB2020/053231, Dec. 18, 2020, 13 pages. [cited by applicant]
J. Snoek et al., “Practical Bayesian Optimization of Machine Learning Algorithms,” 2012, 9 pages, NIPS'2012. [cited by applicant]
M. Feurer et al., “Chapter 1, Hyperparameter Optimization,” 2019, 31 pages, Automated Machine Learning, The Springer Series on Challenges in Machine Learning. [cited by applicant]
O. G. Aliu et al., “A Survey of Self Organisation in Future Cellular Networks,” 2013, pp. 336-361, IEEE Communications Surveys & Tutorials, vol. 15, No. 1, First Quarter 2013. [cited by applicant]
P. V. Klaine et al., “A Survey of Machine Learning Techniques Applied to Self-Organizing Cellular Networks,” 2017, pp. 2392-2431, IEEE Communications Surveys & Tutorials, vol. 19, No. 4, Fourth Quarter 2017. [cited by applicant]
Written Opinion of the International Preliminary Examining Authority, PCT App. No. PCT/IB2020/053231, Feb. 23, 2022, 11 pages. [cited by applicant]
Y. Yang et al., “Generative-Adversarial-Network-Based Wireless Channel Modeling: Challenges and Opportunities,” Mar. 2019, pp. 22-27, IEEE Communications Magazine, vol. 57, No. 3. [cited by applicant]