IP Library › Granted Patent US 12,349,061
Granted Patent B2
US 12,349,061 · App. 18/205,053 · Granted Jul 1, 2025

Power saving in radio access network

Inventors: Vaibhav Singh (Bangalore, IN); Anand Bedekar (Glenview, IL); Jun He (Shanghai, CN)
Assignee: NOKIA SOLUTIONS AND NETWORKS OY
H04W52/0206H04W24/10H04W52/02
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,349,061
App. No.
18/205,053
Granted
Jul 1, 2025
Kind
B2
Abstract

To maximize power saving in a radio access network comprising cells, an optimal action amongst actions comprising switching on one or more cells, switching off one or more cells, and doing nothing is determined using a trained model, which maximizes a long term reward on tradeoff between throughput and power, the trained model taking as input a load estimate. The trained model may be updated online using measurement results on load, throughput and power consumption.

Claims (49)

1. An apparatus comprising:

at least one processor; and

at least one memory including computer program code, the at least one memory and computer program code configured to, with the at least one processor, cause the apparatus at least to perform:

determining, for a group of cells in a radio access network, an optimal action as an output of a first trained model, using the first trained model, which is based on reinforcement learning and maximizes a long term reward on tradeoff between throughput and power saving within the group of cells, the first trained model taking as input a state, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses in cells of the group of cells, and wherein the state comprises at least one of a load estimate and, per a cell in the group of cells, a current cell status;

causing the optimal action to be performed in response to the optimal action being modifying power settings of one or more cells, or switching on one or more cells, or switching off one or more cells;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on.

2. The apparatus of claim 1 , wherein the at least one memory and computer program code are configured to, with the at least one processor, cause the apparatus further at least to perform the determining in response to receiving, as a new load estimate, a new load prediction from a second trained model comprised in the apparatus or in another apparatus, the second trained model outputting periodically, using at least measured load data from the radio access network as input, load predictions.

3. The apparatus of claim 1 , wherein the at least one memory and computer program code are configured to, with the at least one processor, cause the apparatus further at least to perform:

instantiating and running the first trained model as a service on top of a radio intelligent controller near real time platform; and

using a data write application programming interface of the radio intelligent controller near real time platform, when causing the optimal action to be performed.

4. An apparatus, comprising:

at least one processor; and

at least one memory including computer program code, the at least one memory and computer program code configured to, with the at least one processor, cause the apparatus at least to perform:

initializing a first trainable model, which maximizes a long term reward on tradeoff between throughput and power saving in a radio access network comprising cells and which first trainable model outputs an optimal action, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses;

acquiring historical data comprising a plurality of time series of evolution of at least load data, power consumption data, and cell throughput data in the radio access network, time series comprising a plurality of time steps;

training the first trainable model to a first trained model using reinforcement learning and iterating the plurality of time series and by iterating, per a time series, the plurality of time steps;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on.

5. The apparatus of claim 4 , wherein the at least one memory and computer program code configured to, with the at least one processor, cause the apparatus further at least to perform:

using Q learning as the reinforcement learning.

6. A method comprising:

determining, for a group of cells in a radio access network, an optimal action as an output of a first trained model, using the first trained model, which is based on reinforcement learning and maximizes a long term reward on tradeoff between throughput and power saving within the group of cells, the first trained model taking as input a state, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses in cells of the group of cells, and wherein the state comprises at least one of a load estimate and, per a cell in the group of cells, a current cell status;

causing the optimal action to be performed in response to the optimal action being modifying power settings of one or more cells, switching on one or more cells, or switching off one or more cells;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on.

7. The method of claim 6 , the method further comprising performing the determining in response to receiving, as a new load estimate, a new load prediction from a second trained model comprised in the apparatus or in another apparatus, the second trained model outputting periodically, using at least measured load data from the radio access network as input, load predictions.

8. The method of claim 6 , the method further comprising:

instantiating and running the first trained model as a service on top of a radio intelligent controller near real time platform; and

using a data write application programming interface of the radio intelligent controller near real time platform, when causing the optimal action to be performed.

9. The method of claim 6 , the method further comprising using Q learning as the reinforcement learning.

10. A method comprising:

initializing a first trainable model, which maximizes a long term reward on tradeoff between throughput and power saving in a radio access network comprising cells and which first trainable model outputs an optimal action, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses;

acquiring historical data comprising a plurality of time series of evolution of at least load data, power consumption data, and cell throughput data in the radio access network, time series comprising a plurality of time steps;

training the first trainable model to a first trained model using reinforcement learning and iterating the plurality of time series and by iterating, per a time series, the plurality of time steps;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on.

11. A non-transitory computer readable medium comprising program instructions for causing an apparatus to perform at least one of a first process and a second process,

wherein the first process comprises at least:

determining, for a group of cells in a radio access network, an optimal action as an output of a first trained model, using the first trained model, which is based on reinforcement learning and maximizes a long term reward on tradeoff between throughput and power saving within the group of cells, the first trained model taking as input a state, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses in cells of the group of cells, and wherein the state comprises at least one of a load estimate and, per a cell in the group of cells, a current cell status;

causing the optimal action to be performed in response to the optimal action being modifying power settings of one or more cells, or switching on one or more cells, or switching off one or more cells;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on,

wherein the second process comprises at least:

initializing a first trainable model, which maximizes a long term reward on tradeoff between throughput and power saving in a radio access network comprising cells and which first trainable model outputs an optimal action, wherein the optimal action is one of actions comprising at least modifying power settings of one or more cells, switching on one or more cells, switching off one or more cells, and retaining the current cell statuses;

acquiring historical data comprising a plurality of time series of evolution of at least load data, power consumption data, and cell throughput data in the radio access network, time series comprising a plurality of time steps;

training the first trainable model to a first trained model using reinforcement learning and iterating the plurality of time series and by iterating, per a time series, the plurality of time steps;

receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on; and

updating the first trained model in response to the receiving load and performance metrics of cells that are switched on, and power consumed by the cells that are switched on.

Priority Claims (1)
FI 20216111 · Oct 28, 2021 · national
Continuity (2)
Continuation 17976376 · Oct 28, 2022
Related Publication 20230319707A1 · Oct 5, 2023
References Cited (35)
US 10484904B2 · Chow · 2019 [cited by examiner]
US 11765654B2 · Singh · 2023 [cited by examiner]
US 20150078261A1 · Yu et al. · 2015 [cited by applicant]
US 20170026888A1 · Kwan · 2017 [cited by examiner]
US 20200313434A1 · Khanna · 2020 [cited by examiner]
US 20210119881A1 · Shirazipour et al. · 2021 [cited by applicant]
US 20210258866A1 · Chou · 2021 [cited by applicant]
US 20210297876A1 · Bellamkonda · 2021 [cited by applicant]
US 20220141766A1 · Choi et al. · 2022 [cited by applicant]
US 20220343220A1 · Sawabe et al. · 2022 [cited by applicant]
US 20230116202A1 · Mendo Mateo · 2023 [cited by examiner]
US 20230171592A1 · Han · 2023 [cited by examiner]
US 20240196178A1 · Ying · 2024 [cited by examiner]
US 20240298225A1 · Hyde · 2024 [cited by examiner]
CN 112752327A · 2021 [cited by applicant]
EP 3869847A1 · 2021 [cited by applicant]
WO 2020131128A1 · 2020 [cited by applicant]
WO 2022207402A1 · 2022 [cited by applicant]
Office Action and Search Report dated Sep. 24, 2023, corresponding to Chinese Patent Application No. 202211334661.9. [cited by applicant]
Intel Corporation, “Use cases for AI/ML enabled NG-RAN”, 3GPP TSG-RAN WG3 Meeting #112-e Electronic meeting, May 17-28, 2021. R3-212301. [cited by applicant]
Office Action dated Mar. 7, 2022 corresponding to Finnish Patent Application No. 20216111. [cited by applicant]
Search Report dated Mar. 7, 2022 corresponding to Finnish Patent Application No. 20216111. [cited by applicant]
Communication of Acceptance—section 29 a of Patents Decree dated Oct. 12, 2022 corresponding to Finnish Patent Application No. 20216111. [cited by applicant]
Y. Gao et al., “Machine Learning based Energy Saving Scheme in Wireless Access Networks,” In: 2020 International Wireless Communications and Mobile Computing (WCMC), Jun. 15, 2020, pp. 1573-1578. [cited by applicant]
I. Donevski et al., “Neural networks for cellular base station switching,” In: 2019 INFOCOM IEEE Conference on Computer Communications Workshops, INFOCOM WKSHPS 2019, Apr. 29, 2019, pp. 738-743. [cited by applicant]
H. Farooq et al., “Mobility prediction-based autonomous proactive energy saving (AURORA) framework for emerging ultra-dense networks,” In: IEEE Trans. on Green Communications and Networking, Dec. 2018, vol. 2, No. 4, pp… [cited by applicant]
S. Wu et al., “Deep convolutional neural network assisted reinforcement learning based mobile network power saving,” In: IEEE Access, May 15, 2020, vol. 8, pp. 93671-93681. [cited by applicant]
The partial European Search Report dated Mar. 20, 2023 corresponding to European Patent Application No. 22203716.0. [cited by applicant]
Intel Corporation: “Use cases for AI/ML enabled NG-RAN,” 3GPP Draft; R3-212301, 3rd Generation Partnership Project (3GPP), Mobile Competence Centre; 650 Route Des Lucioles; F-06921 Sophia-Antipolis Cedex; France, vol. R… [cited by applicant]
Kaloxylos Alexandros et al., “AI and ML—Enablers for Beyond 5G Networks,” Dec. 1, 2020. [cited by applicant]
Notice of Reasons for Rejection issued in corresponding Japanese Patent Application No. 2022-172404 dated Apr. 17, 2023, with English language summary thereof. [cited by applicant]
Ericsson, “Network Energy Saving”, 3GPP TSG-RAN WG3 Meeting #144-e, R3-215216, Oct. 21, 2021. [cited by applicant]
Mismar et al. “Deep Reinforcement Learning for 5G Networks”, Mar. 2020, IEEE transactions on communications, vol. 68, No. 3, 12 pages. [cited by applicant]
Rejection Decision dated Jul. 6, 2024, in corresponding Chinese Patent Application No. 202211334661.9, with English translation thereof. [cited by applicant]
Second Office Action issued in Chinese Patent Application No. 202211334661.9 dated Mar. 16, 2024, with English language summary thereof. [cited by applicant]