IP Library › Granted Patent US 12,744,569
Granted Patent B2
US 12,744,569 · App. 18/578,896 · Granted Sep 22, 2026

Reinforcement learning of beam codebooks for millimeter wave and terahertz MIMO systems

Inventors: Ahmed Alkhateeb (Chandler, AZ); Yu Zhang (Tempe, AZ); Muhammad Alrabeiah (Tempe, AZ)
Assignee: ARIZONA BOARD OF REGENTS ON BEHALF OF ARIZONA STATE UNIVERSITY
H04B7/0617H04B7/0684
View Patent ↗
Loading inventors, assignments & file history…
Monitor This Case
Get email alerts when status or documents change.
Order Certified Copies
Most orders are placed with the USPTO same day — all within 24 business hours.
Order via The Patent Place →
Pre-filled with this patent's details
Quick Facts
Patent No.
US 12,744,569
App. No.
18/578,896
Granted
Sep 22, 2026
Kind
B2
Abstract

Reinforcement learning of beam codebooks for millimeter wave and terahertz multiple-input-multiple-output (MIMO) systems is provided. Millimeter wave (mmWave) and terahertz (THz) MIMO systems rely on predefined beamforming codebooks for both initial access and data transmission. These predefined codebooks, however, are commonly not optimized for specific environments, user distributions, and/or possible hardware impairments. To overcome these limitations, this disclosure develops a deep reinforcement learning framework that learns how to optimize the codebook beam patterns relying only on receive power measurements. The developed model learns how to adapt the beam patterns based on the surrounding environment, user distribution, hardware impairments, and array geometry. Further, this approach does not require any knowledge about the channel, radio frequency (RF) hardware, or user positions.

Claims (33)

1 . A method for intelligently learning a beam codebook for multi-antenna wireless communications, the method comprising:

obtaining receive power measurements from a plurality of antennas; and

training the beam codebook using deep learning and the receive power measurements,

wherein the training of the beam codebook is achieved without knowledge of positions of wireless users within an environment accessed by a network node associated with the beam codebook.

2 . The method of claim 1 , further comprising beamforming wireless communications with a wireless device using the trained beam codebook.

3 . The method of claim 2 , further comprising initiating the wireless communications with the wireless device using the trained beam codebook.

4 . The method of claim 1 , wherein the training of the beam codebook uses the deep learning and the receive power measurements only.

5 . The method of claim 1 , wherein the training of the beam codebook is achieved without employing channel estimation.

6 . The method of claim 1 , wherein no knowledge of hardware details of communication circuitry that is employed at the network node associated with the beam codebook is used to train the beam codebook.

7 . The method of claim 1 , wherein the training of the beam codebook is achieved without knowledge of hardware details of communication circuitry that is employed at the network node associated with the beam codebook.

8 . A neural network for training of a beam codebook for multi-antenna wireless communications, the neural network comprising:

an actor network configured to predict one or more beam patterns for the beam codebook; and

a critic network configured to evaluate the one or more beam patterns predicted by the actor network based on receive power measurements of an environment,

wherein the training of the beam codebook is achieved without knowledge of positions of wireless users within an environment accessed by a network node associated with the beam codebook.

9 . The neural network of claim 8 , wherein the neural network comprises a Wolpertinger architecture.

10 . The neural network of claim 8 , wherein the training of the beam codebook uses deep learning and the receive power measurements only.

11 . The neural network of claim 8 , wherein the training of the beam codebook is achieved without employing channel estimation.

12 . The neural network of claim 8 , wherein the training of the beam codebook is achieved without knowledge of hardware details of communication circuitry that is employed at the network node associated with the beam codebook.

13 . The neural network of claim 8 , further comprising beamforming wireless communications with a wireless device using the trained beam codebook.

14 . A network node, comprising:

communication circuitry coupled to a plurality of antennas and configured to establish communications with a wireless device in an environment; and

a processing system configured to:

obtain receive power measurements from the plurality of antennas;

perform a machine learning-based analysis of the environment based on the receive power measurements; and

adapt the communications with the wireless device in accordance with the machine learning-based analysis of the environment,

wherein training of a beam codebook associated with the network node is achieved without knowledge of positions of wireless users of the wireless device within the environment accessed by the network node.

15 . The network node of claim 14 , wherein the communication circuitry comprises a radio frequency (RF) transceiver.

16 . The network node of claim 15 , wherein the RF transceiver is configured to communicate via at least one of a terahertz (THz) band or a millimeter wave (mm Wave) band.

17 . The network node of claim 14 , wherein the training of the beam codebook uses deep learning and the receive power measurements only.

18 . The network node of claim 14 , wherein the training of the beam codebook is achieved without employing channel estimation.

19 . The network node of claim 14 , wherein the training of the beam codebook is achieved without knowledge of hardware details of the communication circuitry that is employed at the network node associated with the beam codebook.

20 . The network node of claim 14 , wherein the processing system is further configured to:

beamform wireless communications with the wireless device using the trained beam codebook.

Assignments (1)
ASSIGNMENT OF ASSIGNOR'S INTEREST Recorded Jan 17, 2024
From: ALKHATEEB, AHMED; ZHANG, YU; ALRABEIAH, MUHAMMED
To: ARIZONA BOARD OF REGENTS ON BEHALF OF ARIZONA STATE UNIVERSITY
Reel/Frame 066148/0395 →
Continuity (2)
Provisional Application 63221192 · Jul 13, 2021
Related Publication 20250007579A1 · Jan 2, 2025
References Cited (8)
US 20090238156A1 · Yong et al. · 2009 [cited by applicant]
US 20190140730A1 · Oteri et al. · 2019 [cited by applicant]
US 20200358514A1 · Landis et al. · 2020 [cited by applicant]
US 20210194551A1 · Raghavan et al. · 2021 [cited by applicant]
US 20210250068A1 · Lee · 2021 [cited by examiner]
US 20230006718A1 · Kostas · 2023 [cited by examiner]
International Search Report and Written Opinion mailed Oct. 4, 2022 in corresponding International Application No. PCT/US2022/036795, 13 pages. [cited by applicant]
Zhang et al., “Reinforcement Learning for Beam Pattern Design in Millimeter Wave and Massive MIMO Systems,” 2020 (2020), [Retrieved from online on Sep. 6, 2022 (Sep. 6, 2022)]; [Retrieved from URL: https://ieeexplore.ie… [cited by applicant]