Handoff Design in User-Centric Cell-Free Massive MIMO Networks Using DRL

Fuente: arXiv
Gespeichert in:
Bibliographische Detailangaben
Hauptverfasser: Ammar, Hussein A., Adve, Raviraj, Shahbazpanahi, Shahram, Boudreau, Gary, Bahceci, Israfil
Format: Preprint
Veröffentlicht: 2025
Schlagworte:
Online-Zugang:
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
_version_ 1866915422841864192
author Ammar, Hussein A.
Adve, Raviraj
Shahbazpanahi, Shahram
Boudreau, Gary
Bahceci, Israfil
author_facet Ammar, Hussein A.
Adve, Raviraj
Shahbazpanahi, Shahram
Boudreau, Gary
Bahceci, Israfil
contents In the user-centric cell-free massive MIMO (UC-mMIMO) network scheme, user mobility necessitates updating the set of serving access points to maintain the user-centric clustering. Such updates are typically performed through handoff (HO) operations; however, frequent HOs lead to overheads associated with the allocation and release of resources. This paper presents a deep reinforcement learning (DRL)-based solution to predict and manage these connections for mobile users. Our solution employs the Soft Actor-Critic algorithm, with continuous action space representation, to train a deep neural network to serve as the HO policy. We present a novel proposition for a reward function that integrates a HO penalty in order to balance the attainable rate and the associated overhead related to HOs. We develop two variants of our system; the first one uses mobility direction-assisted (DA) observations that are based on the user movement pattern, while the second one uses history-assisted (HA) observations that are based on the history of the large-scale fading (LSF). Simulation results show that our DRL-based continuous action space approach is more scalable than discrete space counterpart, and that our derived HO policy automatically learns to gather HOs in specific time slots to minimize the overhead of initiating HOs. Our solution can also operate in real time with a response time less than 0.4 ms.
format Preprint
id arxiv_https___arxiv_org_abs_2507_20966
institution arXiv
publishDate 2025
record_format arxiv
spellingShingle Handoff Design in User-Centric Cell-Free Massive MIMO Networks Using DRL
Ammar, Hussein A.
Adve, Raviraj
Shahbazpanahi, Shahram
Boudreau, Gary
Bahceci, Israfil
Information Theory
Artificial Intelligence
Machine Learning
Networking and Internet Architecture
Signal Processing
In the user-centric cell-free massive MIMO (UC-mMIMO) network scheme, user mobility necessitates updating the set of serving access points to maintain the user-centric clustering. Such updates are typically performed through handoff (HO) operations; however, frequent HOs lead to overheads associated with the allocation and release of resources. This paper presents a deep reinforcement learning (DRL)-based solution to predict and manage these connections for mobile users. Our solution employs the Soft Actor-Critic algorithm, with continuous action space representation, to train a deep neural network to serve as the HO policy. We present a novel proposition for a reward function that integrates a HO penalty in order to balance the attainable rate and the associated overhead related to HOs. We develop two variants of our system; the first one uses mobility direction-assisted (DA) observations that are based on the user movement pattern, while the second one uses history-assisted (HA) observations that are based on the history of the large-scale fading (LSF). Simulation results show that our DRL-based continuous action space approach is more scalable than discrete space counterpart, and that our derived HO policy automatically learns to gather HOs in specific time slots to minimize the overhead of initiating HOs. Our solution can also operate in real time with a response time less than 0.4 ms.
title Handoff Design in User-Centric Cell-Free Massive MIMO Networks Using DRL
topic Information Theory
Artificial Intelligence
Machine Learning
Networking and Internet Architecture
Signal Processing
url https://arxiv.org/abs/2507.20966