Deep Reinforcement Learning for Multi-User RF Charging with Non-linear Energy Harvesters

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Azarbahram, Amirhossein, López, Onel L. A., Popovski, Petar, Pandey, Shashi Raj, Latva-aho, Matti
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929336951504896
author Azarbahram, Amirhossein
López, Onel L. A.
Popovski, Petar
Pandey, Shashi Raj
Latva-aho, Matti
author_facet Azarbahram, Amirhossein
López, Onel L. A.
Popovski, Petar
Pandey, Shashi Raj
Latva-aho, Matti
contents Radio frequency (RF) wireless power transfer (WPT) is a promising technology for sustainable support of massive Internet of Things (IoT). However, RF-WPT systems are characterized by low efficiency due to channel attenuation, which can be mitigated by precoders that adjust the transmission directivity. This work considers a multi-antenna RF-WPT system with multiple non-linear energy harvesting (EH) nodes with energy demands changing over discrete time slots. This leads to the charging scheduling problem, which involves choosing the precoders at each slot to minimize the total energy consumption and meet the EH requirements. We model the problem as a Markov decision process and propose a solution relying on a low-complexity beamforming and deep deterministic policy gradient (DDPG). The results show that the proposed beamforming achieves near-optimal performance with low computational complexity, and the DDPG-based approach converges with the number of episodes and reduces the system's power consumption, while the outage probability and the power consumption increase with the number of devices.
format Preprint
id arxiv_https___arxiv_org_abs_2405_04218
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Deep Reinforcement Learning for Multi-User RF Charging with Non-linear Energy Harvesters
Azarbahram, Amirhossein
López, Onel L. A.
Popovski, Petar
Pandey, Shashi Raj
Latva-aho, Matti
Signal Processing
Radio frequency (RF) wireless power transfer (WPT) is a promising technology for sustainable support of massive Internet of Things (IoT). However, RF-WPT systems are characterized by low efficiency due to channel attenuation, which can be mitigated by precoders that adjust the transmission directivity. This work considers a multi-antenna RF-WPT system with multiple non-linear energy harvesting (EH) nodes with energy demands changing over discrete time slots. This leads to the charging scheduling problem, which involves choosing the precoders at each slot to minimize the total energy consumption and meet the EH requirements. We model the problem as a Markov decision process and propose a solution relying on a low-complexity beamforming and deep deterministic policy gradient (DDPG). The results show that the proposed beamforming achieves near-optimal performance with low computational complexity, and the DDPG-based approach converges with the number of episodes and reduces the system's power consumption, while the outage probability and the power consumption increase with the number of devices.
title Deep Reinforcement Learning for Multi-User RF Charging with Non-linear Energy Harvesters
topic Signal Processing
url https://arxiv.org/abs/2405.04218