arXiv · 2512.15062
Deep Reinforcement Learning for Joint Time and Power Management in SWIPT-EH CIoT
Abstract
This letter presents a novel deep reinforcement learning (DRL) approach for joint time allocation and power control in a cognitive Internet of Things (CIoT) system with simultaneous wireless information and power transfer (SWIPT). The CIoT transmitter autonomously manages energy harvesting (EH) and transmissions using a learnable time switching factor while optimizing power to enhance throughput and lifetime. The joint optimization is modeled as a Markov decision process under small-scale fading, realistic EH, and interference constraints. We develop a double deep Q-network (DDQN) enhanced with an upper confidence bound. Simulations benchmark our approach, showing superior performance over existing DRL methods.
Explore related subjects
Keep this discovery
Nadia Abdolkhani, Nada Abdel Khalek, Walaa Hamouda, Iyad Dayoub. 2025-12-17. Deep Reinforcement Learning for Joint Time and Power Management in SWIPT-EH CIoT. https://doi.org/10.1109/lcomm.2025.3536182.
Cite the original work for its findings. Save a collection to share your selection of sources.