arXiv · 2411.01766
Lyapunov-guided Multi-Agent Reinforcement Learning for Delay-Sensitive Wireless Scheduling
Abstract
In this paper, a two-stage intelligent scheduler is proposed to minimize the packet-level delay jitter while guaranteeing delay bound. Firstly, Lyapunov technology is employed to transform the delay-violation constraint into a sequential slot-level queue stability problem. Secondly, a hierarchical scheme is proposed to solve the resource allocation between multiple base stations and users, where the multi-agent reinforcement learning (MARL) gives the user priority and the number of scheduled packets, while the underlying scheduler allocates the resource. Our proposed scheme achieves lower delay jitter and delay violation rate than the Round-Robin Earliest Deadline First algorithm and MARL with delay violation penalty.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Cheng Zhang, Lan Wei, Ji Fan, Zening Liu, Yongming Huang. 2024-11-04. Lyapunov-guided Multi-Agent Reinforcement Learning for Delay-Sensitive Wireless Scheduling. https://arxiv.org/abs/2411.01766
Cite the original work for its findings. Save a collection to share your selection of sources.