arXiv · 2610.09213
Mission-critical spectrum sharing with decentralized Multi-Agent Reinforcement Learning
Abstract
Motivated by emerging mission-critical applications and an increasingly congested spectrum, we develop a decentralized multi-agent reinforcement learning (MARL) model for dynamic spectrum access. The model enables secondary users to learn effective transmission strategies across shared frequency bands while minimizing collisions with high-priority primary users and among themselves. We design the agent-level learners following a Markov potential game approach, connecting independent local updates to system-level improvement. We instantiate this design using lightweight linear actor-critic learners suitable for resource-constrained edge devices, rather than computationally intensive centralized or deep multi-agent architectures. Across spectrum environments with different incumbent activities, the learned policies adapt their transmission policy and waiting behavior to preserve throughput while greatly reducing transmission collisions relative to random and forecast-aware heuristic baselines. The results establish the value of decentralized MARL and shows up to 96.8% reduction in overall collisions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Dimitrios Pylorof, Imtiaz Nasim, Humberto E. Garcia, Vivek Agarwal, Jasni A. Mannil, Mingyue Ji. 2026-10-06. Mission-critical spectrum sharing with decentralized Multi-Agent Reinforcement Learning. https://arxiv.org/abs/2610.09213
Cite the original work for its findings. Save a collection to share your selection of sources.