arXiv · 2609.07618
Decentralized Safe Multi-Agent Reinforcement Learning via Predictive Shielding
Abstract
Environments are increasingly populated by multiple robots performing independent tasks with limited prior knowledge of each other. Deploying such multi-agent systems presents significant challenges. Specifically, shifts in deployment states compared to training data can lead to poor policy performance and compromised safety. While safety shields exist to mitigate these risks, they are typically reactive, which degrades performance near unseen obstacles,and centralized, limiting their scalability. To address this, we propose a decentralized framework that integrates predictive shielding with model-based finite horizon Q-learning. This approach allows agents to safely adapt their pre-trained policies during deployment. Furthermore, to mitigate livelocks in symmetric scenarios, we introduce a communication- free protocol for conflict resolution
Explore related subjects
Keep this discovery
Yacine El Yamani, Hanna Krasowski, Elena Vanneaux. 2026-09-07. Decentralized Safe Multi-Agent Reinforcement Learning via Predictive Shielding. https://arxiv.org/abs/2609.07618
Cite the original work for its findings. Save a collection to share your selection of sources.
Discover connections
Connections use source metadata and explicit phrase matches, not verified experimental comparisons.