2026
Safe Multi-Agent Reinforcement Learning via Distributional Safety Critic and Maximum Entropy Optimization
AAAI 2026technical
Deploying multi-agent reinforcement learning (MARL) in safety-critical systems faces significant challenges due to insufficient agent exploration and inadequate safety constraint guarantees. Current approaches are constrained by two fundamental limitations: inefficient exploration leading to subopti