2025
Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning
AAAI 2025technical
Guiding the policy of multi-agent reinforcement learning to align with human common sense is a difficult problem, largely due to the complexity of modeling common sense as a reward, especially in complex and long-horizon multi-agent tasks. Recent works have shown the effectiveness of reward shaping,…