← Search

Siyao Zhao

1 accepted papers

2025

Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning

AAAI 2025technical

Guiding the policy of multi-agent reinforcement learning to align with human common sense is a difficult problem, largely due to the complexity of modeling common sense as a reward, especially in complex and long-horizon multi-agent tasks. Recent works have shown the effectiveness of reward shaping,…