← Search

Hu Fu

5 accepted papers

2026

HPS: Hyperspherical Parameter Sharing for Efficient Multi-Agent Reinforcement Learning

ICML 2026poster

Parameter Sharing (PS) is widely used to improve efficiency in Multi-Agent Reinforcement Learning (MARL), but it can limit behavioral diversity and degrade performance. This limitation stems from gradient conflicts among agents on shared weights, which hinders effective policy learning. To fully cha…

Cited by 0SourceScholar
2025

Incentives for Early Arrival in Cooperative Games (Extended Abstract)

IJCAI 2025

We study cooperative games where players join sequentially, and the value generated by those who have joined at any point must be irrevocably divided among these players. We introduce two desiderata for the value division mechanism: that the players should have incentives to join as early as possibl

Cited by 0SourcePDFScholar
2023

On the Last-iterate Convergence in Time-varying Zero-sum Games: Extra Gradient Succeeds where Optimism Fails

NeurIPS 2023poster

Last-iterate convergence has received extensive study in two player zero-sum games starting from bilinear, convex-concave up to settings that satisfy the MVI condition. Typical methods that exhibit last-iterate convergence for the aforementioned games include extra-gradient (EG) and optimistic gradi…

Cited by 12SourcePDFScholar