← Search

Tianxu Li

3 accepted papers

2024

Learning Distinguishable Trajectory Representation with Contrastive Loss

NeurIPS 2024poster

Policy network parameter sharing is a commonly used technique in advanced deep multi-agent reinforcement learning (MARL) algorithms to improve learning efficiency by reducing the number of policy parameters and sharing experiences among agents. Nevertheless, agents that share the policy parameters t…

Cited by 0SourcePDFScholar