← Search

Enze Zhang

2 accepted papers

2025

Mean-Field Aided QMIX: A Scalable and Flexible Q-Learning Approach for Large-Scale Agent Groups

ICASSP 2025accepted

Value decomposition methods are effective for multi-agent reinforcement learning (MARL), with QMIX being one of the most advanced. However, it struggles with scalability and flexibility in large-scale agent systems. The introduction of mean-field theory into MARL provides a potential solution to bot…

Cited by 0SourceScholar
2025

Offline-to-Online Reinforcement Learning with Classifier-Free Diffusion Generation

ICML 2025poster

Offline-to-online Reinforcement Learning (O2O RL) aims to perform online fine-tuning on an offline pre-trained policy to minimize costly online interactions. Existing work used offline datasets to generate data that conform to the online data distribution for data augmentation. However, generated da…

Cited by 0SourcePDFScholar