2025
GenPO: Generative Diffusion Models Meet On-Policy Reinforcement Learning
NeurIPS 2025poster
Recent advances in reinforcement learning (RL) have demonstrated the powerful exploration capabilities and multimodality of generative diffusion-based policies. While substantial progress has been made in offline RL and off-policy RL settings, integrating diffusion policies into on-policy frameworks…