2026
Lookahead Sample Reward Guidance for Test-Time Scaling of Diffusion Models
ICML 2026spotlight
Diffusion models have demonstrated strong generative performance; however, generated samples often fail to fully align with human intent. This paper studies a test-time scaling method that enables sampling from regions with higher human-aligned reward values. Existing gradient guidance methods appro…