← Search

Haoyou Deng

1 accepted papers

2026

DenseGRPO: From Sparse to Dense Reward for Flow Matching Model Alignment

ICLR 2026poster

Recent GRPO-based approaches built on flow matching models have shown remarkable improvements in human preference alignment for text-to-image generation. Nevertheless, they still suffer from the sparse reward problem: the terminal reward of the entire denoising trajectory is applied to all intermedi…

Cited by 0SourceScholar