2026
Value Matching: Scalable and Gradient-Free Reward-Guided Flow Adaptation
ICLR 2026poster
Adapting large-scale flow and diffusion models to downstream tasks through reward optimization is essential for their adoption in real-world applications, including scientific discovery and image generation. While recent fine-tuning methods based on reinforcement learning and stochastic optimal cont…