2025
Iterative Foundation Model Fine-Tuning on Multiple Rewards
NeurIPS 2025poster
Fine-tuning foundation models has emerged as a powerful approach for generating objects with specific desired properties. Reinforcement learning (RL) provides an effective framework for this purpose, enabling models to generate outputs that maximize a given reward function. However, in many applicat…