← Search

Felipe Pinto Coelho Nuti

2 accepted papers

2025

TuCo: Measuring the Contribution of Fine-Tuning to Individual Responses of LLMs

ICML 2025poster

Past work has studied the effects of fine-tuning on large language models' (LLMs) overall performance on certain tasks. However, a way to quantitatively and systematically analyze its effect on individual outputs is still lacking. In this work, we propose a new method for measuring the contribution…

Cited by 0SourcePDFScholar
2023

Extracting Reward Functions from Diffusion Models

NeurIPS 2023poster

Diffusion models have achieved remarkable results in image generation, and have similarly been used to learn high-performing policies in sequential decision-making tasks. Decision-making diffusion models can be trained on lower-quality data, and then be steered with a reward function to generate ne…