2025
Preference Learning with Response Time: Robust Losses and Guarantees
NeurIPS 2025poster
This paper investigates the integration of response time data into human preference learning frameworks for more effective reward model elicitation. While binary preference data has become fundamental in fine-tuning foundation models, generative AI systems, and other large-scale models, the valuable…