← Search

Sahasrajit Sarmasarkar

1 accepted papers

2025

Preference Learning with Response Time: Robust Losses and Guarantees

NeurIPS 2025poster

This paper investigates the integration of response time data into human preference learning frameworks for more effective reward model elicitation. While binary preference data has become fundamental in fine-tuning foundation models, generative AI systems, and other large-scale models, the valuable…

Cited by 0SourceScholar