← Search

Yukti Makhija

3 accepted papers

2025

FRACTAL: Fine-Grained Scoring from Aggregate Text Labels

ACL 2025long

Fine-Tuning of LLMs using RLHF / RLAIF has been shown as a critical step to improve the performance of LLMs in complex generation tasks. These methods typically use response-level human or model feedback for alignment. Recent works indicate that finer sentence or span-level labels provide more accur…

Cited by 0SourcePDFScholar