← Search

Priyanka Agrawal

3 accepted papers

2025

FRACTAL: Fine-Grained Scoring from Aggregate Text Labels

ACL 2025long

Fine-Tuning of LLMs using RLHF / RLAIF has been shown as a critical step to improve the performance of LLMs in complex generation tasks. These methods typically use response-level human or model feedback for alignment. Recent works indicate that finer sentence or span-level labels provide more accur…

Cited by 0SourcePDFScholar
2023

Benchmarking Large Language Model Capabilities for Conditional Generation

ACL 2023long

Pre-trained large language models (PLMs) underly most new developments in natural language processing. They have shifted the field from application-specific model pipelines to a single model that is adapted to a wide range of tasks. Autoregressive PLMs like GPT-3 or PaLM and associated techniques li…

2018

On Controllable Sparse Alternatives to Softmax

NeurIPS 2018poster

Converting an n-dimensional vector to a probability distribution over n objects is a commonly used component in many machine learning tasks like multiclass classification, multilabel classification, attention mechanisms etc. For this, several probability mapping functions have been proposed and empl…

Cited by 76SourcePDFScholar