← Search

Michelle S. Lam

1 accepted papers

2025

Aligning Language Models with Demonstrated Feedback

ICLR 2025poster

Language models are aligned to emulate the collective voice of many, resulting in outputs that align with no one in particular. Steering LLMs away from generic output is possible through supervised finetuning or RLHF, but requires prohibitively large datasets for new ad-hoc tasks. We argue that it i…