← Search

Zubin Trivadi Aysola

2 accepted papers

2025

Rejected Dialects: Biases Against African American Language in Reward Models

NAACL 2025findings

Preference alignment via reward models helps build safe, helpful, and reliable large language models (LLMs). However, subjectivity in preference judgments and the lack of representative sampling in preference data collection can introduce new biases, hindering reward models’ fairness and equity. In…

2021

RedCaps: Web-curated image-text data created by the people, for the people

NeurIPS 2021poster

Large datasets of paired images and text have become increasingly popular for learning generic representations for vision and vision-and-language tasks. Such datasets have been built by querying search engines or collecting HTML alt-text – since web data is noisy, they require complex filtering pipe…

Cited by 183SourcecodeScholar