← Search

Negar Mokhberian

3 accepted papers

2025

A Systematic Analysis of Base Model Choice for Reward Modeling

EMNLP 2025

Reinforcement learning from human feedback (RLHF) and, at its core, reward modeling have become a crucial part of training powerful large language models (LLMs). One commonly overlooked factor in training high-quality reward models (RMs) is the effect of the base model, which is becoming more challe

Cited by 0SourcePDFScholar
2024

Capturing Perspectives of Crowdsourced Annotators in Subjective Learning Tasks

NAACL 2024long

Supervised classification heavily depends on datasets annotated by humans. However, in subjective tasks such as toxicity classification, these annotations often exhibit low agreement among raters. Annotations have commonly been aggregated by employing methods like majority voting to determine a sing…

2021

Detecting Polarized Topics Using Partisanship-aware Contextualized Topic Embeddings

EMNLP 2021finding

Growing polarization of the news media has been blamed for fanning disagreement, controversy and even violence. Early identification of polarized topics is thus an urgent matter that can help mitigate conflict. However, accurate measurement of topic-wise polarization is still an open research challe…