← Search

Gusti Triandi Winata

1 accepted papers

2025

M-RewardBench: Evaluating Reward Models in Multilingual Settings

ACL 2025long

Reward models (RMs) have driven the state-of-the-art performance of LLMs today by enabling the integration of human feedback into the language modeling process. However, RMs are primarily trained and evaluated in English, and their capabilities in multilingual settings remain largely understudied. I…