← Search

Rishanth Rajendhran

2 accepted papers

2025

VeriFastScore: Speeding up long-form factuality evaluation

EMNLP 2025

Metrics like FactScore and VeriScore that evaluate long-form factuality operate by decomposing an input response into atomic claims and then individually verifying each claim. While effective and interpretable, these methods incur numerous LLM calls and can take upwards of 100s to evaluate a single

2024

Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness

NAACL 2024long

*Do larger and more performant models resolve NLP’s longstanding robustness issues?* We investigate this question using over 20 models of different sizes spanning different architectural choices and pretraining objectives. We conduct evaluations using (a) out-of-domain and challenge test sets, (b) b…