← Search

Nathan Stringham

1 accepted papers

2024

Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness

NAACL 2024long

*Do larger and more performant models resolve NLP’s longstanding robustness issues?* We investigate this question using over 20 models of different sizes spanning different architectural choices and pretraining objectives. We conduct evaluations using (a) out-of-domain and challenge test sets, (b) b…