← Search

Jonathan Zheng

3 accepted papers

2024

NEO-BENCH: Evaluating Robustness of Large Language Models with Neologisms

ACL 2024long

The performance of Large Language Models (LLMs) degrades from the temporal drift between data used for model training and newer text seen during inference. One understudied avenue of language change causing data drift is the emergence of neologisms – new word forms – over time. We create a diverse r…

2022

Stanceosaurus: Classifying Stance Towards Multicultural Misinformation

EMNLP 2022main

We present Stanceosaurus, a new corpus of 28,033 tweets in English, Hindi and Arabic annotated with stance towards 250 misinformation claims. As far as we are aware, it is the largest corpus annotated with stance towards misinformation claims. The claims in Stanceosaurus originate from 15 fact-check…

Cited by 18SourcePDFScholar