2025
FactEval: Evaluating the Robustness of Fact Verification Systems in the Era of Large Language Models
NAACL 2025long
Whilst large language models (LLMs) have made significant advances in every natural language processing task, studies have shown that these models are vulnerable to small perturbations in the inputs, raising concerns about their robustness in the real-world. Given the rise of misinformation online a…