2023
LM vs LM: Detecting Factual Errors via Cross Examination
EMNLP 2023long main
A prominent weakness of modern language models (LMs) is their tendency to generate factually incorrect text, which hinders their usability. A natural question is whether such factual errors can be detected automatically. Inspired by truth-seeking mechanisms in law, we propose a factuality evaluation…