2025
How to Generalize the Detection of AI-Generated Text: Confounding Neurons
EMNLP 2025
Detectors of LLM-generated text suffer from poor domain shifts generalization ability. Yet, reliable text detection methods in the wild are of paramount importance for plagiarism detection, integrity of the public discourse, and AI safety. Linguistic and domain confounders introduce spurious correla