2025
Attacking Misinformation Detection Using Adversarial Examples Generated by Language Models
EMNLP 2025
Large language models have many beneficial applications, but can they also be used to attack content-filtering algorithms in social media platforms? We investigate the challenge of generating adversarial examples to test the robustness of text classification algorithms detecting low-credibility cont