2026
CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing
AAAI 2026technical
Recent advancements in Large Language Models (LLMs) have led to their widespread adoption in daily applications. Despite their impressive capabilities, they remain vulnerable to adversarial attacks, as even minor meaning-preserving changes such as synonym substitutions can lead to incorrect predicti