2026
ReMoE: Region-Mixture Experts for Adversarially-Robust Vision Transformers
CVPR 2026
Vision Transformers (ViTs) achieve state-of-the-art performance on a wide range of vision tasks, yet they remain highly vulnerable to adversarial perturbations due to the lack of explicit region-level semantic modeling. Adversarial perturbations are typically local and spatially structured, whereas