2025
Remarkable Robustness of LLMs: Stages of Inference?
NeurIPS 2025poster
We investigate the robustness of Large Language Models (LLMs) to structural interventions by deleting and swapping adjacent layers during inference. Surprisingly, models retain 72–95\% of their original top-1 prediction accuracy without any fine-tuning. We find that performance degradation is not un…