2026
Learning Rewrite-Invariant Reasoning with Targeted Alternation Training
ICML 2026poster
Large language models (LLMs) often fail in systematic, model-specific ways under meaning-preserving question rewrites (paraphrases, format changes, benign distractors). In this work, we address this instability by identifying where the model's reasoning process diverges across semantically-equivalen…