2026
Finding the Translation Switch: Discovering and Exploiting the Task-Initiation Features in LLMs
AAAI 2026technical
Large Language Models (LLMs) frequently exhibit strong translation abilities, even without task-specific fine-tuning. However, the internal mechanisms governing this innate capability remain largely opaque. To demystify this process, we leverage Sparse Autoencoders (SAEs) and introduce a novel frame