2025
Interpretable Next-token Prediction via the Generalized Induction Head
NeurIPS 2025poster
While large transformer models excel in predictive performance, their lack of interpretability restricts their usefulness in high-stakes domains. To remedy this, we propose the Generalized Induction-Head Model (GIM), an interpretable model for next-token prediction inspired by the observation of “in…