A Survey on Actionable Interpretability in Large Language Models
Large Language Models (LLMs) have become central to modern AI, with interpretability serving as a critical means of investigating the opaque and highly nonlinear mechanisms encoded within billions of parameters and ensuring trustworthy deployment. However, descriptive interpretability approaches for