2025
The Computational Complexity of Circuit Discovery for Inner Interpretability
ICLR 2025spotlight
Many proposed applications of neural networks in machine learning, cognitive/brain science, and society hinge on the feasibility of inner interpretability via circuit discovery. This calls for empirical and theoretical explorations of viable algorithmic options. Despite advances in the design and te…