← Search

Areeb Ahmad

4 accepted papers

2025

Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits

NeurIPS 2025poster

Transformer-based language models exhibit complex behavior, but their internal computations remain poorly understood. Most mechanistic interpretability approaches treat components, such as attention heads and MLPs, as atomic units, ignoring potential functional substructure. We propose a finer-grain…

Cited by 0SourceScholar
2023

ScriptWorld: Text Based Environment for Learning Procedural Knowledge

IJCAI 2023poster

Text-based games provide a framework for developing natural language understanding and commonsense knowledge about the world in reinforcement learning based agents. Existing text-based environments often rely on fictional situations and characters to create a gaming framework and are far from real-w…