← Search

Rohan Gupta

4 accepted papers

2025

Efficient Evaluation of Multi-Task Robot Policies With Active Experiment Selection

CoRL 2025poster

Evaluating learned robot control policies to determine their performance costs the experimenter time and effort. As robots become more capable in accomplishing diverse tasks, evaluating across all these tasks becomes more difficult as it is impractical to test every policy on every task multiple tim…

Cited by 0SourceScholar
2025

MIB: A Mechanistic Interpretability Benchmark

ICML 2025poster

How can we know whether new mechanistic interpretability methods achieve real improvements? In pursuit of lasting evaluation standards, we propose MIB, a Mechanistic Interpretability Benchmark, with two tracks spanning four tasks and five models. MIB favors methods that precisely and concisely recov…

2024

InterpBench: Semi-Synthetic Transformers for Evaluating Mechanistic Interpretability Techniques

NeurIPS 2024poster

Mechanistic interpretability methods aim to identify the algorithm a neural network implements, but it is difficult to validate such methods when the true algorithm is unknown. This work presents InterpBench, a collection of semi-synthetic yet realistic transformers with known circuits for evaluatin…

Cited by 4SourcePDFScholar