← Search

Thomas Dooms

2 accepted papers

2025

Bilinear MLPs enable weight-based mechanistic interpretability

ICLR 2025spotlight

A mechanistic understanding of how MLPs do computation in deep neural net- works remains elusive. Current interpretability work can extract features from hidden activations over an input dataset but generally cannot explain how MLP weights construct features. One challenge is that element-wise nonli…

2025

Parameterized Synthetic Text Generation with SimpleStories

NeurIPS 2025poster

We present SimpleStories, a large synthetic story dataset in simple language, consisting of 2 million samples each in English and Japanese. Through parameterizing prompts at multiple levels of abstraction, we achieve control over story characteristics at scale, inducing syntactic and semantic divers…

Cited by 0SourcecodeScholar