← Search

Neil Rohit Mallinar

4 accepted papers

2026

FACT: a first-principles alternative to the Neural Feature Ansatz for how networks learn representations

ICLR 2026poster

It is a central challenge in deep learning to understand how neural networks learn representations. A leading approach is the Neural Feature Ansatz (NFA) (Radhakrishnan et al., 2024), a conjectured mechanism for how feature learning occurs. Although the NFA is empirically validated, it is an educate…

Cited by 0SourceScholar
2025

Emergence in non-neural models: grokking modular arithmetic via average gradient outer product

ICML 2025oral

Neural networks trained to solve modular arithmetic tasks exhibit grokking, a phenomenon where the test accuracy starts improving long after the model achieves 100% training accuracy in the training process. It is often taken as an example of "emergence", where model ability manifests sharply throug…

Cited by 6SourcePDFScholar
2022

Benign, Tempered, or Catastrophic: Toward a Refined Taxonomy of Overfitting

NeurIPS 2022accept

The practical success of overparameterized neural networks has motivated the recent scientific study of \emph{interpolating methods}-- learning methods which are able fit their training data perfectly. Empirically, certain interpolating methods can fit noisy training data without catastrophically ba…

Cited by 45SourcePDFScholar