← Search

Freya Behrens

4 accepted papers

2026

Dataset Distillation for Memorized Data: Soft Labels can Leak Held-Out Teacher Knowledge

ICLR 2026poster

Dataset distillation aims to compress training data into fewer examples via a teacher, from which a student can learn effectively. While its success is often attributed to structure in the data, modern neural networks also memorize specific facts, but if and how such memorized information can be tra…

Cited by 0SourcecodeScholar
2025

Counting in Small Transformers: The Delicate Interplay between Attention and Feed-Forward Layers

ICML 2025poster

Next to scaling considerations, architectural design choices profoundly shape the solution space of transformers. In this work, we analyze the solutions simple transformer blocks implement when tackling the histogram task: counting items in sequences. Despite its simplicity, this task reveals a comp…

2024

A Phase Transition between Positional and Semantic Learning in a Solvable Model of Dot-Product Attention

NeurIPS 2024spotlight

Many empirical studies have provided evidence for the emergence of algorithmic mechanisms (abilities) in the learning of language models, that lead to qualitative improvements of the model capabilities. Yet, a theoretical characterization of how such mechanisms emerge remains elusive. In this paper,…

Cited by 13SourcePDFScholar