← Search

Stefan Ruseti

3 accepted papers

2025

The Strawberry Problem: Emergence of Character-level Understanding in Tokenized Language Models

EMNLP 2025

Despite their remarkable progress across diverse domains, Large Language Models (LLMs) consistently fail at simple character-level tasks, such as counting letters in words, due to a fundamental limitation: tokenization. In this work, we frame this limitation as a problem of low mutual information an

2024

How Hard is this Test Set? NLI Characterization by Exploiting Training Dynamics

EMNLP 2024main

Natural Language Inference (NLI) evaluation is crucial for assessing language understanding models; however, popular datasets suffer from systematic spurious correlations that artificially inflate actual model performance. To address this, we propose a method for the automated creation of a challeng…