AAAI 2026technical0 citations
CharBench: Evaluating the Role of Tokenization in Character-Level Tasks
Abstract
Tasks that require character-level reasoning, such as counting or locating characters within words, remain challenging for contemporary language models. A common conjecture is that language models
BibTeX
@inproceedings{aaai2026_charbenchevaluat,
title = {CharBench: Evaluating the Role of Tokenization in Character-Level Tasks},
author = {Omri Uzan and Yuval Pinter},
booktitle = {AAAI 2026},
year = {2026}
}