← Search

Richard Rutmann

1 accepted papers

2024

Tokenizer Choice For LLM Training: Negligible or Crucial?

NAACL 2024findings

The recent success of large language models (LLMs) has been predominantly driven by curating the training dataset composition, scaling of model architectures and dataset sizes and advancements in pretraining objectives, leaving tokenizer influence as a blind spot.Shedding light on this underexplored…