← Search

Jaime Sevilla

2 accepted papers

2024

Algorithmic progress in language models

NeurIPS 2024poster

We investigate the rate at which algorithms for pre-training language models have improved since the advent of deep learning. Using a dataset of over 200 language model evaluations on Wikitext and Penn Treebank spanning 2012-2023, we find that the compute required to reach a set performance threshol…

2024

Position: Will we run out of data? Limits of LLM scaling based on human-generated data

ICML 2024poster

We investigate the potential constraints on LLM scaling posed by the availability of public human-generated text data. We forecast the growing demand for training data based on current trends and estimate the total stock of public human text data. Our findings indicate that if current LLM developmen…

Cited by 182SourcePDFScholar