← Search

Laura Pérez-Mayos

1 accepted papers

2021

How much pretraining data do language models need to learn syntax?

EMNLP 2021main

Transformers-based pretrained language models achieve outstanding results in many well-known NLU benchmarks. However, while pretraining methods are very convenient, they are expensive in terms of time and resources. This calls for a study of the impact of pretraining data size on the knowledge of th…

Cited by 42SourcePDFScholar