2022
The Inductive Bias of In-Context Learning: Rethinking Pretraining Example Design
ICLR 2022spotlight
Pretraining Neural Language Models (NLMs) over a large corpus involves chunking the text into training examples, which are contiguous text segments of sizes processable by the neural architecture. We highlight a bias introduced by this common practice: we prove that the pretrained NLM can model much…