2024
DiLM: Distilling Dataset into Language Model for Text-level Dataset Distillation
NAACL 2024findings
Dataset distillation aims to compress a training dataset by creating a small number of informative synthetic samples such that neural networks trained on them perform as well as those trained on the original training dataset. Current text dataset distillation methods create each synthetic sample as…