← Search

Hsun-Yu Kuo

1 accepted papers

2025

Not All LLM-Generated Data Are Equal: Rethinking Data Weighting in Text Classification

ICLR 2025spotlight

Synthetic data augmentation via Large Language Models (LLMs) allows researchers to leverage additional training data, thus enhancing the performance of downstream tasks, especially when real-world data is scarce. However, the generated data can deviate from the real-world data, and this misalignment…

Cited by 1SourcePDFScholar