2026
Holdout-Loss-Based Data Selection for LLM Finetuning via In-Context Learning
ICLR 2026poster
Fine-tuning large pretrained language models is a common approach for aligning them with human preferences, but noisy or off-target examples can dilute supervision. While small, well-chosen datasets often match the performance of much larger ones, systematic and efficient ways to identify high-value…