2025
Advantage-Guided Distillation for Preference Alignment in Small Language Models
ICLR 2025spotlight
Alignment techniques enable Large Language Models (LLMs) to generate outputs that align with human preferences and play a crucial role in their effectiveness. However, their impact often diminishes when applied to Small Language Models (SLMs), likely due to the limited capacity of these models. Inst…