2025
Param$\Delta$ for Direct Mixing: Post-Train Large Language Model At Zero Cost
ICLR 2025poster
The post-training phase of large language models is essential for enhancing capabilities such as instruction-following, reasoning, and alignment with human preferences. However, it demands extensive high-quality data and poses risks like overfitting, alongside significant computational costs due to…