2025
Icon2: Aligning Large Language Models Using Self-Synthetic Preference Data via Inherent Regulation
EMNLP 2025
Large Language Models (LLMs) require high quality preference datasets to align with human preferences. However, conventional methods for constructing such datasets face significant challenges: reliance on pre-collected instructions often leads to distribution mismatches with target models, while the