2026
Think-While-Generating: On-the-Fly Reasoning for Personalized Long-Form Generation
ICLR 2026poster
Preference alignment has enabled large language models (LLMs) to better reflect human expectations, but current methods mostly optimize for population-level preferences, overlooking individual users. Personalization is essential, yet early approaches—such as prompt customization or fine-tuning—strug…