← Search

Seojin Park

1 accepted papers

2026

Semi-Supervised Preference Optimization with Limited Feedback

ICLR 2026oral

The field of preference optimization has made outstanding contributions to the alignment of language models with human preferences. Despite these advancements, recent methods still rely heavily on substantial paired (labeled) feedback data, leading to substantial resource expenditures. To address th…

Cited by 0SourcecodeScholar