2024
ULTRAFEEDBACK: Boosting Language Models with Scaled AI Feedback
ICML 2024poster
Learning from human feedback has become a pivot technique in aligning large language models (LLMs) with human preferences. However, acquiring vast and premium human feedback is bottlenecked by time, labor, and human capability, resulting in small sizes or limited topics of current datasets. This fur…