2025
Beyond Human Data: Aligning Multimodal Large Language Models by Iterative Self-Evolution
AAAI 2025technical
Human preference alignment can significantly enhance the capabilities of Multimodal Large Language Models (MLLMs). However, collecting high-quality preference data remains costly. One promising solution is the self-evolution strategy, where models are iteratively trained on data they generate. Curre…