2025
InterMT: Multi-Turn Interleaved Preference Alignment with Human Feedback
NeurIPS 2025spotlight
As multimodal large models (MLLMs) continue to advance across challenging tasks, a key question emerges: \textbf{\textit{What essential capabilities are still missing? }} A critical aspect of human learning is continuous interaction with the environment -- not limited to language, but also involving…