2025
ReFLAIR: Enhancing Multimodal Reasoning via Structured Reflection and Reward-Guided Learning
EMNLP 2025
Large models can achieve higher performance on complex problems through iterative self-reflection. Yet when reflection is uncontrolled, it often leads to longer outputs, higher inference cost, and an increased risk of hallucination. Existing training methods rarely address this trade off. We introdu