2026
RLSF-V: Mitigating Hallucinations in MLLMs via Fuzzy Semantic Self-Feedback
ICML 2026poster
Multimodal large language models (MLLMs) extend large language models (LLMs) with visual perception for open-world understanding, but exacerbate LLMs' hallucinations, in which generated text contradicts visual evidence or common sense. To mitigate hallucinations, a dominant strategy is Direct Prefer…