← Search

Alberto Compagnoni

1 accepted papers

2026

ReAG: Reasoning-Augmented Generation for Knowledge-based Visual Question Answering

CVPR 2026

Multimodal Large Language Models (MLLMs) have shown impressive capabilities in jointly understanding text, images, and videos, often evaluated via Visual Question Answering (VQA). However, even state-of-the-art MLLMs struggle with domain-specific or knowledge-intensive queries, where relevant inform

Cited by 0SourcecodeScholar