← Search

Zhaoyi Zhang

1 accepted papers

2025

When Open-Vocabulary Visual Question Answering Meets Causal Adapter: Benchmark and Approach

AAAI 2025technical

Visual Question Answering (VQA) is a multifaceted task that integrates computer vision and natural language processing to produce textual answers from images and questions. Existing VQA benchmarks predominantly adhere to a closed-set paradigm, limiting their ability to address arbitrary, unseen answ…

Cited by 0SourcePDFScholar