← Search

Yongquan Feng

2 accepted papers

2024

Modality Re-Balance for Visual Question Answering: A Causal Framework

ICASSP 2024accepted

Visual Question Answering (VQA) models often prioritize language cues over visual knowledge, leading to the "language prior" phenomenon. To address this, researchers have proposed methods to balance language and image information during training and inference. However, these approaches often struggl…

Cited by 0SourceScholar
2023

Memory-based Exploration-value Evaluation Model for Visual Navigation

ICRA 2023poster

We propose a hierarchical visual navigation solution, called Memory-based Exploration-value Evaluation Model (MEEM), to improve the agent's navigation performance. MEEM employs a hierarchical policy to tackle the challenge of sparse rewards, holds an episodic memory to store the historical informati…

Cited by 1SourceScholar