2024
Human Guided Cross-Modal Reasoning with Semantic Attention Learning for Visual Question Answering
ICASSP 2024accepted
One of the major difficulties in the Visual Question Answering (VQA) task of real-world images is the long-tailed distribution of concepts which makes the model vulnerable to negative linguistic biases. To imitate human learning and reasoning, researchers have designed reasoning models, which, howev…