2018
Overcoming Language Priors in Visual Question Answering with Adversarial Regularization
NeurIPS 2018poster
Modern Visual Question Answering (VQA) models have been shown to rely heavily on superficial correlations between question and answer words learned during training -- \eg overwhelmingly reporting the type of room as kitchen or the sport being played as tennis, irrespective of the image. Most alarmin…