2022
UTC: A Unified Transformer With Inter-Task Contrastive Learning for Visual Dialog
CVPR 2022poster
Visual Dialog aims to answer multi-round, interactive questions based on the dialog history and image content. Existing methods either consider answer ranking and generating individually or only weakly capture the relation across the two tasks implicitly by two separate models. The research on a uni…