2022
ITA: Image-Text Alignments for Multi-Modal Named Entity Recognition
NAACL 2022long
Recently, Multi-modal Named Entity Recognition (MNER) has attracted a lot of attention. Most of the work utilizes image information through region-level visual representations obtained from a pretrained object detector and relies on an attention mechanism to model the interactions between image and…