2024
Granular Entity Mapper: Advancing Fine-grained Multimodal Named Entity Recognition and Grounding
EMNLP 2024finding
Multimodal Named Entity Recognition and Grounding (MNERG) aims to extract paired textual and visual entities from texts and images. It has been well explored through a two-step paradigm: initially identifying potential visual entities using object detection methods and then aligning the extracted te…