2024
LLMs as Bridges: Reformulating Grounded Multimodal Named Entity Recognition
ACL 2024findings
Grounded Multimodal Named Entity Recognition (GMNER) is a nascent multimodal task that aims to identify named entities, entity types and their corresponding visual regions. GMNER task exhibits two challenging properties: 1) The weak correlation between image-text pairs in social media results in a s…