2023
Rethinking Multimodal Entity and Relation Extraction from a Translation Point of View
ACL 2023long
We revisit the multimodal entity and relation extraction from a translation point of view. Special attention is paid on the misalignment issue in text-image datasets which may mislead the learning. We are motivated by the fact that the cross-modal misalignment is a similar problem of cross-lingual d…