2022
An Anchor-based Relative Position Embedding Method for Cross-Modal Tasks
EMNLP 2022main
Position Embedding (PE) is essential for transformer to capture the sequence ordering of input tokens. Despite its general effectiveness verified in Natural Language Processing (NLP) and Computer Vision (CV), its application in cross-modal tasks remains unexplored and suffers from two challenges: 1)…