2022
MedCLIP: Contrastive Learning from Unpaired Medical Images and Text
EMNLP 2022main
Existing vision-text contrastive learning like CLIP aims to match the paired image and caption embeddings while pushing others apart, which improves representation transferability and supports zero-shot prediction. However, medical image-text datasets are orders of magnitude below the general images…