← Search

Paul Kinahan

1 accepted papers

2026

Boosting Medical Visual Understanding From Multi-Granular Language Learning

ICLR 2026poster

Recent advances in image-text pretraining have significantly enhanced visual understanding by aligning visual and textual representations. Contrastive Language-Image Pretraining (CLIP) has played a pivotal role in multimodal learning. However, its focus on single-label, single-granularity alignment…

Cited by 3SourcecodeScholar