← Search

Kevin Minghan Zhang

1 accepted papers

2025

MedicalNarratives: Connecting Medical Vision and Language with Localized Narratives

NeurIPS 2025poster

Multi-modal models are data hungry. While datasets with natural images are abundant, medical image datasets can not afford the same luxury. To enable representation learning for medical images at scale, we turn to YouTube, a platform with a large reservoir of open-source medical pedagogical videos.…

Cited by 0SourceScholar