MedSIGHT: Towards Grounded Visual Comprehension in Medical Large Vision-Language Models
Medical large vision-language models (Med-LVLMs) have recently achieved remarkable progress in vision–language comprehension and medical image segmentation. However, existing models still struggle to unify these two capabilities, which is essential for achieving clinically reasoning that connects vi…