CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation
Medical vision-language models can automate the generation of radiology reports but struggle with accurate visual grounding and factual consistency. Existing models often misalign textual findings with visual evidence, leading to unreliable or weakly grounded predictions. We present "CURE", an error