2024
Enhancing Image-to-Text Generation in Radiology Reports through Cross-modal Multi-Task Learning
COLING 2024main
Image-to-text generation involves automatically generating descriptive text from images and has applications in medical report generation. However, traditional approaches often exhibit a semantic gap between visual and textual information. In this paper, we propose a multi-task learning framework to…