← Search

Zhongzhen Huang

4 accepted papers

2025

MMXU: A Multi-Modal and Multi-X-ray Understanding Dataset for Disease Progression

ACL 2025finding

Large vision-language models (LVLMs) have shown great promise in medical applications, particularly in visual question answering (MedVQA) and diagnosis from medical images. However, existing datasets and models often fail to consider critical aspects of medical diagnostics, such as the integration o…

2024

CAT: Coordinating Anatomical-Textual Prompts for Multi-Organ and Tumor Segmentation

NeurIPS 2024poster

Existing promptable segmentation methods in the medical imaging field primarily consider either textual or visual prompts to segment relevant objects, yet they often fall short when addressing anomalies in medical images, like tumors, which may vary greatly in shape, size, and appearance. Recognizin…

2024

ZePT: Zero-Shot Pan-Tumor Segmentation via Query-Disentangling and Self-Prompting

CVPR 2024poster

The long-tailed distribution problem in medical image analysis reflects a high prevalence of common conditions and a low prevalence of rare ones which poses a significant challenge in developing a unified model capable of identifying rare or novel tumor categories not encountered during training. In…

2023

KiUT: Knowledge-Injected U-Transformer for Radiology Report Generation

CVPR 2023poster

Radiology report generation aims to automatically generate a clinically accurate and coherent paragraph from the X-ray image, which could relieve radiologists from the heavy burden of report writing. Although various image caption methods have shown remarkable performance in the natural image field,…

Cited by 96SourcePDFScholar