← Search

Rongsheng Wang

6 accepted papers

2026

MedGMAE: Gaussian Masked Autoencoders for Medical Volumetric Representation Learning

ICLR 2026poster

Self-supervised pre-training has emerged as a critical paradigm for learning transferable representations from unlabeled medical volumetric data. Masked autoencoder based methods have garnered significant attention, yet their application to volumetric medical image faces fundamental limitations from…

Cited by 0SourcecodeScholar
2026

MicroVerse: A Preliminary Exploration Toward a Micro-World Simulation

ICLR 2026poster

Recent advances in video generation have opened new avenues for macroscopic simulation of complex dynamic systems, but their application to microscopic phenomena remains largely unexplored. Microscale simulation holds great promise for biomedical applications such as drug discovery, organ-on-chip sy…

Cited by 0SourcecodeScholar
2025

A General Knowledge Injection Framework for ICD Coding

ACL 2025finding

ICD Coding aims to assign a wide range of medical codes to a medical text document, which is a popular and challenging task in the healthcare domain. To alleviate the problems of long-tail distribution and the lack of annotations of code-specific evidence, many previous works have proposed incorpora…

2025

Exploring Compositional Generalization of Multimodal LLMs for Medical Imaging

ACL 2025long

Medical imaging provides essential visual insights for diagnosis, and multimodal large language models (MLLMs) are increasingly utilized for its analysis due to their strong generalization capabilities; however, the underlying factors driving this generalization remain unclear. Current research sugg…

2025

Towards Medical Complex Reasoning with LLMs through Medical Verifiable Problems

ACL 2025finding

The breakthrough of OpenAI o1 highlights the potential of enhancing reasoning to improve LLM. Yet, most research in reasoning has focused on mathematical tasks, leaving domains like medicine underexplored. The medical domain, though distinct from mathematics, also demands robust reasoning to provide…

2024

CARZero: Cross-Attention Alignment for Radiology Zero-Shot Classification

CVPR 2024poster

The advancement of Zero-Shot Learning in the medical domain has been driven forward by using pre-trained models on large-scale image-text pairs focusing on image-text alignment. However existing methods primarily rely on cosine similarity for alignment which may not fully capture the complex relatio…