← Search

Feipeng Ma

3 accepted papers

2025

Incomplete Multi-modal Brain Tumor Segmentation via Learnable Sorting State Space Model

CVPR 2025poster

Brain tumor segmentation plays a crucial role in clinical diagnosis, yet the frequent unavailability of certain MRI modalities poses a significant challenge. In this paper, we introduce the Learnable Sorting State Space Model (LS3M), a novel framework designed to maximize the utilization of availabl…

Cited by 0SourcePDFScholar
2024

Image Captioning with Multi-Context Synthetic Data

AAAI 2024technical

Image captioning requires numerous annotated image-text pairs, resulting in substantial annotation costs. Recently, large models (e.g. diffusion models and large language models) have excelled in producing high-quality images and text. This potential can be harnessed to create synthetic image-text p…

Cited by 13SourcePDFScholar
2024

Visual Perception by Large Language Model’s Weights

NeurIPS 2024poster

Existing Multimodal Large Language Models (MLLMs) follow the paradigm that perceives visual information by aligning visual features with the input space of Large Language Models (LLMs) and concatenating visual tokens with text tokens to form a unified sequence input for LLMs. These methods demonstra…