← Search

Xiwen Wei

3 accepted papers

2026

Fuel Gauge: Estimating Chain-of-Thought Length Ahead of Time in Large Multimodal Models

CVPR 2026

Reasoning Large Multi-modality Models (LMMs) have become the de facto choice for many applications. However, these models rely on a Chain-of-Thought (CoT) process that is lengthy and unpredictable at runtime, often resulting in inefficient use of computational resources (due to memory fragmentation)

Cited by 0SourceScholar
2025

Mitigating Intra- and Inter-modal Forgetting in Continual Learning of Unified Multimodal Models

NeurIPS 2025poster

Unified Multimodal Generative Models (UMGMs) unify visual understanding and image generation within a single autoregressive framework. However, their ability to continually learn new tasks is severely hindered by catastrophic forgetting, both within a modality (intra-modal) and across modalities (in…

Cited by 0SourceScholar
2023

Text-Guided Unsupervised Latent Transformation for Multi-Attribute Image Manipulation

CVPR 2023poster

Great progress has been made in StyleGAN-based image editing. To associate with preset attributes, most existing approaches focus on supervised learning for semantically meaningful latent space traversal directions, and each manipulation step is typically determined for an individual attribute. To a…

Cited by 3SourcePDFScholar