← Search

Yuhao Xu

5 accepted papers

2026

AutoQVLA: Not All Channels Are Equal in Vision-Language-Action Model's Quantization

ICLR 2026poster

The advent of Vision-Language-Action (VLA) models represents a significant leap for embodied intelligence, yet their immense computational demands critically hinder deployment on resource-constrained robotic platforms. Intuitively, low-bit quantization is a prevalent and preferred technique for larg…

Cited by 0SourcecodeScholar
2026

SE-Diff: Simulator and Experience Enhanced Diffusion Model for Comprehensive ECG Generation

ICLR 2026poster

Cardiovascular disease (CVD) is a leading cause of mortality worldwide. Electrocardiograms (ECGs) are the most widely used non-invasive tool for cardiac assessment, yet large, well-annotated ECG corpora are scarce due to cost, privacy, and workflow constraints. Generating ECGs can aid mechanistic un…

Cited by 0SourcecodeScholar
2025

Each Complexity Deserves a Pruning Policy

NeurIPS 2025poster

The established redundancy in visual tokens within large vision–language models (LVLMs) allows for pruning to effectively reduce their substantial computational demands. Empirical evidence from previous works indicates that visual tokens in later decoder stages receive less attention than shallow la…

Cited by 0SourcecodeScholar
2025

OOTDiffusion: Outfitting Fusion Based Latent Diffusion for Controllable Virtual Try-On

AAAI 2025technical

We present OOTDiffusion, a novel network architecture for realistic and controllable image-based virtual try-on (VTON). We leverage the power of pretrained latent diffusion models, designing an outfitting UNet to learn the detailed garment features. Without a redundant warping process, the garment f…

2023

AugTarget Data Augmentation for Infrared Small Target Detection

ICASSP 2023accepted

Sample shortage has always been a frequently-faced problem for the machine-learning models in infrared small target detection. As one of main limitations, it is hampering the further promotion of target detection performance. In this paper, we propose a simple and effective data augmentation scheme,…

Cited by 0SourceScholar