← Search

Yongqi Chen

6 accepted papers

2025

Distributed LLM Serving on Consumer-Grade GPUs by Reconciling Computation and Communication

EMNLP 2025

Large language models are reshaping internet services. Serving these models is often costly, as it requires multiple high-end GPUs. Consumer-grade GPUs offer cheaper computational power, providing an opportunity for more cost-efficient LLM serving.Prior efforts have explored distributed serving at s

Cited by 0SourcePDFScholar
2025

Fast Video Generation with Sliding Tile Attention

ICML 2025poster

Diffusion Transformers (DiTs) with 3D full attention power state-of-the-art video generation, but suffer from prohibitive compute cost -- when generating just a 5-second 720P video, attention alone takes 800 out of 950 seconds of total inference time. This paper introduces sliding tile attention (ST…

Cited by 5SourcePDFScholar
2025

Faster Video Diffusion with Trainable Sparse Attention

NeurIPS 2025poster

Scaling video diffusion transformers (DiTs) is limited by their quadratic 3D attention, even though most of the attention mass concentrates on a small subset of positions. We turn this observation into VSA, a trainable, hardware-efficient sparse attention that replaces full attention at both traini…

Cited by 0SourcecodeScholar
2025

Scalable Benchmarking and Robust Learning for Noise-Free Ego-Motion and 3D Reconstruction from Noisy Video

ICLR 2025poster

We aim to redefine robust ego-motion estimation and photorealistic 3D reconstruction by addressing a critical limitation: the reliance on noise-free data in existing models. While such sanitized conditions simplify evaluation, they fail to capture the unpredictable, noisy complexities of real-world…

2024

A Lightweight Change Detection Method Based on Feature Interaction and Transformer for High Resolution Remote Sensing Images

ICASSP 2024accepted

Change detection has consistently been a prominent direction in the field of remote sensing. As for high resolution remote sensing images (HRRSI), despite the notable achievements of change detection models, the majority of their impressive performance stems from their large scale architecture or co…

Cited by 0SourceScholar
2024

CMA: A Chromaticity Map Adapter for Robust Detection of Screen-Recapture Document Images

CVPR 2024poster

The rebroadcasting of screen-recaptured document images introduces a significant risk to the confidential documents processed in government departments and commercial companies. However detecting recaptured document images subjected to distortions from online social networks (OSNs) is challenging si…

Cited by 2SourcePDFScholar