← Search

Wang Shen

4 accepted papers

2026

LearniBridge: Learnable Calibration of Feature Caching for Diffusion Models Acceleration

ICML 2026poster

Diffusion Transformers (DiTs) have driven substantial progress in image and video generation but suffer from prohibitive computational costs. Feature caching accelerates inference by reusing intermediate representations. Existing methods rely on historical features for implementation simplicity, yet…

Cited by 0SourceScholar
2025

RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations

IJCAI 2025

Key-Value (KV) cache facilitates efficient large language models (LLMs) inference by avoiding recomputation of past KVs. As the batch size and context length increase, the oversized KV caches become a significant memory bottleneck, highlighting the need for efficient compression. Existing KV quantiz

Cited by 0SourcePDFScholar
2021

Dual Attention Guided Gaze Target Detection in the Wild

CVPR 2021poster

Gaze target detection aims to infer where each person in a scene is looking. Existing works focus on 2D gaze and 2D saliency, but fail to exploit 3D contexts. In this work, we propose a three-stage method to simulate the human gaze inference behavior in 3D space. In the first stage, we introduce a c…

Cited by 91PDFcodeScholar