← Search

Canyu Zhang

6 accepted papers

2026

InfoFlow KV: Information-Flow-Aware KV Recomputation for Long Context

ICML 2026poster

Retrieval-augmented generation (RAG) for long-context question answering is bottlenecked by inference-time prefilling over large retrieved contexts. A common strategy is to precompute key–value (KV) caches for individual documents and selectively recompute a small subset of tokens to restore global …

Cited by 0SourceScholar
2026

VIVA: VLM-Guided Instruction-Based Video Editing with Reward Optimization

CVPR 2026

Instruction-based video editing aims to modify an input video according to a natural-language instruction while preserving content fidelity and temporal coherence. However, existing diffusion-based approaches are often trained on paired data of simple editing operations, which fundamentally limits t

Cited by 0SourcecodeScholar
2024

Bidirectional Autoregessive Diffusion Model for Dance Generation

CVPR 2024poster

Dance serves as a powerful medium for expressing human emotions but the lifelike generation of dance is still a considerable challenge. Recently diffusion models have showcased remarkable generative abilities across various domains. They hold promise for human motion generation due to their adaptabl…

Cited by 8SourcePDFScholar
2024

EINet: Point Cloud Completion via Extrapolation and Interpolation

ECCV 2024poster

"Scanned point clouds are often sparse and incomplete due to the limited field of view of sensing devices, significantly impeding the performance of downstream applications. Therefore, the task of point cloud completion is introduced to obtain a dense and complete point cloud from the incomplete inp…

2023

Few-Shot 3D Point Cloud Semantic Segmentation via Stratified Class-Specific Attention Based Transformer Network

AAAI 2023technical

3D point cloud semantic segmentation aims to group all points into different semantic categories, which benefits important applications such as point cloud scene reconstruction and understanding. Existing supervised point cloud semantic segmentation methods usually require large-scale annotated poin…

2021

Fast Local Representation Learning with Adaptive Anchor Graph

ICASSP 2021accepted

Dimension reduction is an effective technology to embed data with high dimension to lower dimension space, where Linear Discriminant Analysis (LDA), one of representative methods, only works with Gaussian distribution data. However, in order to solve non-Gaussian issue that only one cluster cannot w…

Cited by 0SourceScholar