← Search

Changming Sun

10 accepted papers

2026

D2Cache: Second-Order Delta Caching for Higher Video Diffusion Acceleration

CVPR 2026

Video diffusion models achieve impressive visual fidelity but remain computationally prohibitive for real-time or interactive generation due to their sequential denoising process. Recent caching methods accelerate inference by reusing outputs across timesteps, typically estimating each new output fr

Cited by 0SourcecodeScholar
2026

MDCS-MoAME: Multi-directional Composite Scanning with Mixture of Attention and Mamba Experts for Cancer Survival Prediction

CVPR 2026

Multi-modal learning approaches that integrate pathological images with genomic profiles have significantly enhanced the accuracy of survival prediction tasks. However, previous methods often struggle to effectively process long-range gigapixel whole slide images (WSIs) and sparse genomic profiles d

Cited by 0SourceScholar
2023

Calibrating a Deep Neural Network with Its Predecessors

IJCAI 2023poster

Confidence calibration - the process to calibrate the output probability distribution of neural networks - is essential for safety-critical applications of such networks. Recent works verify the link between mis-calibration and overfitting. However, early stopping, as a well-known technique to mitig…

2023

Learning Partial Correlation Based Deep Visual Representation for Image Classification

CVPR 2023poster

Visual representation based on covariance matrix has demonstrates its efficacy for image classification by characterising the pairwise correlation of different channels in convolutional feature maps. However, pairwise correlation will become misleading once there is another channel correlating with…

2020

Instance-Aware Embedding for Point Cloud Instance Segmentation

ECCV 2020poster

Although recent works have made significant progress in encoding meaningful context information for instance segmentation in 2D images, the works for 3D point cloud counterpart lag far behind. Conventional methods use radius search or other similar methods for aggregating local information. However,…

Cited by 24SourcePDFScholar
2020

ReDro: Efficiently Learning Large-sized SPD Visual Representation

ECCV 2020poster

Symmetric positive definite (SPD) matrix has recently been used as an effective visual representation. When learning this representation in deep networks, eigen-decomposition of covariance matrix is usually needed for a key step called matrix normalisation. This could result in significant computati…

Cited by 12SourcePDFScholar
2019

Knowledge Adaptation for Efficient Semantic Segmentation

CVPR 2019poster

Both accuracy and efficiency are of significant importance to the task of semantic segmentation. Existing deep FCNs suffer from heavy computations due to a series of high-resolution feature maps for preserving the detailed knowledge in dense estimation. Although reducing the feature map resolution (…

Cited by 291PDFScholar
2018

An End-to-End TextSpotter With Explicit Alignment and Attention

CVPR 2018poster

Text detection and recognition in natural images have long been considered as two separate tasks that are processed sequentially. Jointly training two tasks is non-trivial due to significant differences in learning difficulties and convergence rates. In this work, we present a conceptually simple ye…

2017

Can Walking and Measuring Along Chord Bunches Better Describe Leaf Shapes?

CVPR 2017poster

Effectively describing and recognizing leaf shapes under arbitrary deformations, particularly from a large database, remains an unsolved problem. In this research, we attempted a new strategy of describing shape by walking along a bunch of chords that pass through the shape to measure the regions tr…

Cited by 36PDFScholar