← Search

Anzhe Cheng

7 accepted papers

2026

ERMoE: Eigen-Reparameterized Mixture-of-Experts for Stable Routing and Interpretable Specialization

CVPR 2026

Mixture-of-Experts (MoE) models expand capacity via sparse expert activation, but routing logits can misalign with expert structure (unstable routing, underutilization) and load imbalance can create stragglers. Auxiliary load-balancing losses reduce disparity but often weaken specialization and down

Cited by 0SourceScholar
2026

Multi-scale Conditional Generative Modeling for Microscopic Image Restoration

ICASSP 2026oral

The advance of diffusion-based generative models in recent years has revolutionized state-of-the-art (SOTA) techniques in a wide variety of image analysis and synthesis tasks, whereas their adaptation on image restoration, particularly within computational microscopy remains theoretically and empiri…

Cited by 0SourcePDFScholar
2026

Multi-scale Generative Modeling for Fast Sampling

ICASSP 2026oral

While working within the spatial domain can pose problems associated with ill-conditioned scores caused by power-law decay, recent advances in diffusion-based generative models have shown that transitioning to the wavelet domain offers a promising alternative. However, within the wavelet domain, we…

Cited by 0SourcePDFScholar
2026

STRUCTURAL COMPLEXITY OF BRAIN MRI REVEALS AGE-ASSOCIATED PATTERNS

ICASSP 2026poster

We adapt structural complexity analysis to three-dimensional signals, with an emphasis on brain magnetic resonance imaging (MRI). This framework captures the multiscale organization of volumetric data by coarse-graining the signal at progressively larger spatial scales and quantifying the informatio…

Cited by 0SourcePDFScholar
2025

Exploiting Application-to-Architecture Dependencies for Designing Scalable OS

ICASSP 2025accepted

With the advent of hundreds of cores on a chip to accelerate applications, the operating system (OS) needs to exploit the existing parallelism provided by the underlying hardware resources to determine the right amount of processes to be mapped on the multi-core systems. However, the existing OS is…

Cited by 0SourceScholar
2024

Unlocking Deep Learning: A BP-Free Approach for Parallel Block-Wise Training of Neural Networks

ICASSP 2024accepted

Backpropagation (BP) has been a successful optimization technique for deep learning models. However, its limitations, such as backward- and update-locking, and its biological implausibility, hinder the concurrent updating of layers and do not mimic the local learning processes observed in the human…

Cited by 0SourceScholar