← Search

Shaohua Wan

2 accepted papers

2026

CasMoE: A Cascaded Framework for Efficient MoE Inference on Resource-constrained Devices

AAAI 2026technical

The Mixture-of-Experts (MoE) architecture has emerged as a key enabler for scaling large language models (LLMs), empowering increased model capacity with minimal computational overhead through gating-based dynamic expert activation. However, due to the memory demands introduced by expert modules, Mo

Cited by 0SourcePDFScholar
2025

Deconfound Semantic Shift and Incompleteness in Incremental Few-shot Semantic Segmentation

AAAI 2025technical

Incremental few-shot semantic segmentation (IFSS) expands segmentation capacity of the trained model to segment new-class images with few samples. However, semantic meanings may shift from background to object class or vice versa during incremental learning. Moreover, new-class samples often lack re…

Cited by 0SourcePDFScholar