← Search

Guiqin Wang

3 accepted papers

2026

MAR: EFFICIENT LARGE LANGUAGE MODELS VIA MODULE-AWARE ARCHITECTURE REFINEMENT

ICASSP 2026poster

Large Language Models (LLMs) excel across diverse domains but suffer from high energy costs due to quadratic attention and dense Feed-Forward Network (FFN) operations. To address these issues, we propose Module-aware Architecture Refinement (MAR), a two-stage framework that integrates State Space Mo…

Cited by 0SourcePDFScholar
2024

Generative Model-Based Feature Knowledge Distillation for Action Recognition

AAAI 2024technical

Knowledge distillation (KD), a technique widely employed in computer vision, has emerged as a de facto standard for improving the performance of small neural networks. However, prevailing KD-based approaches in video tasks primarily focus on designing loss functions and fusing cross-modal informatio…

2023

Weakly-Supervised Action Localization by Hierarchically-Structured Latent Attention Modeling

ICCV 2023poster

Weakly-supervised action localization aims to recognize and localize action instancese in untrimmed videos with only video-level labels. Most existing models rely on multiple instance learning(MIL), where the predictions of unlabeled instances are supervised by classifying labeled bags. The MIL-base…

Cited by 4PDFcodeScholar