← Search

Ying Peng

3 accepted papers

2026

Revisiting Cross-Architecture Distillation: Adaptive Dual-Teacher Transfer for Lightweight Video Models

AAAI 2026technical

Vision Transformers (ViTs) have achieved strong performance in video action recognition, but their high computational cost limits their practicality. Lightweight CNNs are more efficient but suffer from accuracy gaps. Cross-Architecture Knowledge Distillation (CAKD) addresses this by transferring kno

Cited by 0SourcePDFScholar
2025

GEVRM: Goal-Expressive Video Generation Model For Robust Visual Manipulation

ICLR 2025poster

With the rapid development of embodied artificial intelligence, significant progress has been made in vision-language-action (VLA) models for general robot decision-making. However, the majority of existing VLAs fail to account for the inevitable external perturbations encountered during deployment.…

Cited by 2SourcePDFScholar
2024

Signal Transformer: Complex-Valued Attention and Meta-Learning for Signal Recognition

ICASSP 2024accepted

Deep neural networks have been shown as a class of useful tools for addressing signal recognition issues in recent years, especially for identifying the nonlinear feature structures of signals. However, this power of most deep learning techniques heavily relies on an abundant amount of training data…

Cited by 0SourceScholar