← Search

Muyi Sun

11 accepted papers

2026

A3fford-HOI: Anatomy-Aligned Affordance Disentanglement for Fine-grained and Generalizable Hand Object Interaction

IJCAI 2026

Hand Object Interaction (HOI) generation provides an efficient solution for virtual reality simulation and embodied AI deployment.Recent studies have explored instruction-driven HOI synthesis, yet they overlooked the fine-grained interactive contact and struggled with robust generalization in data-s

Cited by 0Scholar
2026

SEMITOOTH: A GENERALIZABLE SEMI-SUPERVISED FRAMEWORK FOR MULTI-SOURCE TOOTH SEGMENTATION

ICASSP 2026poster

With the rapid advancement of artificial intelligence, intelligent dentistry for clinical diagnosis and treatment has become increasingly promising. As the primary clinical dentistry task, tooth structure segmentation for Cone-Beam Computed Tomography (CBCT) has made significant progress in recent y…

Cited by 0SourcePDFScholar
2026

VividListener: Expressive and Controllable Listener Dynamics Modeling for Multi-Modal Responsive Interaction

AAAI 2026technical

Generating responsive listener head dynamics with nuanced emotions and expressive reactions is crucial for dialogue modeling in various virtual avatar animations. Previous studies mainly focus on the direct short-term production of listener behavior. They overlook the fine-grained control over motio

Cited by 0SourcePDFScholar
2025

DanceEditor: Towards Iterative Editable Music-driven Dance Generation with Open-Vocabulary Descriptions

ICCV 2025poster

Generating coherent and diverse human dances from music signals has gained tremendous progress in animating virtual avatars. While existing methods support direct dance synthesis, they fail to recognize that enabling users to edit dance movements is far more practical in real-world choreography scen…

2024

Contrmix: Progressive Mixed Contrastive Learning for Semi-Supervised Medical Image Segmentation

ICASSP 2024accepted

While medical image segmentation has achieved impressive progress, it usually being constrained by labor-intensive and costly pixel-wise annotations. The existing semi-supervised learning methods ignore the inherent imbalance and high similarity of different categories in medical images. To address…

Cited by 0SourceScholar
2024

MoPE-CLIP: Structured Pruning for Efficient Vision-Language Models with Module-wise Pruning Error Metric

CVPR 2024poster

Vision-language pre-trained models have achieved impressive performance on various downstream tasks. However their large model sizes hinder their utilization on platforms with limited computational resources. We find that directly using smaller pre-trained models and applying magnitude-based pruning…

Cited by 21SourcePDFScholar
2024

PTM-VQA: Efficient Video Quality Assessment Leveraging Diverse PreTrained Models from the Wild

CVPR 2024poster

Video quality assessment (VQA) is a challenging problem due to the numerous factors that can affect the perceptual quality of a video e.g. content attractiveness distortion type motion pattern and level. However annotating the Mean opinion score (MOS) for videos is expensive and time-consuming which…

Cited by 5SourcePDFScholar
2023

Diverse 3D Hand Gesture Prediction From Body Dynamics by Bilateral Hand Disentanglement

CVPR 2023poster

Predicting natural and diverse 3D hand gestures from the upper body dynamics is a practical yet challenging task in virtual avatar creation. Previous works usually overlook the asymmetric motions between two hands and generate two hands in a holistic manner, leading to unnatural results. In this wor…

2023

Lightvessel: Exploring Lightweight Coronary Artery Vessel Segmentation Via Similarity Knowledge Distillation

ICASSP 2023accepted

In recent years, deep convolution neural networks (DCNNs) have achieved great prospects in coronary artery vessel segmentation. However, it is difficult to deploy complicated models in clinical scenarios since high-performance approaches have excessive parameters and high computation costs. To tackl…

Cited by 0SourceScholar