← Search

Jia Sun

6 accepted papers

2026

MoReL: A Generalizable Framework for Dexterous Hand Retargeting via Modular Residual Reinforcement Learning

RA-L 2026

Effective motion retargeting is essential for robotic hands to perform fine-grained teleoperated manipulation. However, existing methods face several key challenges: optimization-based approaches offer accurate reproduction but suffer from high computational latency; learning-based methods provide f

Cited by 0SourceScholar
2026

ScaleADFG: Affordance-Based Dexterous Functional Grasping via Scalable Dataset

RA-L 2026

Dexterous functional tool-use grasping is essential for effective robotic manipulation of tools. However, existing approaches face significant challenges in efficiently constructing large-scale datasets and ensuring generalizability to everyday object scales. These issues primarily arise from size m

Cited by 1SourcecodeScholar
2025

DataSIR: A Benchmark Dataset for Sensitive Information Recognition

NeurIPS 2025poster

With the rapid development of artificial intelligence technologies, the demand for training data has surged, exacerbating risks of data leakage. Despite increasing incidents and costs associated with such leaks, data leakage prevention (DLP) technologies lag behind evolving evasion techniques that b…

Cited by 0SourcecodeScholar
2025

Refer and Grasp: Vision-Language Guided Continuous Dexterous Grasping

IROS 2025

Robotic grasping guided by natural language instructions faces challenges due to ambiguities in object descriptions and the need to interpret complex spatial context. Existing visual grounding methods often rely on datasets that fail to capture these complexities, particularly when object categories

Cited by 0SourcecodeScholar
2024

EVE: Efficient Vision-Language Pre-training with Masked Prediction and Modality-Aware MoE

AAAI 2024technical

Building scalable vision-language models to learn from diverse, multimodal data remains an open challenge. In this paper, we introduce an Efficient Vision-languagE foundation model, namely EVE, which is one unified multimodal Transformer pre-trained solely by one unified pre-training task. Specifica…

Cited by 11SourcePDFScholar
2019

Led3D: A Lightweight and Efficient Deep Approach to Recognizing Low-Quality 3D Faces

CVPR 2019poster

Due to the intrinsic invariance to pose and illumination changes, 3D Face Recognition (FR) has a promising potential in the real world. 3D FR using high-quality faces, which are of high resolutions and with smooth surfaces, have been widely studied. However, research on that with low-quality input i…

Cited by 70PDFScholar