← Search

Runyu Ding

8 accepted papers

2025

Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning

IROS 2025

Teleoperation is a crucial tool for collecting human demonstrations, but controlling robots with bimanual dexterous hands remains a challenge. Existing teleoperation systems struggle to handle the complexity of coordinating two hands for intricate manipulations. We introduce Bunny-VisionPro, a real-

Cited by 129SourcecodeScholar
2024

ACE: A Cross-platform and visual-Exoskeletons System for Low-Cost Dexterous Teleoperation

CoRL 2024poster

Bimanual robotic manipulation with dexterous hands has a large potential workability and a wide workspace as it follows the most natural human workflow. Learning from human demonstrations has proven highly effective for learning a dexterous manipulation policy. To collect such data, teleoperation se…

Cited by 37SourceScholar
2024

RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding

CVPR 2024poster

We propose a lightweight and scalable Regional Point-Language Contrastive learning framework namely RegionPLC for open-world 3D scene understanding aiming to identify and recognize open-set objects and categories. Specifically based on our empirical studies we introduce a 3D-aware SFusion strategy t…

2024

V-IRL: Grounding Virtual Intelligence in Real Life

ECCV 2024poster

"There is a sensory gulf between the Earth that humans inhabit and the digital realms in which modern AI agents are created. To develop AI agents that can sense, think, and act as flexibly as humans in real-world settings, it is imperative to bridge the realism gap between the digital and physical w…

2023

PLA: Language-Driven Open-Vocabulary 3D Scene Understanding

CVPR 2023poster

Open-vocabulary scene understanding aims to localize and recognize unseen categories beyond the annotated label space. The recent breakthrough of 2D open-vocabulary perception is largely driven by Internet-scale paired image-text data with rich vocabulary concepts. However, this success cannot be di…

2022

DODA: Data-Oriented Sim-to-Real Domain Adaptation for 3D Semantic Segmentation

ECCV 2022poster

"Deep learning approaches achieve prominent success in 3D semantic segmentation. However, collecting densely annotated real-world 3D datasets is extremely time-consuming and expensive. Training models on synthetic data and generalizing on real-world scenarios becomes an appealing alternative, but un…

2022

Towards Efficient 3D Object Detection with Knowledge Distillation

NeurIPS 2022accept

Despite substantial progress in 3D object detection, advanced 3D detectors often suffer from heavy computation overheads. To this end, we explore the potential of knowledge distillation (KD) for developing efficient 3D object detectors, focusing on popular pillar- and voxel-based detectors. In the a…

2021

PAConv: Position Adaptive Convolution With Dynamic Kernel Assembling on Point Clouds

CVPR 2021poster

We introduce Position Adaptive Convolution (PAConv), a generic convolution operation for 3D point cloud processing. The key of PAConv is to construct the convolution kernel by dynamically assembling basic weight matrices stored in Weight Bank, where the coefficients of these weight matrices are self…

Cited by 555PDFcodeScholar