← Search

Shaoting Zhu

6 accepted papers

2025

GS-Occ3D: Scaling Vision-only Occupancy Reconstruction with Gaussian Splatting

ICCV 2025poster

Occupancy is crucial for autonomous driving, providing essential geometric priors for perception and planning. However, existing methods predominantly rely on LiDAR-based occupancy annotations, which limits scalability and prevents leveraging vast amounts of potential crowdsourced data for auto-labe…

2025

RoboEngine: Plug-and-Play Robot Data Augmentation with Semantic Robot Segmentation and Background Generation

IROS 2025

Visual augmentation has become a crucial technique for enhancing the visual robustness of imitation learning. However, existing methods are often limited by prerequisites such as camera calibration or the need for controlled environments (e.g., green screen setups). In this work, we introduce RoboEn

Cited by 36SourcecodeScholar
2025

SARO: Space-Aware Robot System for Terrain Crossing via Vision-Language Model

ICRA 2025

The application of vision-language models (VLMs) has achieved impressive success in various robotics tasks. However, there are few explorations for foundation models used in quadruped robot navigation through terrains in 3D environments. We introduce SARO (Space-Aware Robot System for Terrain Crossi

Cited by 5SourcecodeScholar
2025

VR-Robo: A Real-to-Sim-to-Real Framework for Visual Robot Navigation and Locomotion

RA-L 2025

Recent success in legged robot locomotion is attributed to the integration of reinforcement learning and physical simulators. However, these policies often encounter challenges when deployed in real-world environments due to sim-to-real gaps, as simulators typically fail to replicate visual realism

Cited by 26SourceScholar