← Search

Shutong Jin

4 accepted papers

2026

Physically-Based Lighting Generation for Robotic Manipulation

ICRA 2026poster

We propose the first framework that leverages physically-based inverse rendering for novel lighting generation on existing real-world human demonstrations of robotic manipulation tasks. Specifically, inverse rendering decomposes the first frame in each demonstration into geometric (surface normal, d…

2025

Feature Extractor or Decision Maker: Rethinking the Role of Visual Encoders in Visuomotor Policies

ICRA 2025

An end-to-end (E2E) visuomotor policy is typically treated as a unified whole, but recent approaches using out-of-domain (OOD) data to pretrain the visual encoder have cleanly separated the visual encoder from the network, with the remainder referred to as the policy. We propose Visual Alignment Tes

Cited by 1SourceScholar
2024

How Physics and Background Attributes Impact Video Transformers in Robotic Manipulation: A Case Study on Planar Pushing

IROS 2024poster

As model and dataset sizes continue to scale in robot learning, the need to understand how the composition and properties of a dataset affect model performance becomes increasingly urgent to ensure cost-effective data collection and model performance. In this work, we empirically investigate how phy…

Cited by 1SourceScholar
2022

SectionKey: 3-D Semantic Point Cloud Descriptor for Place Recognition

IROS 2022poster

Place recognition is seen as a crucial factor to correct cumulative errors in Simultaneous Localization and Mapping (SLAM) applications. Most existing studies focus on visual place recognition, which is inherently sensitive to environmental changes such as illumination, weather and seasons. Consider…

Cited by 19SourceScholar