← Search

Guangqi Jiang

6 accepted papers

2026

Cross-Hand Latent Representation for Vision-Language-Action Models

CVPR 2026

Dexterous manipulation is essential for real-world robot autonomy, mirroring the central role of human hand coordination in daily activity. Humans rely on rich multimodal perception--vision, sound, and language-guided intent--to perform dexterous actions, motivating vision-based, language-conditione

Cited by 0SourceScholar
2026

GSWorld: Closed-Loop Photo-Realistic Simulation Suite for Robotic Manipulation

ICRA 2026poster

This paper presents GSWorld, a robust, photo-realistic simulator for robotics manipulation that combines 3D Gaussian Splatting with physics engines. Our framework advocates ‘closing the loop’ of developing manipulation policies with reproducible evaluation of policies learned from real-robot data an…

2025

RoboDuet: Learning a Cooperative Policy for Whole-Body Legged Loco-Manipulation

RA-L 2025

Fully leveraging the loco-manipulation capabilities of a quadruped robot equipped with a robotic arm is non-trivial, as it requires controlling all degrees of freedom (DoFs) of the quadruped robot to achieve effective whole-body coordination. In this letter, we propose a novel framework RoboDuet, wh

Cited by 13SourceScholar
2025

Robots Pre-train Robots: Manipulation-Centric Robotic Representation from Large-Scale Robot Datasets

ICLR 2025poster

The pre-training of visual representations has enhanced the efficiency of robot learning. Due to the lack of large-scale in-domain robotic datasets, prior works utilize in-the-wild human videos to pre-train robotic visual representation. Despite their promising results, representations from human vi…

2024

Diffusion Reward: Learning Rewards via Conditional Video Diffusion

ECCV 2024poster

"Learning rewards from expert videos offers an affordable and effective solution to specify the intended behaviors for reinforcement learning (RL) tasks. In this work, we propose , a novel framework that learns rewards from expert videos via conditional video diffusion models for solving complex vis…

2024

Make-An-Agent: A Generalizable Policy Network Generator with Behavior-Prompted Diffusion

NeurIPS 2024poster

Can we generate a control policy for an agent using just one demonstration of desired behaviors as a prompt, as effortlessly as creating an image from a textual description? In this paper, we present **Make-An-Agent**, a novel policy parameter generator that leverages the power of conditional diffus…