← Search

Yizheng Zhang

7 accepted papers

2026

Cooperative-Competitive Team Play of Real-World Craft Robots

ICRA 2026poster

Multi-agent deep Reinforcement Learning (RL) has made significant progress in developing intelligent game-playing agents in recent years. However, the efficient training of collective robots using multi-agent RL and the transfer of learned policies to real-world applications remain open research que…

2025

AgentWorld: An Interactive Simulation Platform for Scene Construction and Mobile Robotic Manipulation

CoRL 2025poster

We introduce AgentWorld, an interactive simulation platform for developing household mobile manipulation capabilities. Our platform combines automated scene construction that encompasses layout generation, semantic asset placement, visual material configuration, and physics simulation, with a dual-m…

Cited by 0SourceScholar
2024

HumanVLA: Towards Vision-Language Directed Object Rearrangement by Physical Humanoid

NeurIPS 2024poster

Physical Human-Scene Interaction (HSI) plays a crucial role in numerous applications. However, existing HSI techniques are limited to specific object dynamics and privileged information, which prevents the development of more comprehensive applications. To address this limitation, we introd…

2024

Learning Highly Dynamic Behaviors for Quadrupedal Robots

ICRA 2024poster

Learning highly dynamic behaviors for robots has been a longstanding challenge. Traditional approaches have demonstrated robust locomotion, but the exhibited behaviors lack diversity and agility. They employ approximate models, which lead to compromises in performance. Data-driven approaches have be…

Cited by 5SourceScholar
2024

Relative Policy-Transition Optimization for Fast Policy Transfer

AAAI 2024technical

We consider the problem of policy transfer between two Markov Decision Processes (MDPs). We introduce a lemma based on existing theoretical results in reinforcement learning to measure the relativity gap between two arbitrary MDPs, that is the difference between any two cumulative expected returns d…

Cited by 0SourcePDFScholar
2023

Learning Terrain-Adaptive Locomotion with Agile Behaviors by Imitating Animals

IROS 2023poster

In this paper, we present a general learning framework for controlling a quadruped robot that can mimic the behavior of real animals and traverse challenging terrains. Our method consists of two steps: an imitation learning step to learn from motions of real animals, and a terrain adaptation step to…

Cited by 13SourceScholar
2022

RECCraft System: Towards Reliable and Efficient Collective Robotic Construction

IROS 2022poster

This research presents a novel Collective Robotic Construction (CRC) system named RECCraft. The RECCraft hardware system is composed of the mobile manipulation vehicles, the cubic blocks, and the folding ramp blocks. Solid connection and easy removal of the blocks are achieved by an electropermanent…

Cited by 4SourceScholar