← Search

Takuma Yoneda

8 accepted papers

2025

STACKGEN: Generating Stable Structures from Silhouettes via Diffusion

IROS 2025

Humans naturally obtain intuition about the interactions between and the stability of rigid objects by observing and interacting with the world. It is this intuition that governs the way in which we regularly configure objects in our environment, allowing us to build complex structures from simple,

Cited by 2SourcecodeScholar
2024

Blending Imitation and Reinforcement Learning for Robust Policy Improvement

ICLR 2024spotlight

While reinforcement learning (RL) has shown promising performance, its sample complexity continues to be a substantial hurdle, restricting its broader application across a variety of domains. Imitation learning (IL) utilizes oracles to improve sample efficiency, yet it is often constrained by the qu…

Cited by 12SourcePDFScholar
2024

Statler: State-Maintaining Language Models for Embodied Reasoning

ICRA 2024poster

There has been a significant research interest in employing large language models to empower intelligent robots with complex reasoning. Existing work focuses on harnessing their abilities to reason about the histories of their actions and observations. In this paper, we explore a new dimension in wh…

Cited by 41SourcecodeScholar
2023

Active Policy Improvement from Multiple Black-box Oracles

ICML 2023poster

Reinforcement learning (RL) has made significant strides in various complex domains. However, identifying an effective policy via RL often necessitates extensive exploration. Imitation learning aims to mitigate this issue by using expert demonstrations to guide exploration. In real-world scenarios,…

2023

Cold Diffusion on the Replay Buffer: Learning to Plan from Known Good States

CoRL 2023poster

Learning from demonstrations (LfD) has successfully trained robots to exhibit remarkable generalization capabilities. However, many powerful imitation techniques do not prioritize the feasibility of the robot behaviors they generate. In this work, we explore the feasibility of plans produced by LfD.…

Cited by 6SourceScholar
2023

To the Noise and Back: Diffusion for Shared Autonomy

RSS 2023poster

Shared autonomy is an operational concept in which a user and an autonomous agent collaboratively control a robotic system. It provides a number of advantages over the extremes of full-teleoperation and full-autonomy in many settings. Traditional approaches to shared autonomy rely on knowledge of th…

2022

Benchmarking Structured Policies and Policy Optimization for Real-World Dexterous Object Manipulation

RA-L 2022

Dexterous manipulation is a challenging and important problem in robotics. While data-driven methods are a promising approach, current benchmarks require simulation or extensive engineering support due to the sample inefficiency of popular methods. We present benchmarks for the TriFinger system, an

Cited by 39SourcecodeScholar
2022

Invariance Through Latent Alignment

RSS 2022poster

A robot's deployment environment often involves perceptual changes that differ from what it has experienced during training. Standard practices such as data augmentation attempt to bridge this gap by augmenting source images in an effort to extend the support of the training distribution to better c…

Cited by 9SourcePDFScholar