← Search

Gaurav Datta

5 accepted papers

2025

ICRT: In-Context Imitation Learning via Next-Token Prediction

ICRA 2025

In-context imitation learning is the capability to perform novel tasks when prompted with task demonstration examples. In-Context Robot Transformer (ICRT) is a causal transformer that performs autoregressive prediction on sensorimotor trajectories, which include images, proprioceptive states, and ac

Cited by 53SourceScholar
2024

A Touch, Vision, and Language Dataset for Multimodal Alignment

ICML 2024oral

Touch is an important sensing modality for humans, but it has not yet been incorporated into a multimodal generative language model. This is partially due to the difficulty of obtaining natural language labels for tactile data and the complexity of aligning tactile readings with both visual observat…

2023

Efficient Preference-Based Reinforcement Learning Using Learned Dynamics Models

ICRA 2023poster

Preference-based reinforcement learning (PbRL) can enable robots to learn to perform tasks based on an individual's preferences without requiring a hand-crafted re-ward function. However, existing approaches either assume access to a high-fidelity simulator or analytic model or take a model-free app…

Cited by 24SourceScholar
2023

IIFL: Implicit Interactive Fleet Learning from Heterogeneous Human Supervisors

CoRL 2023poster

Imitation learning has been applied to a range of robotic tasks, but can struggle when robots encounter edge cases that are not represented in the training data (i.e., distribution shift). Interactive fleet learning (IFL) mitigates distribution shift by allowing robots to access remote human supervi…

Cited by 5SourcecodeScholar
2023

Learning to Efficiently Plan Robust Frictional Multi-Object Grasps

IROS 2023poster

We consider a decluttering problem where multiple rigid convex polygonal objects rest in randomly placed positions and orientations on a planar surface and must be efficiently transported to a packing box using both single and multi-object grasps. Prior work considered frictionless multi-object gras…

Cited by 13SourceScholar