← Search

Pranav Kumar

1 accepted papers

2024

Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning

RSS 2024poster

A common failure mode for policies trained with imitation is compounding execution errors at test time. When the learned policy encounters states that are not present in the expert demonstrations, the policy fails, leading to degenerate behavior. The Dataset Aggregation, or DAgger approach to this p…

Cited by 15SourcePDFScholar