2024
Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning
RSS 2024poster
A common failure mode for policies trained with imitation is compounding execution errors at test time. When the learned policy encounters states that are not present in the expert demonstrations, the policy fails, leading to degenerate behavior. The Dataset Aggregation, or DAgger approach to this p…