← Search

Joao Candido Ramos

1 accepted papers

2024

Mimicking Better by Matching the Approximate Action Distribution

ICML 2024poster

In this paper, we introduce MAAD, a novel, sample-efficient on-policy algorithm for Imitation Learning from Observations. MAAD utilizes a surrogate reward signal, which can be derived from various sources such as adversarial games, trajectory matching objectives, or optimal transport criteria. To co…