2026
Provably Efficient Policy-Reward Co-Pretraining for Adversarial Imitation Learning
ICML 2026poster
Adversarial imitation learning (AIL) demonstrates superior expert sample efficiency compared to behavioral cloning (BC), yet requires substantial online environment interaction. While recent empirical work has explored initializing AIL algorithms with BC-pretrained policies to address this limitatio…