2026
Consistent Zero-Shot Imitation with Contrastive Goal Inference
ICML 2026poster
In the same way that generative models today conduct most of their training in a self-supervised fashion, how can agentic models conduct their training in a self-supervised fashion, interactively exploring, learning, and preparing to quickly adapt to new tasks? The problem of reward-free exploration…