2023
Action Inference by Maximising Evidence: Zero-Shot Imitation from Observation with World Models
NeurIPS 2023poster
Unlike most reinforcement learning agents which require an unrealistic amount of environment interactions to learn a new behaviour, humans excel at learning quickly by merely observing and imitating others. This ability highly depends on the fact that humans have a model of their own embodiment that…