2025
PWM: Policy Learning with Multi-Task World Models
ICLR 2025poster
Reinforcement Learning (RL) has made significant strides in complex tasks but struggles in multi-task settings with different embodiments. World model methods offer scalability by learning a simulation of the environment but often rely on inefficient gradient-free optimization methods for policy ext…