2024
No Representation, No Trust: Connecting Representation, Collapse, and Trust Issues in PPO
NeurIPS 2024poster
Reinforcement learning (RL) is inherently rife with non-stationarity since the states and rewards the agent observes during training depend on its changing policy. Therefore, networks in deep RL must be capable of adapting to new observations and fitting new targets. However, previous works have obs…