2026
Reliability-Adjusted Prioritized Experience Replay
ICLR 2026poster
Experience replay enables data-efficient learning from past experiences in online reinforcement learning agents. Traditionally, experiences were sampled uniformly from a replay buffer, regardless of differences in experience-specific learning potential. In an effort to sample more efficiently, resea…