2026
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making
ICML 2026poster
Offline reinforcement learning (RL) learns policies from fixed datasets, thereby avoiding costly or unsafe environment interactions. However, its reliance on finite static datasets inherently restricts the ability to generalize beyond the training distribution. Prior solutions based on synthetic dat…