← Search

Lu Guo

2 accepted papers

2026

RAD: Retrieval High-quality Demonstrations to Enhance Decision-making

ICML 2026poster

Offline reinforcement learning (RL) learns policies from fixed datasets, thereby avoiding costly or unsafe environment interactions. However, its reliance on finite static datasets inherently restricts the ability to generalize beyond the training distribution. Prior solutions based on synthetic dat…

Cited by 0SourceScholar