2025
Dynamic Uncertainty Estimation for Offline Reinforcement Learning
AAAI 2025technical
Offline reinforcement learning confronts the distributional shift challenge, a consequence of learning policy from static datasets. Current methods primarily handle this issue by aligning the learned policy with the behavior policy or conservatively estimating Q-values for out-of-distribution (OOD)…