2024
Augmenting Decision with Hypothesis in Reinforcement Learning
ICML 2024poster
Value-based reinforcement learning is the current State-Of-The-Art due to high sampling efficiency. However, our study shows it suffers from low exploitation in early training period and bias sensitiveness. To address these issues, we propose to augment the decision-making process with hypothesis, a…