← Search

Jaegu Choy

1 accepted papers

2020

No-Regret Shannon Entropy Regularized Neural Contextual Bandit Online Learning for Robotic Grasping

IROS 2020poster

In this paper, we propose a novel contextual bandit algorithm that employs a neural network as a reward estimator and utilizes Shannon entropy regularization to encourage exploration, which is called Shannon entropy regularized neural contextual bandits (SERN). In many learning-based algorithms for…

Cited by 2SourceScholar