← Search

Yun-En Liu

1 accepted papers

2017

Trading off Rewards and Errors in Multi-Armed Bandits

AISTATS 2017poster

In multi-armed bandits, the most common objective is the maximization of the cumulative reward. Alternative settings include active exploration, where a learner tries to gain accurate estimates of the rewards of all arms. While these objectives are contrasting, in many scenarios it is desirable to t…

Cited by 35SourcePDFScholar