2018
Why so gloomy? A Bayesian explanation of human pessimism bias in the multi-armed bandit task
NeurIPS 2018poster
How humans make repeated choices among options with imperfectly known reward outcomes is an important problem in psychology and neuroscience. This is often studied using multi-armed bandits, which is also frequently studied in machine learning. We present data from a human stationary bandit experime…