2020
A Farewell to Arms: Sequential Reward Maximization on a Budget with a Giving Up Option
AISTATS 2020poster
We consider a sequential decision-making problem where an agent can take one action at a time and each action has a stochastic temporal extent, i.e., a new action cannot be taken until the previous one is finished. Upon completion, the chosen action yields a stochastic reward. The agent seeks to max…