← Search

Nishant Mehta

5 accepted papers

2024

On the price of exact truthfulness in incentive-compatible online learning with bandit feedback: a regret lower bound for WSU-UX

AISTATS 2024poster

In one view of the classical game of prediction with expert advice with binary outcomes, in each round, each expert maintains an adversarially chosen belief and honestly reports this belief. We consider a recently introduced, strategic variant of this problem with selfish (reputation-seeking) expert…

Cited by 2SourcePDFScholar
2020

A Farewell to Arms: Sequential Reward Maximization on a Budget with a Giving Up Option

AISTATS 2020poster

We consider a sequential decision-making problem where an agent can take one action at a time and each action has a stochastic temporal extent, i.e., a new action cannot be taken until the previous one is finished. Upon completion, the chosen action yields a stochastic reward. The agent seeks to max…

Cited by 3SourcePDFScholar