← Search

Krishna Acharya

2 accepted papers

2024

One Shot Inverse Reinforcement Learning for Stochastic Linear Bandits

UAI 2024poster

The paradigm of inverse reinforcement learning (IRL) is used to specify the reward function of an agent purely from its actions and is critical for value alignment and AI safety. While IRL is successful in practice, theoretical guarantees remain nascent. Motivated by the need for IRL in large action…

Cited by 1SourcePDFScholar
2024

Oracle Efficient Algorithms for Groupwise Regret

ICLR 2024poster

We study the problem of online prediction, in which at each time step $t \in \{1,2, \cdots T\}$, an individual $x_t$ arrives, whose label we must predict. Each individual is associated with various groups, defined based on their features such as age, sex, race etc., which may intersect. Our goal is…

Cited by 5SourcePDFScholar