← Search

Pratik Gajane

3 accepted papers

2023

Autonomous Exploration for Navigating in MDPs Using Blackbox RL Algorithms

IJCAI 2023poster

We consider the problem of navigating in a Markov decision process where extrinsic rewards are either absent or ignored. In this setting, the objective is to learn policies to reach all the states that are reachable within a given number of steps (in expectation) from a starting state. We introduce…

Cited by 0SourcePDFScholar
2015

A Relative Exponential Weighing Algorithm for Adversarial Utility-based Dueling Bandits

ICML 2015poster

We study the K-armed dueling bandit problem which is a variation of the classical Multi-Armed Bandit (MAB) problem in which the learner receives only relative feedback about the selected pairs of arms. We propose a new algorithm called Relative Exponential-weight algorithm for Exploration and Exploi…

Cited by 55SourcePDFScholar