← Search

Sanjay P. Bhat

5 accepted papers

2022

Identifying near-optimal decisions in linear-in-parameter bandit models with continuous decision sets

UAI 2022poster

We consider an online optimization problem in a bandit setting in which a learner chooses decisions from a continuous decision set at discrete decision epochs, and receives noisy rewards from the environment in response. While the noise samples are assumed to be independent and sub-Gaussian, the…

Cited by 1SourcePDFScholar
2021

Computing an Efficient Exploration Basis for Learning with Univariate Polynomial Features

AAAI 2021technical

Barycentric spanners have been used as an efficient exploration basis in online linear optimization problems in a bandit framework. We characterise the barycentric spanner for decision problems in which the cost (or reward) is a polynomial in a single decision variable. Our characterisation of the b…

Cited by 6SourcePDFScholar