← Search

P. N. Karthik

4 accepted papers

2026

Asymptotically Optimal Sequential Testing with Markovian Data

ICML 2026poster

We study one-sided and $\alpha$-correct sequential hypothesis testing for data generated by an ergodic Markov chain. The *null* hypothesis is that the unknown transition matrix belongs to a prescribed set $\cal P$ of stochastic matrices, and the *alternative* corresponds to a disjoint set $\cal Q$. …

Cited by 1SourceScholar
2025

Optimal Multi-Objective Best Arm Identification with Fixed Confidence

AISTATS 2025poster

We consider a multi-armed bandit setting with finitely many arms, in which each arm yields an $M$-dimensional vector reward upon selection. We assume that the reward of each dimension (a.k.a. {\em objective}) is generated independently of the others. The best arm of any given objective is the arm wi…

Cited by 0SourceScholar
2024

Fixed-Budget Differentially Private Best Arm Identification

ICLR 2024poster

We study best arm identification (BAI) in linear bandits in the fixed-budget regime under differential privacy constraints, when the arm rewards are supported on the unit interval. Given a finite budget $T$ and a privacy parameter $\varepsilon>0$, the goal is to minimise the error probability in f…

Cited by 0SourcePDFScholar
2023

Almost Cost-Free Communication in Federated Best Arm Identification

AAAI 2023technical

We study the problem of best arm identification in a federated learning multi-armed bandit setup with a central server and multiple clients. Each client is associated with a multi-armed bandit in which each arm yields i.i.d. rewards following a Gaussian distribution with an unknown mean and known va…

Cited by 9SourcePDFScholar