← Search

Vidya K Muthukumar

3 accepted papers

2022

Harmless interpolation in regression and classification with structured features

AISTATS 2022poster

Overparametrized neural networks tend to perfectly fit noisy training data yet generalize well on test data. Inspired by this empirical observation, recent work has sought to understand this phenomenon of benign overfitting or harmless interpolation in the much simpler linear model. Previous theoret…

Cited by 16SourcePDFScholar
2022

Learning from an Exploring Demonstrator: Optimal Reward Estimation for Bandits

AISTATS 2022poster

We introduce the “inverse bandit” problem of estimating the rewards of a multi-armed bandit instance from observing the learning process of a low-regret demonstrator. Existing approaches to the related problem of inverse reinforcement learning assume the execution of an optimal policy, and thereby s…

2022

Universal and data-adaptive algorithms for model selection in linear contextual bandits

ICML 2022spotlight

Model selection in contextual bandits is an important complementary problem to regret minimization with respect to a fixed model class. We consider the simplest non-trivial instance of model-selection: distinguishing a simple multi-armed bandit problem from a linear contextual bandit problem. Even i…

Cited by 7SourcePDFScholar