← Search

Mark D. Reid

2 accepted papers

2016

Causal Bandits: Learning Good Interventions via Causal Inference

NeurIPS 2016poster

We study the problem of using causal models to improve the rate at which good interventions can be learned online in a stochastic environment. Our formalism combines multi-arm bandits and causal inference to model a novel type of bandit feedback that is not exploited by existing approaches. We propo…