← Search

Ramesh Johari

11 accepted papers

2026

Estimation of Treatment Effects Under Nonstationarity via the Truncated Policy Gradient Estimator

ICML 2026poster

Randomized experiments (or A/B tests) are widely used to evaluate interventions in dynamic systems such as recommendation platforms, marketplaces, and digital health. In these settings, interventions affect both current and future system states, so estimating the global average treatment effect (GAT…

Cited by 0SourceScholar
2024

Hybrid$^2$ Neural ODE Causal Modeling and an Application to Glycemic Response

ICML 2024oral

Hybrid models composing mechanistic ODE-based dynamics with flexible and expressive neural network components have grown rapidly in popularity, especially in scientific domains where such ODE-based modeling offers important interpretability and validated causal grounding (e.g., for counterfactual re…

2023

Online Learning for Traffic Routing under Unknown Preferences

AISTATS 2023poster

In transportation networks, road tolling schemes are a method to cope with the efficiency losses due to selfish user routing, wherein users choose routes to minimize individual travel costs. However, the efficacy of tolling schemes often relies on access to complete information on users’ trip attrib…

2020

Adaptive Experimental Design with Temporal Interference: A Maximum Likelihood Approach

NeurIPS 2020poster

Suppose an online platform wants to compare a treatment and control policy (e.g., two different matching algorithms in a ridesharing system, or two different inventory management algorithms in an online retail site). Standard experimental approaches to this problem are biased (due to temporal inter…

Cited by 46SourcePDFScholar
2020

Unreasonable Effectiveness of Greedy Algorithms in Multi-Armed Bandit with Many Arms

NeurIPS 2020spotlight

We study the structure of regret-minimizing policies in the {\em many-armed} Bayesian multi-armed bandit problem: in particular, with $k$ the number of arms and $T$ the time horizon, we consider the case where $k \geq \sqrt{T}$. We first show that {\em subsampling} is a critical step for designing o…