← Search

Jai Moondra

2 accepted papers

2025

Navigating the Social Welfare Frontier: Portfolios for Multi-objective Reinforcement Learning

ICML 2025poster

In many real-world applications of Reinforcement Learning (RL), deployed policies have varied impacts on different stakeholders, creating challenges in reaching consensus on how to effectively aggregate their preferences. Generalized $p$-means form a widely used class of social welfare functions for…

Cited by 0SourcePDFScholar
2021

Reusing Combinatorial Structure: Faster Iterative Projections over Submodular Base Polytopes

NeurIPS 2021poster

Optimization algorithms such as projected Newton's method, FISTA, mirror descent and its variants enjoy near-optimal regret bounds and convergence rates, but suffer from a computational bottleneck of computing ``projections" in potentially each iteration (e.g., $O(T^{1/2})$ regret of online mirror…