2026
Scalable Multi-Objective and Meta Reinforcement Learning via Gradient Estimation
AAAI 2026technical
We study the problem of efficiently estimating policies that simultaneously optimize multiple objectives in reinforcement learning (RL). Given n objectives (or tasks), we seek the optimal partition of these objectives into k groups, which is much smaller than n, where each group comprises related ob