2023
ContraBAR: Contrastive Bayes-Adaptive Deep RL
ICML 2023poster
In meta reinforcement learning (meta RL), an agent seeks a Bayes-optimal policy -- the optimal policy when facing an unknown task that is sampled from some known task distribution. Previous approaches tackled this problem by inferring a $\textit{belief}$ over task parameters, using variational infer…