← Search

Ana L. C. Bazzan

3 accepted papers

2025

Constructing an Optimal Behavior Basis for the Option Keyboard

NeurIPS 2025poster

Multi-task reinforcement learning aims to quickly identify solutions for new tasks with minimal or no additional interaction with the environment. Generalized Policy Improvement (GPI) addresses this by combining a set of base policies to produce a new one that is at least as good—though not necessar…

Cited by 0SourceScholar
2023

A Toolkit for Reliable Benchmarking and Research in Multi-Objective Reinforcement Learning

NeurIPS 2023poster

Multi-objective reinforcement learning algorithms (MORL) extend standard reinforcement learning (RL) to scenarios where agents must optimize multiple---potentially conflicting---objectives, each represented by a distinct reward function. To facilitate and accelerate research and benchmarking in mult…

2023

Multi-Step Generalized Policy Improvement by Leveraging Approximate Models

NeurIPS 2023poster

We introduce a principled method for performing zero-shot transfer in reinforcement learning (RL) by exploiting approximate models of the environment. Zero-shot transfer in RL has been investigated by leveraging methods rooted in generalized policy improvement (GPI) and successor features (SFs). Alt…

Cited by 6SourcePDFScholar