← Search

Guillermo A. Pérez

2 accepted papers

2025

Revelations: A Decidable Class of POMDPs with Omega-Regular Objectives

AAAI 2025technical

Partially observable Markov decision processes (POMDPs) form a prominent model for uncertainty in sequential decision making. We are interested in constructing algorithms with theoretical guarantees to determine whether the agent has a strategy ensuring a given specification with probability 1. This…

2022

Distillation of RL Policies with Formal Guarantees via Variational Abstraction of Markov Decision Processes

AAAI 2022technical

We consider the challenge of policy simplification and verification in the context of policies learned through reinforcement learning (RL) in continuous environments. In well-behaved settings, RL algorithms have convergence guarantees in the limit. While these guarantees are valuable, they are insuf…

Cited by 13SourcePDFScholar