← Search

Patrick Wienhöft

2 accepted papers

2025

Solving Robust Markov Decision Processes: Generic, Reliable, Efficient

AAAI 2025technical

Markov decision processes (MDP) are a well-established model for sequential decision-making in the presence of probabilities. In *robust* MDP (RMDP), every action is associated with an *uncertainty set* of probability distributions, modelling that transition probabilities are not known precisely. Ba…

Cited by 1SourcePDFScholar
2023

More for Less: Safe Policy Improvement with Stronger Performance Guarantees

IJCAI 2023poster

In an offline reinforcement learning setting, the safe policy improvement (SPI) problem aims to improve the performance of a behavior policy according to which sample data has been generated. State-of-the-art approaches to SPI require a high number of samples to provide practical probabilistic guar…