← Search

Olivier Buffet

8 accepted papers

2025

\(\varepsilon\)-Optimally Solving Two-Player Zero-Sum POSGs

NeurIPS 2025poster

We present a novel framework for \(\varepsilon\)-optimally solving two-player zero-sum partially observable stochastic games (zs-POSGs). These games pose a major challenge due to the absence of a principled connection with dynamic programming (DP) techniques developed for two-player zero-sum stochas…

Cited by 0SourceScholar
2024

Solving Hierarchical Information-Sharing Dec-POMDPs: An Extensive-Form Game Approach

ICML 2024poster

A recent theory shows that a multi-player decentralized partially observable Markov decision process can be transformed into an equivalent single-player game, enabling the application of Bellman's principle of optimality to solve the single-player game by breaking it down into single-stage subgames.…

Cited by 4SourcePDFScholar
2023

Robust Robot Planning for Human-Robot Collaboration

ICRA 2023poster

In human-robot collaboration, the objectives of the human are often unknown to the robot. Moreover, even assuming a known objective, the human behavior is also uncertain. In order to plan a robust robot behavior, a key preliminary question is then: How to derive realistic human behaviors given a kno…

Cited by 7SourceScholar
2020

Optimally Solving Two-Agent Decentralized POMDPs Under One-Sided Information Sharing

ICML 2020poster

Optimally solving decentralized partially observable Markov decision processes under either full or no information sharing received significant attention in recent years. However, little is known about how partial information sharing affects existing theory and algorithms. This paper addresses this…

Cited by 22SourcePDFScholar
2018

rho-POMDPs have Lipschitz-Continuous epsilon-Optimal Value Functions

NeurIPS 2018poster

Many state-of-the-art algorithms for solving Partially Observable Markov Decision Processes (POMDPs) rely on turning the problem into a “fully observable” problem—a belief MDP—and exploiting the piece-wise linearity and convexity (PWLC) of the optimal value function in this new state space (the beli…

Cited by 25SourcePDFScholar