← Search

Dustin Morrill

3 accepted papers

2023

Composing Efficient, Robust Tests for Policy Selection

UAI 2023poster

Modern reinforcement learning systems produce many high-quality policies throughout the learning process. However, to choose which policy to actually deploy in the real world, they must be tested under an intractable number of environmental conditions. We introduce RPOSST, an algorithm to select a s…

Cited by 0SourcePDFScholar
2021

Efficient Deviation Types and Learning for Hindsight Rationality in Extensive-Form Games

ICML 2021spotlight

Hindsight rationality is an approach to playing general-sum games that prescribes no-regret learning dynamics for individual agents with respect to a set of deviations, and further describes jointly rational behavior among multiple agents with mediated equilibria. To develop hindsight rational learn…

2021

Hindsight and Sequential Rationality of Correlated Play

AAAI 2021technical

Driven by recent successes in two-player, zero-sum game solving and playing, artificial intelligence work on games has increasingly focused on algorithms that produce equilibrium-based strategies. However, this approach has been less effective at producing competent players in general-sum games or t…