← Search

Matt MacDermott

3 accepted papers

2025

Can a Bayesian Oracle Prevent Harm from an Agent?

UAI 2025

Is there a way to design powerful AI systems based on machine learning methods that would satisfy probabilistic safety guarantees? With the long-term goal of obtaining a probabilistic guarantee that would apply in every context, we consider estimating a context-dependent bound on the probability of

2024

Discovering Agents (Abstract Reprint)

AAAI 2024technical

Causal models of agents have been used to analyse the safety aspects of machine learning systems. But identifying agents is non-trivial – often the causal model is just assumed by the modeller without much justification – and modelling failures can lead to mistakes in the safety analysis. This paper…

Cited by 0SourcePDFScholar