← Search

Seyed Mohammad Asghari

6 accepted papers

2023

Approximate Thompson Sampling via Epistemic Neural Networks

UAI 2023poster

Thompson sampling (TS) is a popular heuristic for action selection, but it requires sampling from a posterior distribution. Unfortunately, this can become computationally intractable in complex environments, such as those modeled using neural networks. Approximate posterior samples can produce effec…

2023

Epistemic Neural Networks

NeurIPS 2023spotlight

Intelligence relies on an agent's knowledge of what it does not know. This capability can be assessed based on the quality of joint predictions of labels across multiple inputs. In principle, ensemble-based approaches can produce effective joint predictions, but the computational costs of large ense…

Cited by 144SourcePDFScholar
2022

Evaluating high-order predictive distributions in deep learning

UAI 2022poster

Most work on supervised learning research has focused on marginal predictions. In decision problems, joint predictive distributions are essential for good performance. Previous work has developed methods for assessing low-order predictive distributions with inputs sampled i.i.d. from the testing dis…

2022

The Neural Testbed: Evaluating Joint Predictions

NeurIPS 2022accept

Predictive distributions quantify uncertainties ignored by point estimates. This paper introduces The Neural Testbed: an open source benchmark for controlled and principled evaluation of agents that generate such predictions. Crucially, the testbed assesses agents not only on the quality of their ma…

2020

Regret Bounds for Decentralized Learning in Cooperative Multi-Agent Dynamical Systems

UAI 2020poster

Regret analysis is challenging in Multi-Agent Reinforcement Learning (MARL) primarily due to the dynamical environments and the decentralized information among agents. We attempt to solve this challenge in the context of decentralized learning in multi-agent linear-quadratic (LQ) dynamical systems.…

Cited by 12SourcePDFScholar