← Search

Hirokazu Shirado

4 accepted papers

2026

Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs

ICML 2026poster

Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet systematically evaluating this capability has remained challenging. We introduce HiddenBench, a 65-task benchmark grounded in the Hidden Profile paradigm, which i…

Cited by 0SourceScholar
2025

Martingale Score: An Unsupervised Metric for Bayesian Rationality in LLM Reasoning

NeurIPS 2025poster

Recent advances in reasoning techniques have substantially improved the performance of large language models (LLMs), raising expectations for their ability to provide accurate, truthful, and reliable information. However, emerging evidence suggests that iterative reasoning may foster belief entrench…

Cited by 0SourceScholar
2023

Rethinking Safe Control in the Presence of Self-Seeking Humans

AAAI 2023technical

Safe control methods are often designed to behave safely even in worst-case human uncertainties. Such design can cause more aggressive human behaviors that exploit its conservatism and result in greater risk for everyone. However, this issue has not been systematically investigated previously. This…

Cited by 4SourcePDFScholar