← Search

Aoi Naito

1 accepted papers

2026

Systematic Failures in Collective Reasoning under Distributed Information in Multi-Agent LLMs

ICML 2026poster

Multi-agent systems built on large language models (LLMs) are expected to enhance decision-making by pooling distributed information, yet systematically evaluating this capability has remained challenging. We introduce HiddenBench, a 65-task benchmark grounded in the Hidden Profile paradigm, which i…

Cited by 0SourceScholar