2026
What to Ask Next? Probing the Imaginative Reasoning of LLMs with TurtleSoup Puzzles
AAAI 2026technical
We investigate the capacity of Large Language Models (LLMs) for imaginative reasoning—the proactive construction, testing, and revision of hypotheses in information-sparse environments. Existing benchmarks, often static or focused on social deduction, fail to capture the dynamic, exploratory nature