2026
OpenDeception: Learning Deception and Trust in Human–AI Interaction via Multi-Agent Simulation
ICML 2026poster
As large language models (LLMs) are increasingly deployed as interactive agents, open-ended human-AI interactions can involve deceptive behaviors with serious real-world consequences, yet existing evaluations remain largely scenario-specific and model-centric. We introduce *OpenDeception*, a lightwe…