← Search

Chen Chu

6 accepted papers

2026

Adaptive Theory of Mind for LLM-based Multi-Agent Coordination

AAAI 2026technical

Theory of Mind (ToM) refers to the ability to reason about others’ mental states, and higher-order ToM involves considering that others also possess their own ToM. Equipping large language model (LLM)-driven agents with ToM has long been considered to improve their coordination in multiagent collabo

Cited by 0SourcePDFScholar
2024

A Successful Strategy for Multichannel Iterated Prisoner’s Dilemma

IJCAI 2024poster

Iterated prisoner’s dilemma (IPD) and its variants are fundamental models for understanding the evolution of cooperation in human society as well as AI systems. In this paper, we focus on multichannel IPD, and examine how an agent should behave to obtain generally high payoffs under this setting.…

Cited by 0SourcePDFScholar
2023

A Pair-Approximation Method for Modelling the Dynamics of Multi-Agent Stochastic Games

AAAI 2023technical

Developing a dynamical model for learning in games has attracted much recent interest. In stochastic games, agents need to make decisions in multiple states, and transitions between states, in turn, influence the dynamics of strategies. While previous works typically focus either on 2-agent stochast…

2022

A Formal Model for Multiagent Q-Learning Dynamics on Regular Graphs

IJCAI 2022poster

Modeling the dynamics of multi-agent learning has long been an important research topic. The focus of previous research has been either on 2-agent settings or well-mixed infinitely large agent populations. In this paper, we consider the scenario where n Q-learning agents locate on regular graphs, su…

Cited by 36SourcePDFScholar
2022

Modelling the Dynamics of Regret Minimization in Large Agent Populations: a Master Equation Approach

IJCAI 2022poster

Understanding the learning dynamics in multiagent systems is an important and challenging task. Past research on multi-agent learning mostly focuses on two-agent settings. In this paper, we consider the scenario in which a population of infinitely many agents apply regret minimization in repeated sy…

Cited by 144SourcePDFScholar