2026
TeamTR: Trust-Region Fine-Tuning for Multi-Agent LLM Coordination
ICML 2026poster
Multi-agent LLM systems can improve reasoning and tool use, yet recent evidence shows their gains are often unstable and sensitive to interaction design. A promising direction is to \emph{train} collaboration, but team post-training introduces a moving-target effect: when agents interact through a s…