Skip to content

Theory of Scene: Breaking the Symmetry Trap in Multi-Agent LLM Coordination

Sep 2026 · 0 citations · 44 references
Computer Science

TL;DR

Tory of Scene (ToS), a training-free reasoning schema in which each agent reads its public role, the only difference between the agents, and the task context they all observe, is proposed, which outperforms all six baselines on every benchmark, and each baseline falls far behind it in at least one setting.

Abstract

Multi-agent systems built on large language models (LLMs) are largely homogeneous, as their agents behave alike even across distinct LLMs. We show that when such agents act concurrently without communication, they collide on targets they must split and diverge on targets they must take together, a double failure we term the symmetry trap. Theory of Mind (ToM), widely used for coordination without communication, cannot escape this trap, since homogeneous agents form the same prediction of one another and respond to it in the same way. We propose Theory of Scene (ToS), a training-free reasoning schema in which each agent reads its public role, the only difference between the agents, and the task context they all observe. Homogeneous agents thereby derive one division of labor, each taking the part its role fixes, which turns homogeneity from the cause of the trap into the cure. ToS reads the role together with the scene through role gating, which determines whether ownership overlaps or is already divided, and the task context through task coupling, which infers whether the team must converge on each target, divide it, or take its stages in turn. We evaluate on DivvyBench, a controlled environment we introduce, whose target types make an episode Competitive, Cooperative, or Mixed across Tabletop, Airspace, and Household scenarios, and on two established agentic benchmarks, GovSim and Overcooked. ToS outperforms all six baselines on every benchmark, and each baseline falls far behind it in at least one setting. Against ToM given the same inputs, ToS raises the DivvyBench success rate from 71.1% to 99.6%, the GovSim total gain from 207 to 400, and the Overcooked level-normalized throughput from 1.41 to 1.67.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

Prompted Identity Degrades Cooperation in Multi-Agent LLM Systems

Multi-agent LLM systems increasingly mix models from several providers, yet exposing each agent's underlying model identity to its peers significantly impairs cooperation. We show that when agents are aware of each other's model family, the group splits into clusters, where agents prefer interacting with others carryin...

Xavier Del Giudice, Alessio Palma, Matteo Migliarini et al. · 0 citations
Preprint Sep 2026

Absorbing State Phase Transitions in Multi-Agent Search

Nontrivial dynamics can emerge in large language model (LLM)-based multi-agent systems, and preliminary evidence exists that formalisms from statistical mechanics can be effective at modeling and predicting such behaviors. In parallel, designing multi-agent communication topology for optimal task-solving is an active r...

Wen-Wen Zheng, Yuzhe Yang, Helen Qu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

Bilevel Coordinated Reflection: A Game-Theoretic Approach to Multi-Agent LLM Systems

Stochastic Reflective Memory Ascent (SRMA), which accepts a candidate memory only after a grounded evaluation risk strictly decreases, is introduced and provides confidence gating for stochastic evaluation and re-anchoring guarantees for piecewise-stationary environments.

Yi-Hang Chen, Yu-Xiang Chen, Yuxuan Huang et al. · 0 citations
Preprint Aug 2026

One Model, Many Minds: Unlocking Multi-Agent Synergy in a Single Agent via Mixture of Roles

The proposed Mixture of Roles (MoRe), which adaptively composes multiple specializations into a single steering vector for single-turn inference, enables multi-perspective specialization in a single-agent, single-turn inference process.

Zhichen Zeng, Hui-Yuan Chen, Jingru Cheng et al. · 2 citations
Preprint Aug 2026

The Interaction Tax: When Communication Erases Diversity in Multi-Agent Teams

Does multi-agent LLM interaction help or hurt? Some work reports gains from debate (Du et al., 2024), critique loops (Chen et al., 2025), and mixture-of-agents synthesis (Wang et al., 2025), while other work finds that interaction adds cost without improving quality under equal budgets (Tran&Kiela, 2026; Xu et al., 202...

Summer Eunhyung Ann, Haokun Liu, Chen-Hao Tan · 3 citations

Related blog posts

MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.