A mean-field theory of the two-rate dynamics yields an analytical ordering condition that generalizes the consensus threshold of the stochastic Naming Game to a critical line in the $(pi,\phi)$ plane, and emerges as an architecture-dependent control parameter for decentralized LLM populations, quantitatively characterized by the statistical-physics toolkit.
Abstract
Decentralized populations of Large Language Model (LLM) agents can spontaneously reach consensus on shared conventions, yet the microscopic mechanisms by which their internal stochasticity shapes macroscopic ordering remain unexplored. We study a minimal LLM Naming Game in which the listener's decision is a single-token LLM call at decoding temperature $T$, replacing the inventory check of the deterministic Naming Game. Each interaction decomposes into an in-inventory and an out-inventory channel with conditional rates $\pi(T)\!\equiv\!P(\text{YES}\mid w\in P_j)$ and $\phi(T)\!\equiv\!P(\text{YES}\mid w\notin P_j)$, whose balance controls an ordering-disordering drift. A mean-field theory of the two-rate dynamics yields an analytical ordering condition that generalizes the consensus threshold of the stochastic Naming Game to a critical line in the $(\pi,\phi)$ plane. Across three open-weight architectures, consensus is always reached, but through three distinct listener regimes: permissive (repaint-noise dominated), near-deterministic, and conservative (missed-collapse dominated). The effective finite-size exponent $\beta(T)$ in $t_{\rm conv}\!\sim\!N^{\beta}$ shifts with temperature, and the temperature-sensitivity $\alpha$ in $t_c\!\sim\!e^{\alpha T}$ ranges from ${\approx}\,0.67$ to ${\approx}\,0$ across architectures. Decoding temperature thus emerges as an architecture-dependent control parameter for decentralized LLM populations, quantitatively characterized by the statistical-physics toolkit.
A novel framework based on Koopman operator theory is developed and validates its theoretical guarantees on multi-agent consensus dynamics, making spectral certification a practical layer for trustworthy collective reasoning.
Nontrivial dynamics can emerge in large language model (LLM)-based multi-agent systems, and preliminary evidence exists that formalisms from statistical mechanics can be effective at modeling and predicting such behaviors. In parallel, designing multi-agent communication topology for optimal task-solving is an active r...
Wen-Wen Zheng, Yuzhe Yang, Helen Qu et al.· 0 citations
Large Language Model (LLM) agents are increasingly deployed as populations of interacting entities, in which consensus --agreement on a shared answer-- emerges as a collective, unengineered behaviour. Prior work on LLM consensus shows that agents can cross-verify their answers and converge towards more factual response...
Emanuele Ricco, Elia Onofri, Vincenzo Sammartino et al.· 0 citations
Multi-agent LLM systems increasingly mix models from several providers, yet exposing each agent's underlying model identity to its peers significantly impairs cooperation. We show that when agents are aware of each other's model family, the group splits into clusters, where agents prefer interacting with others carryin...
Xavier Del Giudice, Alessio Palma, Matteo Migliarini et al.· 0 citations
Large language models are increasingly used as interacting agents, but it remains unclear how robust their coordination is when public communication is unreliable. We study this question in iterated $N$-player Stag Hunt games played by homogeneous LLM groups under controlled programmatic action inversion, which changes...
Xuan-Yi Liu, Niall J. Dalton, Hairil Amin et al.· 0 citations
We study an asynchronous consensus dynamics on $N$ agents: at each step a uniformly chosen agent replaces its state by $f(Y_1,\dots,Y_r)$, where $f$ is a fixed monotone aggregation rule and $Y_1,\dots,Y_r$ are the states of $r$ agents sampled uniformly with replacement. Let $T$ be the first time at which all agents agr...
Elchanan Mossel· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.