This paper introduces the first framework for evaluating LLM decision making when agents are provided with graded similarity signals, and develops an LLM-behavioral-game-theoretic model that captures some of their reasoning rationale, and can support cooperative outcomes in equilibrium under sufficiently high similarity scores.
Abstract
As LLM-based agents with user-instructed goals are becoming widely deployed, they increasingly encounter each other in strategic interactions, and face challenges of finding mutually beneficial outcomes. Prior literature has argued that cooperation problems such as the Prisoner's Dilemma are resolvable in settings where agents know they follow very similar decision making patterns, as for example in monocultural AI ecosystems. Following that line of work, this paper introduces the first framework for evaluating LLM decision making when agents are provided with graded similarity signals. Among our findings, we establish that different LLM models vary drastically in how they navigate similarity signals, with some modern models showing consistent behavior across cooperation problems, payoff structures, and prompt framing. Perhaps surprisingly, our experiments also show that the dataset based on which the similarity signal is computed has small to no impact on induced cooperation, and that LLM models systematically self-identify as highly similar when asked to evaluate another model's chain-of-thought reasoning by themselves. Finally, we develop an LLM-behavioral-game-theoretic model that captures some of their reasoning rationale, and show that it can support cooperative outcomes in equilibrium under sufficiently high similarity scores.
Can cooperation among large language model (LLM) agents be evolutionarily stable against free-rider invasion? We study an indirect reciprocity donation game where LLM agents observe behavioral traces and donate on a continuous scale. Strategies, represented as natural language prompts, evolve through cultural transmiss...
Costly cooperation and the evolution of costly signals are both difficult to reconcile with simple fitness maximization, yet both are common in biological and social systems. We study a model in which agents emit costly signals and condition their actions on the signals they observe. The signals are arbitrary, non-cost...
Mahdiah Abolhasani, Saman Moghimi-Araghi, Mohammad Salahshour· Frontiers in Ethology· 0 citations
People often cooperate out of concern for their reputation. The corresponding theory of indirect reciprocity predicts that such cooperation can only evolve if people's opinions of each other are sufficiently correlated. This correlation, however, can be difficult to achieve when individuals form their opinions independ...
Yukari Jessica Tham, C. Hilbe, Yohsuke Murase· Proceedings of the National...· 0 citations
Mechanistic analyses show that a moderate neighborhood size enables individuals to strike an optimal balance between information sufficiency and decision-making tractability, which allows them to detect reciprocal opportunities while avoiding the deterioration of decision quality due to information overload.
Yi-Hsin Ku, Xin Ou, Ji-Qiang Zhang et al.· 0 citations
Large language models are increasingly used as interacting agents, but it remains unclear how robust their coordination is when public communication is unreliable. We study this question in iterated $N$-player Stag Hunt games played by homogeneous LLM groups under controlled programmatic action inversion, which changes...
Xuan-Yi Liu, Niall J. Dalton, Hairil Amin et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.