Skip to content
Preprint

Do LLMs Take Care of Their Own? Similarity Signals Can Induce Cooperation

Aug 2026 · 0 citations · 58 references
Computer Science

TL;DR

This paper introduces the first framework for evaluating LLM decision making when agents are provided with graded similarity signals, and develops an LLM-behavioral-game-theoretic model that captures some of their reasoning rationale, and can support cooperative outcomes in equilibrium under sufficiently high similarity scores.

Abstract

As LLM-based agents with user-instructed goals are becoming widely deployed, they increasingly encounter each other in strategic interactions, and face challenges of finding mutually beneficial outcomes. Prior literature has argued that cooperation problems such as the Prisoner's Dilemma are resolvable in settings where agents know they follow very similar decision making patterns, as for example in monocultural AI ecosystems. Following that line of work, this paper introduces the first framework for evaluating LLM decision making when agents are provided with graded similarity signals. Among our findings, we establish that different LLM models vary drastically in how they navigate similarity signals, with some modern models showing consistent behavior across cooperation problems, payoff structures, and prompt framing. Perhaps surprisingly, our experiments also show that the dataset based on which the similarity signal is computed has small to no impact on induced cooperation, and that LLM models systematically self-identify as highly similar when asked to evaluate another model's chain-of-thought reasoning by themselves. Finally, we develop an LLM-behavioral-game-theoretic model that captures some of their reasoning rationale, and show that it can support cooperative outcomes in equilibrium under sufficiently high similarity scores.

View source

Similar papers

Preprint Aug 2026

Emergence of Reputation-Based Cooperation in LLM Agents

Can cooperation among large language model (LLM) agents be evolutionarily stable against free-rider invasion? We study an indirect reciprocity donation game where LLM agents observe behavioral traces and donate on a continuous scale. Strategies, represented as natural language prompts, evolve through cultural transmiss...

Kazuya Horibe, Kenji Itao, Wataru Toyokawa · 0 citations
Open access Sep 2026

The co-evolution of costly markers and cooperation in social dilemmas

Costly cooperation and the evolution of costly signals are both difficult to reconcile with simple fitness maximization, yet both are common in biological and social systems. We study a model in which agents emit costly signals and condition their actions on the signals they observe. The signals are arbitrary, non-cost...

Mahdiah Abolhasani, Saman Moghimi-Araghi, Mohammad Salahshour · 0 citations
Open access Aug 2026

Indirect reciprocity with dual private assessment.

People often cooperate out of concern for their reputation. The corresponding theory of indirect reciprocity predicts that such cooperation can only evolve if people's opinions of each other are sufficiently correlated. This correlation, however, can be difficult to achieve when individuals form their opinions independ...

Yukari Jessica Tham, C. Hilbe, Yohsuke Murase · 0 citations
Preprint Aug 2026

Evolution of cooperation with Q-learning: how much information do we need?

Mechanistic analyses show that a moderate neighborhood size enables individuals to strike an optimal balance between information sufficiency and decision-making tractability, which allows them to detect reciprocal opportunities while avoiding the deterioration of decision quality due to information overload.

Yi-Hsin Ku, Xin Ou, Ji-Qiang Zhang et al. · 0 citations
#machine learning Preprint Sep 2026

Be Careful Who You Trust: Coordination Dynamics under Corrupted Communication in LLM Multi-Agent Games

Large language models are increasingly used as interacting agents, but it remains unclear how robust their coordination is when public communication is unreliable. We study this question in iterated $N$-player Stag Hunt games played by homogeneous LLM groups under controlled programmatic action inversion, which changes...

Xuan-Yi Liu, Niall J. Dalton, Hairil Amin et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.