Skip to content
Open access

What Do You Mean? Exploring the Alleged Theory of Mind Abilities of Large Language Models

2026 · Transactions of the Association for Computational Linguistics · Vol 14, pp. 2391-2410 · 0 citations · 50 references

TL;DR

The results reveal that, although LLMs occasionally succeed in decoding communicative intentions, their performance is not attributable to human-like ToM reasoning, and offers insight into their interpretive biases, contributing to a deeper understanding of their linguistic capabilities.

Abstract

This study explores the capacity of Large Language Models (LLMs) to perform tasks requiring Theory of Mind (ToM), a critical component of pragmatic language understanding. Although previous work suggests that LLMs may exhibit emergent ToM abilities, this research examines whether such capabilities genuinely involve reasoning about beliefs or merely reflect the reliance on shallow statistical cues. Through a series of controlled experiments featuring indirect speech acts and verbal irony, we assess how belief contexts influence LLM interpretations. The results reveal that, although LLMs occasionally succeed in decoding communicative intentions, their performance is not attributable to human-like ToM reasoning. This work underscores the limitations of LLMs in simulating humanlike ToM and offers insight into their interpretive biases, contributing to a deeper understanding of their linguistic capabilities.1

Read PDF

Similar papers

Preprint Aug 2026

Assessing mentalization in humans and large language models

Different capacities for mentalization across LLMs are demonstrated, and cognitive computational modeling is highlighted as a formal method for assessing comparative intelligence across humans and machines.

Aamir Sohail, Xintong Zhong, Arkady Konovalov et al. · 0 citations
Preprint Aug 2026

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning

Large language models (LLMs) have recently shown strong performance on Theory of Mind (ToM) tests, prompting debate about the nature and validity of the underlying capabilities. At the same time, reasoning-oriented LLMs trained via reinforcement learning with verifiable rewards have demonstrated notable improvements ac...

Ian B. de Haan, P. van der Putten, Max van Duijn · 0 citations
Open access Sep 2026

What primates know about other minds and how often they use it: A computational approach to comparative theory of mind.

Can non-human primates (NHPs) represent other minds? Answering this question is difficult because primates can fail tasks due to a lack of motivation or succeed through simpler strategies. Here, we address these challenges through a computational theory-testing framework for NHP Theory of Mind. In this framework, each...

Marlene D. Berke, Daniel J. Horschler, Amanda L. Royka et al. · 0 citations
Review Open access Jun 2025

From Prompts to Constructs: A Dual-Validity Framework for Large Language Model Research in Psychology.

This review argues that robust AI psychological research requires integrating two methodological traditions: psychometric validation of what a score means and causal inference standards for what the results warrant, developing a dual-validity framework in which evidentiary demands scale with scientific ambition.

Zhicheng Lin · 11 citations · ⚡1
Review Open access Sep 2026

What the mental lexicon might be

The nature of the mental lexicon remains one of the central controversies in cognitive science, bearing on fundamental questions concerning the representation and processing of linguistic knowledge. In this article, we examine Gary Libben’s contributions to the study of lexical representation and processing, with p...

R. D. de Almeida, Lori Buchanan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.