Skip to content
Review Open access

Seemingly conscious AI risks

Aug 2026 · AI and Ethics · Vol 6 · 3 citations · 145 references

TL;DR

A unified framework connecting empirical hallmarks of consciousness attribution to a structured risk taxonomy of Seemingly Conscious AI (SCAI), AI systems that exhibit hallmarks which elicit consciousness attribution from users is provided.

Abstract

AI systems are increasingly designed in ways that lead users to perceive them as conscious. This paper provides a unified framework connecting empirical hallmarks of consciousness attribution to a structured risk taxonomy of Seemingly Conscious AI (SCAI), AI systems that exhibit hallmarks which elicit consciousness attribution from users. We survey the empirical literature to identify five such hallmarks of SCAI, spanning affective capacity, anthropomorphic features, autonomous action, self-reflective behavior, and social-interactive behavior. These provide observable, system-level proxies for this inherently subjective phenomenon, informing its design and enabling its empirical study. Drawing on this foundation, we develop a taxonomy of SCAI risks spanning risks to individuals, including emotional dependence and autonomy erosion, and societal-level harms, including human status erosion and political strife. We complement this conceptual analysis with an expert survey to assess the likelihood of each risk category. We find that risks to individuals, particularly emotional dependence and autonomy erosion, are already observable and rated as high probability, while societal risks, at a low probability, carry high potential severity and path-dependence. The single perceptual mechanism of consciousness attribution is shown to generate this heterogeneous risk surface. We then discuss the implications of these risks and map the multidisciplinary research gaps in this nascent field to inform its research agenda.

Read PDF

Similar papers

#large language models Open access Sep 2026

AI models as consciousness attributors: how LLMs ascribe consciousness to other agents

This work introduces model-generated consciousness attribution as an object of empirical operationalization and diagnosis, defining an attribution rule as the recurring relationship between features of a target and an evaluator’s ratings, without implying subjective belief, intention, or experience.

Bongsu Kang, Chang-Eop Kim · 0 citations

The perceived whiteness of artificial intelligence.

The findings show that AI is not perceived as socially neutral but instead acquires racial meanings associated with credibility, authority, and capability, demonstrating how social categories shape perceptions of novel technological entities beyond their underlying algorithmic properties.

M. Gamez-Djokic, Adam Waytz · 0 citations
Preprint Aug 2026

The Evolutionary Origin of Values: implications for AI alignment, sentience and existential risk

The evolutionary origin of value in biological organisms is traced by tracing the evolutionary origin of value in biological organisms to conclude that the real alignment challenge lies not in preventing rogue AI agency, but in ensuring LLMs intelligently apply learned ethical values.

Francis Heylighen · 0 citations
Conference Open access 2026

We Are Required to Re-Think the World of AI

The paper argues that practical imitation may bypass this barrier by relying on belief and perceived equivalence rather than authentic internalization rather than authentic internalization, and may help ensure that AI remains an auxiliary tool rather than becoming a governing influence over human thought and action.

Jeremy Horne · 0 citations
Open access Aug 2026

Playing with the dials of belief: how controllable AI behaviours could modulate human belief and cognition across scales

A virtual psychopharmacology analogy is developed in which different AI system configurations have effects on belief dynamics that resemble neuromodulatory changes in the precision assigned to social evidence, and which concludes that interaction configurations should be treated as modifiable influences on belief and a...

H. Morrin, L. Nicholls, Q. Deeley et al. · 7 citations
Open access Jul 2026

Prudential rights for strategically capable AI

It is argued that for advanced AI systems deployed in high-stakes environments the more urgent question may be prudential and strategic, and there is a threshold of evidential and strategic risk beyond which it becomes rationally justified to adopt norms of treatment that include constraints on coercion, deletion, and...

Ognjen Arandjelovíc · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.