Skip to content
Review

Mind the Gaps: Mixture-of-Minds for Human Simulation

Aug 2026 · 0 citations · 41 references
Computer Science

TL;DR

Anacreon is introduced, an audience simulation model that targets the individual level within a narrow, well-specified domain and reaches a state-of-the-art ordinal alignment of 0.775, the individual-level accuracy measure on which the field has converged, with a small residual bias.

Abstract

Predicting how a population will answer a new question is a long-standing goal. Statistical methods succeed at the level of the mass but falter at the level of the individual. Large language model simulators inherit this gap. They recover a population's central tendencies while flattening its heterogeneity, and they carry social biases and prompt brittleness that distort individual predictions. This paper introduces Anacreon, an audience simulation model that targets the individual level within a narrow, well-specified domain. Anacreon learns an authorship embedding that separates individuals, clusters a real qualitative corpus around seed people, and trains a dedicated adapter for each cluster, a mixture of minds, on a Gemma~4 12B base. It harvests demographics, psychological traits, and survey responses from public text, and augments each record with a chain-of-emotion. It reduces prompt brittleness by shuffling response options and reduces positive bias by balancing the training distribution. On a large, externally sourced survey, Anacreon reaches a state-of-the-art ordinal alignment of 0.775, the individual-level accuracy measure on which the field has converged, with a small residual bias. The work is a step toward drawing aggregate insight from faithfully simulated individuals.

View source

Similar papers

Review Jul 2026

Analyzing and Correcting Benevolence Bias in Large Language Models

Benevolence bias is identified and measure, a small but consistent tendency for aligned LLMs to lean toward the kinder, safer, more socially approved answer on value-laden survey questions, and is easy to diagnose and straightforward to fix.

Yuanzi Li, Jun-Hao Wang, Minghui Liu et al. · 0 citations
Preprint Aug 2026

Small Foundation Models of Human Cognition and Behaviour

Small cognitively fine-tuned models show promise as noise ceiling estimators for psychological experiments, though their scope remains bounded by the paradigms seen in training.

Nick Oh, F. Gobet · 1 citation
Jul 2026

LLMs struggle to simulate human belief updates in controlled environments

LLMs are increasingly deployed as proxies for human study participants in social science experiments, yet the fidelity of this practice has rarely been tested directly. We test whether six LLMs can simulate individual human belief updates, comparing LLM outputs 1-to-1 against ground truth data from 391 UK participants...

Sebastian Pohl, Harsh Mehta, Pranav Mambayil et al. · 0 citations
#natural language process... Preprint Aug 2026

Benchmarking large language model agent societies against human behavioural distributions

SILICA is an open instrument that tests three doubts of large language model agents: whether the agents behave like the humans they stand in for, whether a finding survives changes to the apparatus that leave the rules untouched, and whether apparent social dynamics are interaction at all rather than the reproduction o...

Raad Bin Tareaf · 0 citations
Preprint Jul 2026

ZenGen: Social Mind for LLMs

ZenGen, an integrated framework for measuring, internalizing, and grounding social intelligence, and Actio, a harness-controlled inference architecture that routes four typed supports into reasoning demonstrate the effectiveness of typed runtime support.

ZenGen Team, Xiang Ao, Jingping Bi et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.