Decomposing Wrong-Consensus Agreement in LLM Self-Consistency
This paper asks what information wrong-consensus agreement actually contains, and answers with a quantitative decomposition, and contrasts near-complete mechanical agreement in the open-weights models against a larger preference-unexplained residual in the frontier family.