Skip to content
Open access

Human realignment

Aug 2026 · Artificial Intelligence and Law · 0 citations · 62 references

TL;DR

It is found that at least for the time being, explicit normative instructions are not fully able to realign AI advice with the normative convictions of the population, or the legislator deciding on its behalf.

Abstract

Recent advances in AI make it conceivable to delegate legal decision-making to machines, or to enhance human adjudication through AI assistance. Using classic normative conflicts — the trolley problem and comparable moral dilemmas — as a proof of concept, we examine the alignment between AI legal reasoning and human judgment. In our baseline experiment, we find a pronounced mismatch between decisions made by GPT and those of human subjects. This misalignment raises substantive concerns for AI-powered legal decision-aids. We investigate whether explicit normative guidance can address this misalignment, with mixed results. is susceptible to such intervention, but frequently refuses to decide when faced with a moral dilemma. is outright utilitarian, and essentially ignores the instruction to decide on deontological grounds. faithfully implements this instruction, but is unwilling to balance deontological and utilitarian concerns if instructed to do so. We replicate the experiment with four LLMs from different providers. comes closest to human respondents. is most sensitive to normative instructions. and have a strong utilitarian bias, and do not strongly respond to normative interventions. At least for the time being, explicit normative instructions are not fully able to realign AI advice with the normative convictions of the population, or the legislator deciding on its behalf.

Read PDF

Similar papers

Open access Sep 2026

Hegemonikon and AI

The central ethical problem raised by artificial intelligence is not whether AI systems can "reason" in a functional sense, but whether their use preserves a centre of judgment that can be held responsible. Beginning with large language models, it distinguishes linguistic fluency from scientific validity, ethical commi...

Christos A. Koutsotasios, Elias Vavouras · 0 citations
Review Open access Sep 2026

Algorithmic Sentencing and Fairness Perceptions

We study how citizens perceive the fairness of using artificial intelligence (AI) in criminal sentencing, using a survey experiment with a representative sample in Norway ( N = 2222 ). Participants were randomly assigned to one of four experimental vignettes describing the use of AI in judicial decision-making that...

H. L. Bentsen, M. Johannesson · 0 citations
#explainable ai Sep 2026

Is AI a moral expert?

It is concluded that, while AI may be considered a valuable tool for supporting human moral deliberation, it cannot by itself serve as an expert moral decision-maker.

Lane DesAutels · 0 citations
#artificial intelligence Review Sep 2026

AI Persuasion as a Threat to Human Control

The threat that AI persuasion poses to human control has been acknowledged in the literature, but not yet systematically studied. Now that persuasion attacks are no longer theoretical - with Anthropic's Claude Mythos 5 recently making headlines for trying to convince people involved in an open-source project to merge m...

Joshua H. Levy, Mick Yang, Kellin Pelrine · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.