Skip to content
Open access

From co-creation to technical bias detection methods: an interdisciplinary showcase from the BIAS project

Jul 2026 · Frontiers in Artificial Intelligence · Vol 9 · 0 citations · 61 references
Medicine

TL;DR

BIAS-WEAT and BIAS-SEAT are introduced, two novel metrics designed to detect biases in word embeddings and language models for Dutch, German, Icelandic, Italian, Norwegian and Turkish.

Abstract

Societal stereotypes are often reflected in, and can be reinforced by, machine learning models and linguistic resources such as word embeddings. While various benchmarks and bias detection methods have been proposed, most focus exclusively on English. When applied to other languages, these approaches typically rely on direct translations of English resources, overlooking language- and culture-specific nuances. In this paper, we introduce BIAS-WEAT and BIAS-SEAT, two novel metrics designed to detect biases in word embeddings and language models for Dutch, German, Icelandic, Italian, Norwegian and Turkish. Drawing on real-world biases identified through co-creation workshops with native speakers in the context of a hiring situation, we translated these insights into technical evaluation metrics that are applicable to general-purpose language resources. Our interdisciplinary study demonstrates how language models embed and reproduce biases that are specific to their linguistic and geographic contexts, underscoring the need for culturally grounded approaches to bias detection.1

Read PDF

Similar papers

Preprint Aug 2026

Subjective Multi-Bias Detection with Large Language Models

This project delved into the pervasive challenge of bias detection within the text content by detecting three different types of multi-span biases in corpus WIKIBIAS with more than 4,000 sentence pairs from Wikipedia edits.

Rui-Yu Li, Zhi-Ying Zhu · 0 citations
Aug 2026

Quantifying Social Biases in Language Model Classifiers is Domain-Dependent

This work investigates whether large language models (LLMs) can automatically adapt template-based bias datasets to specific domains using zero-shot prompting and shows that domain-adapted templates capture real-world bias patterns more faithfully than standard templates.

Tamara Quiroga, Felipe Bravo-Marquez, Valentin Barrière · 0 citations
#artificial intelligence Preprint Aug 2026

Evaluating and Mitigating Anti-LGBTQ Biases in German and Multilingual Language Models

A multilingual German-English benchmark dataset that combines community-sourced stereotypes from German-speaking queer individuals with a German translation of WinoQueer is introduced, showing that language models reproduce anti-queer stereotypes, with variation across identities and models.

M. Morch, Daniel Braun · 0 citations

Beyond Good Intentions: When Does the Framing of Multilingual and Low-Resource NLP Research Become a Caricature?

Building language technologies and conducting NLP research for low-resource languages---particularly when led by native speakers or involving participatory research practices---are often framed as means of addressing inequality, serving local communities, and, at times, contributing to *decolonisation*. In this paper,...

Nedjma Djouhra Ousidhoum, Noopur Zambare, Mohamed Abdalla · 0 citations
Review Open access Aug 2026

Artificial Minds, Cultural Shadows: Cultural Alignment, Identity, and Voice Across Multiple Large Language Models

Comparison of five widely used large language models suggests that AI-generated language may shape how culturally situated perspectives are expressed, with differences across models indicating that AI-generated language may shape how culturally situated perspectives are expressed.

Ashkan Goudarzi, Aylar Naderi Zonouz · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.