Skip to content
Preprint

Fairness-Aware Mixture-of-Experts via Subgroup Reweighting and Gate Regularization

Aug 2026 · 0 citations · 22 references
Computer Science

TL;DR

This work identifies routing-induced bias, a failure mode in which subgroup imbalance drives the gating network to route subgroups onto a few experts, and proposes an end-to-end Mixture-of-Experts (MoE) framework that corrects it and improves fairness while maintaining competitive predictive performance.

Abstract

Deep learning models often produce performance disparities across demographic groups, due to the training data imbalance with respect to sensitive attributes such as gender or age. To address this problem, existing work has explored fair representation learning, data re-sampling, and adversarial training, which can be broadly categorized into two main approaches. Single-stage methods typically learn a shared representation for fairness, but often struggle to handle heterogeneous subgroup distributions. Two-stage methods learn representations separately from the final prediction task, which can lead to misalignment between fairness objectives and downstream predictions. We identify routing-induced bias, a failure mode in which subgroup imbalance drives the gating network to route subgroups onto a few experts, and propose an end-to-end Mixture-of-Experts (MoE) framework that corrects it. Specifically, we apply subgroup reweighting to correct data imbalance, and introduce gate entropy regularization to prevent routing from collapsing onto subgroup attributes, keeping expert utilization both balanced and interpretable. Beyond improving fairness, the routing distribution offers an interpretable view of how subgroups are allocated across experts. Experimental results demonstrate that the proposed approach improves fairness while maintaining competitive predictive performance.

View source

Similar papers

Preprint Aug 2026

FairReL: Deepfake Detection using Fairness-Aware Representation Learning

Although recent deepfake detectors achieve high overall accuracy, their errors remain unevenly distributed across demographic subgroups, with real faces from certain groups more often misclassified as fake. Existing fairness-aware detectors typically regularise the entire feature representation, without identifying or...

Xiaoman Lu, Jiaqi Li, Shuntian Zheng et al. · 0 citations
Conference Jul 2026

Enhancing User-Side Fairness on Neural Collaborative Filtering Using Generative Adversarial Network-Based Augmentation

Deep recommender systems frequently suffer from algorithmic biases that lead to unequal recommendation exposure across demographic groups, particularly disadvantaging underrepresented users. Although various debiasing methods have been proposed, they typically rely on invasive architectural adjustments that disrupt the...

Radhofan Azizi Ramdhani, Rita Rismala · 0 citations
Preprint Aug 2026

Fairness-Aware Test-Time Prompt Tuning

FairTPT is developed, a novel fairness-aware episodic TTA method that jointly minimizes target marginal entropy while maximizing spurious marginal entropy through soft-prompt tuning and establishes a foundation for robust TTA, which is essential for achieving fairness in practice.

Yoann L. Launay, Parameswaran Kamalaruban, Tom Kempton et al. · 0 citations
Open access 2026

An explainable fairness-aware deep learning framework for credit score classification on imbalanced financial data

Deep learning-based credit scoring systems face three interrelated challenges typically addressed in isolation: unavoidable class imbalance, model opacity, and demographic inequalities. This paper proposes the Explainable Fairness-Aware Deep Learning (EFADL) framework, which is a unified end-to-end pipeline designed to...

Unknown authors · 0 citations
Preprint Jul 2026

OT-FairBoost: Optimal Transport-Guided Gradient Boosting for Fairness Regularization on Tabular Data

Although neural-based machine learning models have received a lot of attention recently, tree-based models such as gradient boosting are competitive for tabular data and therefore remain widely used in various applications of AI. As when using other machine learning predictive models, they can however yield discriminat...

Veronika Shilova, Abdoulaye Sakho, Younes Boumoussou et al. · 0 citations
Conference 2023

Fairness and Accuracy Under Domain Generalization

As machine learning (ML) algorithms are increasingly used in high-stakes applications, concerns have arisen that they may be biased against certain social groups. Although many approaches have been proposed to make ML models fair, they typically rely on the assumption that data distributions in training and deployment...

Thai-Hoang Pham, Xue-Ru Zhang, Ping Zhang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.