This work identifies routing-induced bias, a failure mode in which subgroup imbalance drives the gating network to route subgroups onto a few experts, and proposes an end-to-end Mixture-of-Experts (MoE) framework that corrects it and improves fairness while maintaining competitive predictive performance.
Abstract
Deep learning models often produce performance disparities across demographic groups, due to the training data imbalance with respect to sensitive attributes such as gender or age. To address this problem, existing work has explored fair representation learning, data re-sampling, and adversarial training, which can be broadly categorized into two main approaches. Single-stage methods typically learn a shared representation for fairness, but often struggle to handle heterogeneous subgroup distributions. Two-stage methods learn representations separately from the final prediction task, which can lead to misalignment between fairness objectives and downstream predictions. We identify routing-induced bias, a failure mode in which subgroup imbalance drives the gating network to route subgroups onto a few experts, and propose an end-to-end Mixture-of-Experts (MoE) framework that corrects it. Specifically, we apply subgroup reweighting to correct data imbalance, and introduce gate entropy regularization to prevent routing from collapsing onto subgroup attributes, keeping expert utilization both balanced and interpretable. Beyond improving fairness, the routing distribution offers an interpretable view of how subgroups are allocated across experts. Experimental results demonstrate that the proposed approach improves fairness while maintaining competitive predictive performance.
Although recent deepfake detectors achieve high overall accuracy, their errors remain unevenly distributed across demographic subgroups, with real faces from certain groups more often misclassified as fake. Existing fairness-aware detectors typically regularise the entire feature representation, without identifying or...
Xiaoman Lu, Jiaqi Li, Shuntian Zheng et al.· 0 citations
Deep recommender systems frequently suffer from algorithmic biases that lead to unequal recommendation exposure across demographic groups, particularly disadvantaging underrepresented users. Although various debiasing methods have been proposed, they typically rely on invasive architectural adjustments that disrupt the...
Radhofan Azizi Ramdhani, Rita Rismala· International Conference on...· 0 citations
FairTPT is developed, a novel fairness-aware episodic TTA method that jointly minimizes target marginal entropy while maximizing spurious marginal entropy through soft-prompt tuning and establishes a foundation for robust TTA, which is essential for achieving fairness in practice.
Yoann L. Launay, Parameswaran Kamalaruban, Tom Kempton et al.· 0 citations
Deep learning-based credit scoring systems face three interrelated challenges typically addressed in isolation: unavoidable class imbalance, model opacity, and demographic inequalities. This paper proposes the Explainable Fairness-Aware Deep Learning (EFADL) framework, which is a unified end-to-end pipeline designed to...
Unknown authors· International Journal of Dat...· 0 citations
Although neural-based machine learning models have received a lot of attention recently, tree-based models such as gradient boosting are competitive for tabular data and therefore remain widely used in various applications of AI. As when using other machine learning predictive models, they can however yield discriminat...
Veronika Shilova, Abdoulaye Sakho, Younes Boumoussou et al.· 0 citations
As machine learning (ML) algorithms are increasingly used in high-stakes applications, concerns have arisen that they may be biased against certain social groups. Although many approaches have been proposed to make ML models fair, they typically rely on the assumption that data distributions in training and deployment...