Skip to content

OSR: Output Space Redistribution for Adaptive Label Removal in Classification Models

Sep 2026 · 0 citations · 30 references
Computer Science

TL;DR

A novel approach that leverages statistical redistribution in the output space to approximate the post-removal confidence vectors of a retrained model, alleviating scalability limitations and potentially mitigates privacy concerns inherent to data-dependent solutions is proposed.

Abstract

Label removal occurs frequently in classification systems with evolving taxonomies, where categories must be dynamically updated or eliminated. To accommodate such changes, classification models must adapt accordingly. Existing solutions, broadly categorized as retraining-based and feature-space-adjustment-based, share common limitations despite their variations, including reliance on access to original data, substantial computational and storage costs, inconsistent results, poor scalability, and degradation of model utility. To address this, we propose a novel approach that leverages statistical redistribution in the output space to approximate the post-removal confidence vectors of a retrained model. Applicable as a modular output filter, our method bypasses the burden of feature-space adjustments or loss-function convergence, alleviating scalability limitations. Furthermore, by requiring only existing labels and prior output confidences, the method potentially mitigates privacy concerns inherent to data-dependent solutions. Extensive experiments demonstrate competitive performance against full retraining, with improvements in computational efficiency and privacy preservation across several classification tasks.

View source

Similar papers

Sep 2026

MLIDSC: A self-adaptive online active learning framework for multiclass imbalanced data stream with concept drift

The MLIDSC introduces an automated labeling mechanism that eliminates the need for prior parameter assumptions and uses a novel weighted scheme that combines the imbalance ratio and the importance of individual instances at a given time, ensuring a focus on critical data points.

Bohnishikhan Halder, K. M. Azharul Hasan, Md. Manjur Ahmed · 0 citations
Open access Aug 2026

Label space reduction for transductive zero-shot classification with large language models

This work proposes distilling the model into a probabilistic classifier, enabling lightweight deployment without repeated LLM calls, and demonstrates that LSR improves macro-F1 scores by an average of 7.0% compared to standard zero-shot classification baselines.

Nathan Vandemoortele, Bram Steenwinckel, F. Ongenae et al. · 0 citations
#machine learning Preprint Sep 2026

Adaptive Ensemble Selection for Noisy Labels on Tabular Data

Incorrect or corrupted labels in tabular datasets can significantly degrade supervised learning performance, particularly when mislabeling is subtle and not easily detectable from feature space alone. In the context of automated or AI-augmented data science workflows, robust detection of such label noise is critical fo...

Faizaan Ali, Inwon Kang, O. Seneviratne · 0 citations
Open access Sep 2026

Adaptive Weighting–Synthetic Minority Oversampling Technique

A novel oversampling algorithm: the adaptive weighting–synthetic minority oversampling technique (AW-SMOTE), which combines the two perspectives of boundary tightness and local density and provides global sample enhancement support.

Shen Yan, Hai-Feng Guo, Xiao-Ming Su · 0 citations
Aug 2026

Active Domain Adaptation Under Concept Shift.

This paper proposes ADA-CS, a plug-and-play module compatible with any ADA or ASFDA framework, and introduces a CSS metric to quantify the Concept Shift Severity across domains, revealing that non-negligible concept shift exists in many transfer tasks.

Zi-Kang Zhu, Yiyan Huang, Xing Yan · 1 citation

Related blog posts

MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.