Skip to content

Adaptive Multi-Prototype Open-Set Recognition Network for Medical Image Classification.

Aug 2026 · IEEE transactions on bio-medical engineering · Vol PP, pp. 1-12 · 0 citations
Medicine

TL;DR

This work proposes a novel framework called Adaptive Multi-Prototype Network with Pretrained Swin Transformer (PSW-AMPN) for OSR on medical images that significantly outperforms existing baselines and achieves state-of-the-art performance on multiple medical image OSR tasks.

Abstract

OBJECTIVE Existing prototype learning methods for open-set recognition (OSR) often use a fixed number of prototypes to represent each class, which struggle to model the inherent intra class variations widely existing in practical scenarios. This limitation is particularly pronounced in medical image applications, where high intra-class heterogeneity makes accurate OSR challenging. To address this issue, we propose a novel framework called Adaptive Multi-Prototype Network with Pretrained Swin Transformer (PSW-AMPN) for OSR on medical images. Specifically, a sparse gated attention module is devised to compute attention scores based on prototype-sample relations, thereby adaptively suppressing redundant sub-prototypes and selectively activating discriminative ones for each class in an end to-end optimization process. By jointly optimizing classification loss and the regularization for open space risk based on multiple prototypes, PSW-AMPN effectively captures complex intra-class structures and enhances class discriminability. Furthermore, PSW-AMPN uses a Pretrained Swin Transformer and a lightweight projector as feature extractor to effectively capture both local and global features. Extensive experiments demonstrate that our approach significantly outperforms existing baselines and achieves state-of-the-art performance on multiple medical image OSR tasks.

View source

Similar papers

Open access Sep 2026

An imbalance-aware class weighted TernausResnet architecture for high-fidelity medical image classification

CWTRNet: A Class weighted TernausResnet framework with adaptive optimization is proposed in this work for high reliability medical image classification. Deep learning(DL) models have made significant advancements in the analysis of medical images, especially in diagnosis and classification. TernausNet, on the other h...

P. P. Lakshmi, M. Sivagami · 0 citations
Open access Aug 2026

A medical image classification algorithm based on a hierarchical and complementary attention-enhanced Swin Transformer model

Comparisons across all datasets confirm that the proposed framework exhibits strong robustness and generalization capability when processing multiple medical imaging modalities, thereby providing reliable technical support for computer-aided medical diagnosis systems.

Ya-Chao Si, Yi Zhang, Ming-Zhan Zhao · 0 citations
Open access Sep 2026

A Unified Dual-Stream Framework for Heterogeneous and Imbalanced Medical Image Classification

Medical image classification models often struggle to generalise across heterogeneous clinical domains owing to variations in visual characteristics, acquisition conditions, and class imbalance. Existing studies largely address these challenges independently, with limited investigation into their combined impact on cla...

Samuel Ovuehor, Adi El-Dalahmeh, Usman Adeel et al. · 0 citations
Open access Aug 2026

Hybrid MobileNetV2–vision transformer approach for multi-label classification of chest X-rays

A hybrid MobileNetV2–Vision Transformer (ViT) framework for multi-label classification of CXR images into 14 disease categories on NIH CXR14 dataset is introduced, which adaptively optimizes key hyperparameters of the MobileNetV2–ViT framework to achieve improved accuracy, faster convergence, and enhanced computational...

R. Raj, Pavan M. P. Kumar, K. N. Manjunath et al. · 0 citations
Open access Sep 2026

DA-MoE: descriptor-attention mixture-of-experts for multi-class gastrointestinal disease classification

Computer-aided diagnosis (CADx) for gastrointestinal (GI) endoscopy increasingly depends on deep models trained end-to-end on raw images. However, raw images are often unavailable in legacy clinical systems or privacy-sensitive settings. This work presents a methodological study on multi-class GI disease classification...

Yi-Liu Xu, Ling-Ling Liu, Mei-Wen Tang · 0 citations
Preprint Aug 2026

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

ARMDIL is an ensemble that uses a multimodal large language model (MLLM) agent to dynamically route each image to the most suitable vision backbone, drastically improves adaptability by allowing new information to be integrated via simple prompt modifications, while enhancing interpretability through natural language r...

Daniel A. Perkins, J. Squires, Janou Milligan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.