Skip to content

UG-FPR: uncertainty-guided feature perturbation and refinement for medical foundation models

Jul 2026 · Journal of Electronic Imaging (JEI) · Vol 35, pp. 043009 - 043009 · 0 citations · 54 references
Engineering

TL;DR

A layer-wise uncertainty-guided feature perturbation and refinement framework that operates directly in representation space and can be seamlessly integrated into pretrained transformer encoders is proposed, validating the effectiveness of uncertainty-guided representation diffusion for medical visual understanding.

Abstract

Abstract. Vision Transformers pretrained via self-supervised learning have demonstrated strong representation capability in natural image analysis and are increasingly adopted in medical imaging tasks. However, when transferred to medical domains, pretrained encoders often exhibit limited adaptability to ambiguous anatomical boundaries and low-contrast structures as contextual dependencies learned from natural images may not adequately capture uncertainty characteristics inherent in medical data. In this work, we propose a layer-wise uncertainty-guided feature perturbation and refinement framework that operates directly in representation space and can be seamlessly integrated into pretrained transformer encoders. The proposed method explicitly estimates spatial uncertainty from encoder features and performs controlled semantic diffusion in feature space, enabling selective refinement of ambiguous regions while maintaining stable representations elsewhere. The enhanced representations are integrated through residual modulation, enabling progressive adaptation without disrupting pretrained dynamics. The proposed framework is fully plug-and-play and can be inserted into each transformer layer without modifying the original architecture. Extensive experiments on multiple medical image analysis tasks demonstrate consistent performance improvements over strong transformer baselines, validating the effectiveness of uncertainty-guided representation diffusion for medical visual understanding.

View source

Similar papers

#artificial intelligence Preprint Sep 2026

MUMINS: Metadata-conditioned Uncertainty-aware Medical Image Next-state Synthesis

An efficient diffusion framework that jointly diffuses a baseline scan and its follow-up residual, summed to synthesize the follow-up scan, while concurrently predicting a spatial uncertainty map, in a single reverse diffusion process is proposed.

A. Oliveras, Roger Marí, Rafael Redondo et al. · 1 citation · ⚡1
Preprint Aug 2026

CiUNet: A Hybrid Swin-CNN UNet for Medical Image Segmentation

Medical image segmentation requires high accuracy and robustness, yet practical commercial deployment also demands privacy preservation and computational efficiency. In this context, the U-Net architecture, which can be inherently decoupled into independent encoder and decoder components, serves as a natural commercial...

Bin Dong, Jing-Hong Chen · 0 citations
Open access Sep 2026

Selective Confidence-Guided Projection-Based Encoding for Medical Image Classification

Selective Confidence-guided Projection-based Encoding (SCOPE) is proposed, a conflict-aware KD framework comprising Selective Relation Alignment (SRA) and Gradient Conflict Resolution (GCR), which demonstrates competitive predictive performance, improved training stability, and low computational overhead.

Tao Chen, Chuan Zhou, Yi-Fan Wang et al. · 0 citations
Preprint Aug 2026

P3CA: Encoder-Agnostic Interpretation of Vision Foundation Model Embeddings via Spatial Probing

Position-prompted PCA (P3CA), an encoder-agnostic method for local probing of channel-rich spatial tensors, is proposed and implemented in EmbedVision, an interactive 3D Slicer-based workflow, and evaluated across natural images, colorectal pathology foundation-model embeddings, and spatial transcriptomic tensors.

A. Jamzad, Dilakshan Srikanthan, F. Akbarifar et al. · 0 citations
Open access Sep 2026

Gaussian Context‐Guided Expert Personalization for Multi‐Rater Medical Image Segmentation

A unified framework is presented that simultaneously models segmentation variability and expert‐specific behavior within a single architecture, enabling personalized predictions while preserving diversity and demonstrating consistent improvements over existing methods in both diversity and personalization metrics.

A. Gharawi, M. Alahmadi · 0 citations
Open access Jul 2026

AFFUNet: adaptive feature fusion Transformer U-Net with joint loss function for medical image segmentation

This work proposes Adaptive Feature Fusion U-Net (AFFUNet), an adaptive feature fusion (AFF) Transformer U-Net with a joint loss function, aiming to improve segmentation accuracy through dynamic feature fusion, hard sample reweighting, and explicit boundary optimization.

Ming-Ming Pang, Chu Yang, Yan-Long Luo · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.