Skip to content
Open access

PSAFM: A parameter-free spatial attention fusion module

Sep 2026 · PLoS ONE · Vol 21, pp. e0357958 · 0 citations · 27 references
Medicine

TL;DR

A parameter-free spatial attention fusion module (PSAFM), where the pointwise average- and max-pooling branches are combined by deterministic weighted fusion, which enables the module to strengthen feature responses without adding trainable parameters.

Abstract

Attention mechanisms improve convolutional neural networks (CNNs) by emphasizing informative features, but many existing modules introduce additional parameters, convolutional or fully connected layers, and non-negligible computational overhead. These costs may limit their use in lightweight or plug-and-play CNN architectures. To address this issue, this paper proposes a parameter-free spatial attention fusion module (PSAFM). Specifically, the pointwise average- and max-pooling branches are combined by deterministic weighted fusion, after which channel statistics and a Tanh activation are used to compute adaptive three-dimensional attention weights. This enables the module to strengthen feature responses without adding trainable parameters. We evaluate PSAFM on CIFAR-10 and CIFAR-100 using ResNet, Pre-ResNet, and MobileNetV2 backbones. On CIFAR-10, PSAFM achieves the best accuracy in 8 of 9 tested backbone settings and improves the original backbones by 0.32 to 1.81 percentage points. On CIFAR-100, PSAFM achieves the best accuracy in 6 of 9 settings and improves the original backbones by 0.05 to 1.68 percentage points. In these experiments, PSAFM is compared with representative attention modules. The results show that its parameter count remains essentially unchanged and that only a small FLOPs overhead is introduced. In other words, when model compactness is important, PSAFM is a simple and efficient attention module that improves CNN feature representation.

Read PDF

Similar papers

Open access Aug 2026

Adaptive Feature Integration in CNN–Transformer Networks for Efficient and Interpretable Visual Classification

A novel adaptive fusion framework that adaptively combines CNN and Transformer features through learnable gating, attention-based feature integration, and explainable-AI methods is developed, intended to improve both computational efficiency and model interpretability.

Komal Sharma, Monika Sainger · 0 citations
Open access Sep 2026

Attn-NucleiSeg: Cross-Attention Fusion with Adaptive Scale Selection for Nuclei Segmentation in HE Images

Nuclei instance segmentation in histopathology images is an important yet challenging task for cancer diagnosis and treatment planning. Existing methods often rely on single-encoder architectures that may have difficulty capturing both local morphological details and global contextual information, while fixed receptive...

Rameesha Javed, Nimra Bukhari, Shabir Hussain · 0 citations
Open access 2026

Deep DPA-MSFNet: A Dual-path Attention and Multi-scale Fusion Network for Robust Eye Tumor Classification

—The rapid ascent of deep learning in medical image analysis has been important in the remarkable advancements made in automated tumor identification. This study presents a novel convolutional neural network architecture for efficient tumor classification by integrating multi-scale fusion techniques with dual-path atte...

S. P. Praveen, Shaik Salma Begum, P. Padmavathi et al. · 0 citations
Open access Sep 2026

MedFuse: dual-stream fusion of convolutional and vision transformer-based features for enhanced medical image classification

Medical image classification is fundamental to computer-aided diagnosis. Limited labeled samples, subtle inter-class differences, and heterogeneous lesion morphology make it difficult for a single representation to capture all relevant cues. This study investigates MedFuse, a simple dual-stream framework that combi...

Ya-Jing Ren, Hai Ling, Zheng Gu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.