Skip to content
Conference

AMDF-UNet: a boundary-enhanced adaptive multi-scale feature fusion network for abdominal CT multi-organ segmentation

Aug 2026 · International Conference on Computer Graphics and Virtuality · Vol 14315, pp. 1431507 - 1431507-9 · 0 citations · 16 references
Engineering

TL;DR

AMDF-UNet is proposed, a novel segmentation network that integrates a Boundary-Enhanced Channel-Prior Convolutional Attention (BE-CPCA) module and an Adaptive Multi-scale Dilated Fusion (AMDF) module into the U-shaped encoder-decoder architecture, thereby enabling effective segmentation of structures with diverse scales.

Abstract

Accurate segmentation of abdominal CT images plays a critical role in clinical diagnosis and surgical planning. However, the significant size variations among abdominal anatomical structures — ranging from large organs such as the spleen and spine to small lesions such as urinary stones — together with blurred boundaries between adjacent organs, pose substantial challenges for automated segmentation methods. To address these issues, we propose AMDF-UNet, a novel segmentation network that integrates a Boundary-Enhanced Channel-Prior Convolutional Attention (BE-CPCA) module and an Adaptive Multi-scale Dilated Fusion (AMDF) module into the U-shaped encoder-decoder architecture. The BE-CPCA module is deployed at the bottleneck layer, combining dual-pathway channel attention with a Sobel-operator-based boundary-enhanced spatial attention mechanism to simultaneously capture global channel dependencies and fine-grained boundary features. The AMDF module is embedded in the decoder, employing parallel dilated convolutions with learnable adaptive weights to fuse multi-scale contextual information, thereby enabling effective segmentation of structures with diverse scales. The overall training objective combines Focal Loss, Dice Loss, and a morphology-based Boundary Loss to jointly optimize pixel-level classification, region-level overlap, and boundary-level accuracy. Experiments on a clinical abdominal CT dataset comprising five anatomical categories demonstrate that AMDF-UNet achieves a mean Dice Coefficient of 93.45% and a mean IoU of 92.13%, outperforming mainstream methods including UNet, UNet++, DeepLabV3+, Attention U-Net, ResUNet, EGE-UNet, TransUNet, Swin-UNet, and VM-UNet.

View source

Similar papers

SAFM-Net: a dual-branch CNN–GNN network with synergistic attention and frequency-domain modulation for ultrasound thyroid nodule segmentation

Accurate segmentation of thyroid nodules in ultrasound images is essential for thyroid cancer risk assessment and computer-aided diagnosis, yet remains challenging due to ambiguous boundaries and significant shape variations. Existing convolutional neural network (CNN)-based methods effectively capture local features b...

Xi-Cheng Fu, Jing-Yao Lei, Qiang Wang et al. · 0 citations
Open access Sep 2026

GSCA-UNet: a gated spatial-channel attention U-net for accurate skin lesion segmentation

Medical image segmentation is a fundamental component of computer-aided diagnosis, where automatic skin lesion segmentation serves as a critical upstream task by providing pixel-wise delineations for subsequent analysis. Deep encoder-decoder architectures, such as U-Net and its variants, have advanced skin lesion seg...

La-Zhen Zhou, Wen-Jie Ou, Xiu-Hua Chen et al. · 0 citations
Review Open access Aug 2026

A Cascaded Deep Learning Framework for Robust Liver CT Segmentation Using ROI Refinement and Patient-Level Cross-Validation

A failure-aware cascaded deep learning framework for automated liver CT segmentation using the publicly available HCC-TACE-Seg dataset is presented and indicates that cascaded localisation and region-of-interest refinement can provide robust liver segmentation while reducing background interference and supporting uncer...

Nisha Joseph, D. Mohan, Jomy George et al. · 0 citations
Open access Aug 2026

REC-UNet: a 2D U-Net model with residual cross-dimensional attention for liver tumor segmentation

A novel residual “Enhancement-Calibration” U-Net architecture, termed REC-UNet, which achieves high overall segmentation accuracy across diverse lesion sizes and contrast conditions without relying on explicit size-stratified optimization.

Zhiyuan Wang, Li-Jun Liang, Wei Wu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.