Skip to content
Preprint

Enhanced Deformable Convolution with Center-invariant Offset and Edge-aware Mask

Sep 2026 · 0 citations · 47 references
Computer Science

TL;DR

Experiments show that EDC outperforms state-of-the-art deformable convolution variants, including Deformable ConvNets V1-V4 and Entire Deformable ConvNets, across mainstream segmentation datasets with various decoder settings, and ablation studies confirm the effectiveness of each component.

Abstract

Deformable convolution networks have recently become popular for many computer vision tasks, especially for semantic segmentation, because of their exceptional capabilities in dynamic spatial modeling. However, due to the dense deformable offsets and the lack of longer-range dependencies, they can not fully adopt proper and precise deformations for feature representations. To tackle the issues, in this paper, we propose Enhanced Deformable ConvNets (EDCN) for semantic segmentation. Specifically, a novel Enhanced Deformable Convolution (EDC) is exploited in the decoder, which integrates the Center-invariant Offset Module (COM) and Edge-aware Mask Module (EMM). The COM employs larger kernels and eliminates deformations at the kernel center, obtaining offsets that are more in line with the target from richer spatial information. Concurrently, the EMM obtains the significance of image content via Sobel edge detection, then selectively applies deformations based on the content significance, minimizing unnecessary deformations associated with relatively less important information, thereby avoiding impact from less informative regions. Experiments show that EDC outperforms state-of-the-art deformable convolution variants, including Deformable ConvNets V1-V4 and Entire Deformable ConvNets, across mainstream segmentation datasets with various decoder settings. Moreover, ablation studies confirm the effectiveness of each component. In addition, visualizations illustrate that EDC enhances spatial adaptation and target focus. We further analyze the extendibility of EDC to larger kernels on the image classification benchmark. Code will be publicly released.

View source

Similar papers

Open access Aug 2026

Weighted multi-scale and wavelet-enhanced Segment Anything Model for salient object detection

Salient Object Detection (SOD) is concerned with isolating the visually most noticeable objects in an image via precise segmentation. Previous approaches, especially those that adapt the Segment Anything Model (SAM), often yield saliency maps that contain incomplete object masks, blurred boundaries, and a lack of fine-...

Zhe Liu, Dan Tian · 0 citations
Sep 2026

ALSRFormer: An adaptive transformer with dynamic window attention and multi-scale deformable feed-forward network for remote sensing image segmentation.

High-resolution remote sensing images present considerable challenges for semantic segmentation due to their complex object structures and extensive spatial distribution. Effective segmentation requires capturing fine-grained local details while simultaneously modeling long-range dependencies. Convolutional Neural Netw...

Zi-Qi Jia, Jia-Wei Zhang, Dong-En Guo et al. · 0 citations
Conference Sep 2026

Understanding Domain-Shift Immunity in Deep Deformable Registration

It is shown that domain-shift immunity is an inherent, largely architecture-agnostic property of deep de-formable registration when trained with a robust pipeline and offered a principled explanation for the cross-domain generalizability of deep registration networks.

Mingzhen Shao, Sarang C. Joshi · 0 citations
Aug 2026

Lightweight Condition-Constrained Salient Object Detection Based on Elastic Pixel Difference Convolution.

The core innovation lies in a novel elastic pixel differential convolution operator, which flexibly captures microscopic gradient cues to effectively coordinate high-level semantics with low-level details during inference with zero additional computational overhead.

Wei-Yi Wei, Yongxin Shi, Shengxia Gao et al. · 0 citations
Preprint Sep 2026

DiffReID: Discriminative Diffusion Model for Object Re-Identification

As a fundamental image processing task, object Re-Identification (ReID) aims to retrieve objects across non-overlapping cameras. Recently, with the development of deep learning, significant advancements have been made in object ReID. However, most existing methods suffer from generalization due to the limited size and...

Ying-Quan Wang, Ping-Ping Zhang, Dong Wang et al. · 0 citations
Aug 2026

Universal Representation for Real-World Misaligned Infrared-Visible Image Fusion.

Infrared and visible image fusion is pivotal for robust visual perception across all weather conditions and scenes. Although deep learning-based methods have made notable progress, most either assume pre-aligned inputs or rely on implicit feature-space alignment, which fails to fundamentally address the amplification o...

Jin-Yuan Liu, Zengxi Zhang, Jiahao Zhang et al. · 1 citation

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.