Skip to content
Open access

Bridging Global Context and Local Detail in State Space Models for Fine-Grained Landslide Segmentation in Remote Sensing Imagery

2026 · IEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing · Vol 19, pp. 29415-29432 · 0 citations · 69 references

TL;DR

This work proposes landslide state-space Mamba, a landslide segmentation network that augments a VSSD backbone with two lightweight local enhancement modules: a structural feature calibration (SFC) module that adaptively calibrates spatial structural discrepancies after global interaction, and a local detail enhancement module (LDEM) that reinforces multiscale textures before spatial downsampling.

Abstract

Accurate landslide mapping from high-resolution remote sensing imagery requires both broad spatial context and precise boundary detail. State space models (SSMs), such as vision state space duality (VSSD), provide global receptive fields with linear complexity, but their global interaction mechanism offers no dedicated support for the fine local structures required by accurate landslide delineation. We propose landslide state-space Mamba (LSMamba), a landslide segmentation network that augments a VSSD backbone with two lightweight local enhancement modules: a structural feature calibration (SFC) module that adaptively calibrates spatial structural discrepancies after global interaction, and a local detail enhancement module (LDEM) that reinforces multiscale textures before spatial downsampling. Both modules are applied only at the shallow encoder stages where spatial detail is richest, introducing limited incremental computational overhead. For decoding, we design an efficient VSSD-based decoder that replaces heavy convolutional fusion heads, improving segmentation accuracy while substantially reducing computational cost. On three diverse benchmarks—Bijie, globally distributed coseismic landslide dataset, and Landslide4Sense—LSMamba consistently achieves the highest mIoU among all compared methods while maintaining a favorable accuracy–efficiency tradeoff. Overall, the results show that lightweight local enhancement is an effective and efficient strategy for fine-grained landslide mapping with SSM backbones.

Read PDF

Similar papers

Open access Sep 2026

HDSMNet: Height-Guided Sparse Cross-Modal Fusion for High-Resolution Remote Sensing Semantic Segmentation

HDSMNet is proposed, a dual-branch multimodal semantic segmentation network designed for optical–nDSM data that enhances discriminative dense feature representations in high-resolution remote sensing images through interaction with a compact set of geometry-guided anchors.

Han-Xun Gu, Jiang-Jie Hu, Li Wang et al. · 0 citations
Preprint Sep 2026

Global-Local Contextual Progressive Expansion Network for Martian Landslide Segmentation in Multimodal Remote Sensing Imagery

Automated landslide segmentation on Mars is one of the important tasks for understanding its surface processes, and all will aid in future space exploration. However, it remains a relatively underexplored open challenge because landslide morphology is highly variable, foreground regions are often sparse or irregular, a...

L. T. Ramos, Sidike Paheding, Abel A. Reyes-Angulo et al. · 0 citations
#artificial intelligence Preprint Sep 2026

SatOV: Restoring Spatial Priors for Training-Free Open-Vocabulary Segmentation in Remote Sensing Imagery

Open-vocabulary semantic segmentation (OVS) of remote sensing imagery is a challenging pixel-level task requiring strong generalization and adaptation to the spatial characteristics of remote sensing data. Although existing vision-language foundation models perform well in general domains, their image-level classificat...

Chang-Hao Zhao, Ling-Lin Zeng, Hai Liu · 0 citations
Conference Sep 2026

Transformer-enhanced land use classification algorithm for high-resolution remote sensing imagery

These findings validate the effectiveness of combining CNNs and Transformer mechanisms in advancing automatic land use recognition and provide a promising pathway for scalable applications in large-scale remote sensing analysis.

Chen-Xi Xu, Rui-Qi Ling, Yi-Chen Sun et al. · 0 citations
Sep 2026

ALSRFormer: An adaptive transformer with dynamic window attention and multi-scale deformable feed-forward network for remote sensing image segmentation.

High-resolution remote sensing images present considerable challenges for semantic segmentation due to their complex object structures and extensive spatial distribution. Effective segmentation requires capturing fine-grained local details while simultaneously modeling long-range dependencies. Convolutional Neural Netw...

Zi-Qi Jia, Jia-Wei Zhang, Dong-En Guo et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.