The introduction of IR275K provides a reproducible foundation for accuracy--efficiency evaluation of infrared MFSR methods, while the architectural analysis offers a concrete starting point for spatially aware SSM design under resource-constrained infrared sensing.
Abstract
Efficient processing is becoming increasingly important in infrared remote sensing, where satellite constellations produce large volumes of observations under constrained detector resolution, power, and downlink bandwidth. Multi-frame super-resolution (MFSR) offers a software-based route to spatial enhancement, but its evaluation in infrared sensing remains fragmented across private datasets and ad-hoc protocols. Existing benchmarks do not explicitly capture the thermal contrast, sensor noise, weak texture, and platform-induced frame-to-frame variation that characterize infrared video. We introduce IR275K, a curated benchmark containing 594 infrared video sequences and 275,196 frames. It provides sequence-level train/validation/test splits and a reproducible X4 evaluation protocol. As an initial architectural probe, we further evaluate CGMamba, a lightweight state-space model with 10.90M parameters and 112.14G FLOPs. CGMamba combines 2D rotary position encoding (2D~RoPE) with center-guided cross-Mamba (CGCM) fusion for implicit multi-frame reconstruction. It achieves 33.19dB PSNR, outperforming infrared single-image super-resolution references by 0.35--0.52~dB at substantially lower computational cost. Ablation results show that removing 2D~RoPE from CGCM causes a 1.53dB drop and severe grid-like artifacts. This indicates that explicit spatial anchoring is critical for stabilizing SSM-based cross-frame gating under infrared conditions. IR275K provides a reproducible foundation for accuracy--efficiency evaluation of infrared MFSR methods, while the architectural analysis offers a concrete starting point for spatially aware SSM design under resource-constrained infrared sensing. Dataset and evaluation resources are available at: https://github.com/InfraRecon7/IR275K.
Object detection in complex lighting and harsh environments significantly benefits from the synergistic deployment of visible and infrared spectra. However, most existing dual-modality methods follow a two-stage pipeline termed "fusion-then-detection", inevitably entailing substantial model complexity and intensive c...
Chao Zeng, Hao Zhao, Ju Zhou· Measurement science and tech...· 0 citations
Visible–thermal object detection benefits from the complementary properties of RGB and thermal imagery, but repeated cross-modal fusion can increase model complexity, particularly in lightweight detectors. This paper proposes SDR-YOLO, a scale-selective detector designed to make better use of shallow spatial details wi...
Li-Juan Wang, Zu-Chao Bao, Bai-Chuan Rong et al.· Remote Sensing· 0 citations
Underwater object detection faces severe challenges caused by light attenuation, scattering, spatially varying turbidity, and boundary blur, which weaken object-related visual signals and reduce localization reliability. This letter presents MED, a Mamba-Enhanced Detector for degradation-aware underwater object detecti...
Yaoming Zhuang, Zi-Rui Fang, Jia-Ming Liu et al.· IEEE Signal Processing Lette...· 0 citations
Infrared small target detection (IRSTD) plays a vital role in remote sensing and automated early-warning infrastructure. However, its performance is deeply constrained by the dim nature of targets and complex background clutter. Although convolutional neural networks (CNNs) have driven advancements in this field, they...
Zhen Huang, Da-Wei Ren, Yan Zhang et al.· IEEE Journal of Selected Top...· 0 citations
Infrared small target detection (IRSTD) is difficult because true targets are sparse and weak, whereas cloud edges, sea clutter, building structures, and sensor noise can produce compact target-like responses. This letter proposes CDMS-Net, which combines context-stable and detail-sensitive predictions through a bounde...
Xi-Jun Wu, Wen-Ding Xiang, Fan Yang· IEEE Geoscience and Remote S...· 0 citations
Small object detection (SOD) in optical remote sensing images (RSI) is essential for aerial perception yet remains challenged by the severe degradation of fine-grained spatial features during downsampling and the high sensitivity of Intersection over Union (IoU) metrics to tiny positional shifts. To address these limit...
Bin Xiao, Jun Yan· Applied Sciences· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.