Skip to content
Open access

DMFNet: exploring diverse mid-feature for visible-infrared person re-identification

Jul 2026 · Frontiers of Computer Science · Vol 8 · 0 citations · 31 references

TL;DR

DMFNet (Diverse Mid-feature Network) is presented, a novel deep learning architecture that effectively harnesses intermediate shared features to bridge this cross-modal gap and enhances cross-modal matching capabilities but also provides interpretable feature visualizations, offering valuable insights into the network's decision-making process.

Abstract

Visible-infrared person re-identification remains a challenging task due to inherent modality discrepancies between RGB and infrared images. Existing methods often struggle to effectively capture both modality-specific and modality-invariant features simultaneously, limiting their cross-modal matching performance. This paper presents DMFNet (Diverse Mid-feature Network), a novel deep learning architecture that effectively harnesses intermediate shared features to bridge this cross-modal gap. DMFNet integrates two key modules: a Multi-layer Feature Cascade Module (MFCM) that aggregates discriminative features across different network stages, and a Dual Feature Generation Module (DFGM) that produces diverse intermediate representations through Instance-Batch Normalization variants. Extensive experiments on the SYSU-MM01 and RegDB datasets demonstrate that DMFNet achieves state-of-the-art performance, with significant improvements in Rank-1 accuracy (up to 8.2% on SYSU-MM01 and 6.5% on RegDB) and mean Average Precision (mAP) over existing methods. Our approach not only enhances cross-modal matching capabilities but also provides interpretable feature visualizations, offering valuable insights into the network's decision-making process. These results pave the way for more robust person re-identification systems in real-world surveillance scenarios, particularly in low-light conditions where traditional visible-only systems often fail.

Read PDF

Similar papers

Open access Aug 2026

ISFNet: Enhancing Cross-Modal Person Re-Identification via Intermediate Shared Feature Learning

Cross-modal person re-identification between visible and infrared domains remains a challenging problem due to significant modality gaps. This paper presents a novel approach termed Intermediate Shared Feature Network (ISFNet) that explicitly addresses this issue by exploiting intermediate feature representations withi...

Aobo Fan, Wangmeng Wang, Zhixin Tie et al. · 0 citations
Aug 2026

Progressively Biased Split Vision Transformer Learning for Visible-Infrared Person Re-Identification

A Progressively Biased Split Vision Transformer (PBSVT) is proposed, which combines a split ViT backbone with progressive bias training to gradually reduce RGB-dominant bias while preserving modality-shared structure and demonstrates the effectiveness of progressive modality transition for robust VI-ReID representation...

Mengru Jiao, Xin-Yue Xu, Jun-Feng Zhang · 0 citations
Preprint Aug 2026

Multi-scale Decomposed Convolution Refinement Network for Visible-Infrared Person Re-Identification

This work proposes MDCRNet, a Multi-scale Decomposed Convolution Refinement Network that enhances cross-modal feature learning and discriminative metric learning, and develops a Joint Discriminative Metric Loss incorporating a novel Granularity Discriminative Loss (GDL).

Mingsheng Zheng, Zirui Jiang, Bo Liu et al. · 0 citations
Open access Aug 2026

Multi Scale Feature Alignment Network for Cross-Modal Visible-Infrared Person Re-Identification

To address the performance degradation of person re-identification (ReID) under complex lighting and day-night conditions, this study proposes a novel dual-path convolution based multi-scale feature alignment (DCMFA) network. The network mainly focuses on addressing the challenges of modality discrepancy and feature al...

Bailiang Huang, Bin Chen, Tian-Ran Sun et al. · 0 citations
Open access Sep 2026

CMIA-Net: Early Cross-Modal Interaction for Visible-Infrared Person Re-Identification

It is argued that VI-ReID should be treated as an early cross-modal correspondence learning problem rather than only a late embedding alignment problem, and CMIA-Net is proposed, a framework that establishes bidirectional visible-infrared interaction at shallow backbone stages and introduces Spectral-Invariant Augmenta...

Dao-Li Zhang, Qi-Cheng Liu · 0 citations
Open access 2026

Multi-Level Details Recovery: A Reinforced-Transformer for Person Re-Identification

Person re-identification (Re-ID) aims to retrieve the same person across non-overlapping cameras. Despite recent progress, Re-ID remains highly challenging due to high inter-class similarity in large-scale datasets and drastic intra-class variations in cross-platform scenarios (e.g., drones and wearable cameras). While...

Meifeng Liu, Hua Han, A. A. M. Muzahid et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.