Skip to content
Open access

Pyramid-guided multi-scale self-attention and channel–spatial refinement for occlusion-robust face recognition

Sep 2026 · Discover Computing · Vol 29 · 0 citations · 34 references

TL;DR

A pyramid-guided multi-scale attention framework based on scale alignment and reliability-aware feature refinement improves occlusion robustness without sacrificing clean-face recognition performance, indicating its practical potential for identity verification and access-control applications involving masks, glasses, and other partial facial occlusions.

Abstract

Facial occlusion degrades face recognition by creating scale-inconsistent identity cues across visible regions and amplifying responses to irrelevant occluders. To address these coupled problems, this study proposes a pyramid-guided multi-scale attention framework based on scale alignment and reliability-aware feature refinement. Hierarchical features are first projected into a shared semantic space, after which local, intermediate, and global dependencies are adaptively weighted according to the available facial information. Channel–spatial refinement is then used to suppress unreliable responses from occluded regions, while an angular-margin objective preserves inter-identity separability from incomplete facial evidence. Experiments on CASIA-WebFace and occluded LFW show that the proposed method achieves an accuracy of 99.26% under clean conditions and 82.72% under occlusion, with an ROC-AUC of 0.8351. At 80% occlusion, the proposed method outperformed the strongest recent baseline, HMPA-GFAF, by 1.41% points and the direct Inception-ResNet-v1 + ArcFace baseline by 5.88% point. These results demonstrate that the proposed framework improves occlusion robustness without sacrificing clean-face recognition performance, indicating its practical potential for identity verification and access-control applications involving masks, glasses, and other partial facial occlusions.

Read PDF

Similar papers

Open access Aug 2026

Compact Occlusion-Robust Facial Expression Recognition via Clean-Anchored Hard Occlusion Fine-Tuning

Control comparisons and ablations indicate that the retained model has the most favorable observed cleanness–robustness trade-off among the tested epoch-matched alternatives; however, fixed-checkpoint comparisons on Occlusion-RAF-DB are not significant after Holm correction, while broader cross-domain validation remain...

Xue-Feng Zhao, Yi-Xuan Dong, Zhao-Man Zhong et al. · 0 citations
Sep 2026

GCA-ODN: A global context-aware dropout network for joint facial landmark detection and emotion recognition under occlusion.

We propose the Global Context-Aware Dropout Network (GCA-ODN), a CNN-based, computationally practical neural architecture for joint facial landmark detection (FLD) and facial expression recognition (FER) under partial facial occlusion. GCA-ODN learns a shared embedding that encodes facial geometry and affective cues, i...

Muhammad Sadiq · 0 citations
Conference Aug 2026

A Quality-Aware and Multi-Scale Representation Network for Low-Quality Face Alignment

Low-quality face images often suffer from blur, occlusion, large pose variations and illumination changes, which make facial landmark localization challenging. Existing HRNetbased face alignment methods preserve high-resolution features, but their fixed multi-scale fusion strategy is insufficient for samples with diffe...

Jun Fang, Keng-Can Feng, Zhan-Yu Lin et al. · 0 citations
Conference 2026

Multi-scale local perception video topic recognition method based on semantic-guided feature pyramid

Video topic recognition faces core challenges such as the limitations of static representations, cross-modal semantic misalignment, insufficient coverage of single-scale features, and weak temporal dynamic modeling. This paper finds that existing methods have significant deficiencies in local detail perception, leading...

Jin-Yao Zhang · 0 citations
Open access Aug 2026

A Reliability-Based Multimodal Framework for 3D Face Recognition under Occlusion

An occlusion-aware hybrid biometric framework for reliable 3D face recognition that reaches an accuracy of up to 98.7%, even in partial occlusions, and significantly reduces the Equal Error Rate, demonstrating its effectiveness and suitability for real-world biometric authentication applications.

M. L. Gangadhar, A. S. Raju, C. R. Roopashree · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.