A pyramid-guided multi-scale attention framework based on scale alignment and reliability-aware feature refinement improves occlusion robustness without sacrificing clean-face recognition performance, indicating its practical potential for identity verification and access-control applications involving masks, glasses, and other partial facial occlusions.
Abstract
Facial occlusion degrades face recognition by creating scale-inconsistent identity cues across visible regions and amplifying responses to irrelevant occluders. To address these coupled problems, this study proposes a pyramid-guided multi-scale attention framework based on scale alignment and reliability-aware feature refinement. Hierarchical features are first projected into a shared semantic space, after which local, intermediate, and global dependencies are adaptively weighted according to the available facial information. Channel–spatial refinement is then used to suppress unreliable responses from occluded regions, while an angular-margin objective preserves inter-identity separability from incomplete facial evidence. Experiments on CASIA-WebFace and occluded LFW show that the proposed method achieves an accuracy of 99.26% under clean conditions and 82.72% under occlusion, with an ROC-AUC of 0.8351. At 80% occlusion, the proposed method outperformed the strongest recent baseline, HMPA-GFAF, by 1.41% points and the direct Inception-ResNet-v1 + ArcFace baseline by 5.88% point. These results demonstrate that the proposed framework improves occlusion robustness without sacrificing clean-face recognition performance, indicating its practical potential for identity verification and access-control applications involving masks, glasses, and other partial facial occlusions.
Control comparisons and ablations indicate that the retained model has the most favorable observed cleanness–robustness trade-off among the tested epoch-matched alternatives; however, fixed-checkpoint comparisons on Occlusion-RAF-DB are not significant after Holm correction, while broader cross-domain validation remain...
Xue-Feng Zhao, Yi-Xuan Dong, Zhao-Man Zhong et al.· Italian National Conference...· 0 citations
We propose the Global Context-Aware Dropout Network (GCA-ODN), a CNN-based, computationally practical neural architecture for joint facial landmark detection (FLD) and facial expression recognition (FER) under partial facial occlusion. GCA-ODN learns a shared embedding that encodes facial geometry and affective cues, i...
Low-quality face images often suffer from blur, occlusion, large pose variations and illumination changes, which make facial landmark localization challenging. Existing HRNetbased face alignment methods preserve high-resolution features, but their fixed multi-scale fusion strategy is insufficient for samples with diffe...
Jun Fang, Keng-Can Feng, Zhan-Yu Lin et al.· 2026 2nd International Confe...· 0 citations
Video topic recognition faces core challenges such as the limitations of static representations, cross-modal semantic misalignment, insufficient coverage of single-scale features, and weak temporal dynamic modeling. This paper finds that existing methods have significant deficiencies in local detail perception, leading...
Jin-Yao Zhang· Poster Volume 0007 The 2026...· 0 citations
An occlusion-aware hybrid biometric framework for reliable 3D face recognition that reaches an accuracy of up to 98.7%, even in partial occlusions, and significantly reduces the Equal Error Rate, demonstrating its effectiveness and suitability for real-world biometric authentication applications.
M. L. Gangadhar, A. S. Raju, C. R. Roopashree· Engineering, Technology &...· 0 citations