Skip to content
Open access

Geometry and mask aware vision transformer for masked face recognition in unconstrained scenarios

Aug 2026 · Engineering Research Express · Vol 8 · 0 citations · 52 references
Physics

TL;DR

Facial identity identification in unrestricted real-world environments may benefit from this model, which performs well in identifying and verifying low-quality and cross-pose masked faces and outperforming the various state-of-the-art methods and previously proposed methods.

Abstract

The concealed facial features make it difficult to identify masked faces. The features are further distorted and deteriorated when masked faces are combined with low resolution and pose variation in unrestricted contexts. Pose variation and low-resolution circumstances, along with face masks, are not assessed for current masked face recognition systems. We developed a transformer-based model a mask-aware geometry-guided vision transformer (MGViT), to address these issues. First, a learnable geometry-guided patch weighting (LGW) is used in the proposed model to suppress occluded regions and concentrate on the key face regions. Second, a mask-aware feature adaptor is created to improve the domain embeddings between masked and unmasked faces. Following that, by combining Identity and Consistency Loss functions to align identity, a strong consistency learning is integrated. Experiments with various challenges are carried out on the various masked face datasets. The proposed model, MGViT, performs well in identifying and verifying low-quality and cross-pose masked faces. Additionally, the model achieves 94.70%, 95.25%, 87.67%, and 78.56% accuracy on RMFRD, Masked LFW, Masked Multi-PIE, and Masked LR, respectively, outperforming the various state-of-the-art methods and previously proposed methods. Facial identity identification in unrestricted real-world environments may benefit from this model.

Read PDF

Similar papers

Conference Jul 2026

A Hybrid CNN–Transformer Network for Robust Masked and Occluded Face Recognition in Smart Surveillance Systems

Face recognition systems applied to smart surveillance settings often experience poor performance when the faces are partially occluded by a mask or other objects. Occlusions eliminate critical facial information, which makes face identification much more difficult for traditional deep learning models. To solve this is...

R. R, Anbalagan E · 0 citations
Open access Aug 2026

A Reliability-Based Multimodal Framework for 3D Face Recognition under Occlusion

An occlusion-aware hybrid biometric framework for reliable 3D face recognition that reaches an accuracy of up to 98.7%, even in partial occlusions, and significantly reduces the Equal Error Rate, demonstrating its effectiveness and suitability for real-world biometric authentication applications.

M. L. Gangadhar, A. S. Raju, C. R. Roopashree · 0 citations
2026

DiEL: Disentangled Evolutionary Learning for Identity-Preserving Face Enhancement and Recognition

The purpose of face enhancement tasks is to improve the recognition of faces, thus adapting to diverse visualization and recognition demands. However, the performance of the majority methods is drastically degraded under extreme conditions, including large pose variations, low resolution, blur, occlusion, and illuminat...

Jingwei Xin, Tian Yang, Jun Hao et al. · 0 citations
Open access Sep 2026

Pyramid-guided multi-scale self-attention and channel–spatial refinement for occlusion-robust face recognition

A pyramid-guided multi-scale attention framework based on scale alignment and reliability-aware feature refinement improves occlusion robustness without sacrificing clean-face recognition performance, indicating its practical potential for identity verification and access-control applications involving masks, glasses,...

Qi-Nan Zhu · 0 citations
Conference Jul 2026

Vision Transformer with Attention Rollout for Deepfake Face Image Detection and Localization

Generative AI and synthetic media generation tools have enabled widespread media manipulation tools and raised important privacy concerns with misinformation, identity fraud and the verification of authenticity of media. Most of the current convolution-based deepfake detection methods are hard to be deployed in real sc...

K. V. Sai Phani, Shaik Mahaboob Jailan, F. Mahammad et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.