Vision Language Model Fusion for Explainable Face Recognition
This work evaluates four VLMs as standalone face verification systems and proposes a fusion framework, where two source models provide similarity scores and textual justifications and a third VLM acts as a decider model, which achieves higher recognition accuracy and fused explanations that are expected to be more robu...