Skip to content

Author

Zannatul Ferdushie

1 paper indexed here

We haven’t gathered this author’s papers yet. Follow them and we’ll fetch their work.

Not the right person? Other researchers publish under this name.

Open access 2026

Vision-Enabled Virtual Robotic Head with Multi-Modal Communication Capability.

AI-enabled intelligent virtual humanoids are critical to ensure natural and socially aware human-machine interactions in areas such as education, customer services, and companionship. In this paper, we develop a vision-enabled virtual robotic head which is entirely running inside a web browser and includes modules like Speech-to-Text, Large Language Models, Text-to-Speech, and vision processing using FastAPI and WebSocket architecture. The proposed approach generates a dynamic robot-like face capable of realistic talking and facial expressions along with a vision module for face detection and visitor recognition and personalized interaction based on their memory. The proposed architecture ensures low-latency (less than one second) despite the active use of vision processing capabilities. User evaluation on 50 people resulted in 92% face recognition accuracy and high satisfaction scores.

Md. Ashiqussalehin, M. Khushi, Zannatul Ferdushie et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.