Toward Sustainable Urban Mobility: A Multimodal Large Language Model (MLLM) Framework for Automated Driver Performance Assessment with YOLOv8-Based Scene Detection
An exploratory proof-of-concept framework for automated driver evaluation that combines real-world dashcam footage, YOLOv8-based object detection, and multimodal large language models (MLLMs), specifically Gemini 1.5 Flash is presented.
Mamatha Byreddy, Yara Zayed, Anas M. R. Alsobeh et al.
· Infrastructures · 0 citations