UniOVA: Universal On-demand Video Analytics with Edge-Cloud Collaborative Multimodal LLM
Video analytics is ubiquitous in modern society, and the emergence of Multimodal Large Language Models (MLLMs) has made its application even more extensive. A common way to support MLLM-based video analytics is to continuously stream video to the cloud and then extracts visual features and samples frames for analysis....