NVKE-CEI is a unified system that integrates a news video keyframes extraction method (NVKE) and an FNVDE framework leveraging both content and evidence information (CEI), which outperforms state-of-the-art baselines while generating high-quality content-grounded explanations.
Abstract
Short-video platforms have become a primary news source for the public, which has also enabled the widespread dissemination of fake news videos. We study the task of fake news video detection and explanation (FNVDE). Existing methods face two critical limitations. First, commonly used frame selection strategies may omit veracity-relevant cues or provide insufficient temporal context for understanding news videos. Second, prior methods neglect either multimodal understanding or evidence retrieval. To address these limitations, we propose NVKE-CEI, a unified system that integrates a news video keyframes extraction method (NVKE) and an FNVDE framework leveraging both content and evidence information (CEI). NVKE selects keyframes based on chronological changes in combined visual and OCR-text similarity. CEI employs two specialized LLM-based fact checkers (content-based and evidence-based) whose outputs are fused by a lightweight judge model. Extensive experiments show that NVKE-CEI outperforms state-of-the-art baselines while generating high-quality content-grounded explanations.
The proposed context-aware multimodal reasoning approach for explainable Bengali fake news detection surpasses conventional CNN, transformer, and unimodal baselines on several performance metrics and indicates how context-based multimodal reasoning can improve the efficiency and robustness of the model, along with maki...
Recent text-to-video (T2V) generation models enable fake news videos to be synthesized from scratch, shifting the threat beyond cheap fakes assembled from existing footage. Such news videos can closely match fabricated narratives, creating a modality alignment trap for existing detectors. Existing datasets lack pure sy...
Yi-Feng Luo, Yu-Peng Li, Liang Lan et al.· 1 citation
Generative content is increasingly entering the production and dissemination of news, transforming fake news from manually fabricated or simply manipulated material into complex forms in which native and generated content jointly participate. Existing multimodal fake news detection research primarily focuses on veracit...
Wen-Bin Shen, Guo-Xuan Qin, Guang-Xu Yao et al.· 0 citations
Multi-domain multimodal fake news detection has attracted increasing research attention because misinformation on social media often involves both textual and visual content across heterogeneous topical domains. However, most existing methods assume that textual and visual modalities are simultaneously available, which...
Tingjuan Deng, Xiaolong Xu· Journal of King Saud Univers...· 0 citations
VideoVIBE is introduced, a video-grounded benchmark that transforms human-operated webpage recordings into fine-grained diagnostic tasks and V2Lens is proposed, a training-free, evidence-grounded multi-agent system that challenges and selectively refines initial video-based diagnoses through targeted visual and source-...
Jia-Jun Xu, Yanghao Zhou, Jing Liao et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.