Skip to content

DiFA: Dual Evidence Fusion and Aggregation for Token-Level Text Anomaly Detection

Aug 2026 · 0 citations · 41 references
Computer Science

TL;DR

A Dual-evidence framework with adaptive Fusion and Aggregation (DiFA) for token-level anomaly detection, which derives anomaly scores from form-structural and semantic views to capture visible structural abnormality and contextual inconsistency, thereby providing complementary evidence for identifying diverse anomalies.

Abstract

Text anomaly detection, the task of identifying text instances that deviate from normal language patterns, is crucial for language-driven applications. However, most existing methods can only perform document-level anomaly detection, making it hard to locate harmful phrases or support targeted prevention. Recently, there has been an emerging trend toward token-level text anomaly detection, which aims to address the above limitation by identifying anomalous words or fragments within a document. Nevertheless, one representative method mainly relies on representation-space distance measurement, neglecting the complementary roles of different anomaly cues in capturing diverse abnormal patterns. To bridge the gaps, we propose a Dual-evidence framework with adaptive Fusion and Aggregation (DiFA) for token-level anomaly detection. DiFA derives anomaly scores from form-structural and semantic views to capture visible structural abnormality and contextual inconsistency, respectively, thereby providing complementary evidence for identifying diverse anomalies. To combine these two scores with varying numerical scales, DiFA incorporates a calibration and fusion mechanism to adaptively balance the two views. Moreover, to obtain a discriminative document-level score, a multivariate aggregation method is designed to summarize token-level anomaly scores from multiple perspectives, preventing rare anomalous tokens from being diluted. Extensive experiments across various text anomaly detection benchmarks demonstrate that DiFA consistently achieves top performance while maintaining strong efficiency, robustness, and interpretability. The code and scripts are available at: https://github.com/qyy11-com/DiFA.

View source

Similar papers

#machine learning Preprint Sep 2026

SIM: Subspace Interaction-based Method for Token-Level Text Anomaly Detection

Token-level text anomaly detection, as an emerging trend of text anomaly detection, moves beyond coarse-grained document-level detection by localizing anomalous tokens within text. By providing fine-grained abnormality prediction, token-level text anomaly detection plays a critical role in various real-world applicatio...

Ke Yan, Yue Tan, Qing-Feng Chen et al. · 0 citations
Preprint Aug 2026

LLM as Detector: An In-context Learning Approach for Tabular Anomaly Detection

LLM-Detector is proposed, a framework that utilizes the in-context learning capacity of LLMs for structured, prompt-conditioned scoring synthesis, enabling LLMs to derive anomaly detection logic from structured normal-state knowledge.

Tu Nguyen, Dang Nguyen, T. D. Le et al. · 0 citations
Preprint Aug 2026

ProTAGAD: A Foundation Model for TAG Anomaly Detection with Decoupled Topological and Textual Prototypes

This work introduces a novel foundation model for TAG anomaly detection featuring decoupled topological and textual prototypes and constructs dual prototype banks to independently model structural normality and semantic consistency, effectively isolating anomaly cues that are otherwise diluted during coupled aggregatio...

Ziyang Wang, Liwen Wu, Cheng Xie et al. · 0 citations
Open access Sep 2026

Relation Prototype Re-Scoring for CLIP-Based Logical Anomaly Detection and Localization

CLIP-based anomaly detectors have markedly advanced training-free and zero-shot industrial anomaly detection and localization, yet their predictions remain dominated by patch-wise vision–language similarity or anomaly-aware feature scoring. This formulation is intrinsically limited for logical anomalies, in which every...

Hanhoon Park · 0 citations
Preprint Aug 2026

TD-VAD: Breaking Visual Dependence in Video Anomaly Detection with Text-Driven Learning

This work proposes a novel Text-Driven Video Anomaly Detection (TD-VAD) approach, which utilizes video-like text descriptions with temporal characteristics generated by LLM to train a VAD model, without any reliance on target-domain anomaly data.

Shuang-Qing Zhang, Lei-Lei Ma, Zhao Wang et al. · 2 citations
Open access Aug 2026

SETAS-VAD: Semantically Enriched Text-Aligned Scoring for Weakly Supervised Video Anomaly Detection

Under fully reproducible conditions on UCF-Crime and XD-Violence, SETAS-VAD achieves state-of-the-art temporal localization, with per-threshold gains increasing at stricter IoU values, indicating improved boundary precision rather than coarse detection sensitivity.

Mohamed Mahmoud, Mostafa Farouk Senussi, Mahmoud Abdalla et al. · 0 citations

Related blog posts

MIT News · Artificial Intelligence Aug 27, 2026

Looking beyond natural sequences

A new machine-learning framework aims to improve the success rate of computational protein design while moving away from results that reproduce sequences found in nature.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.