Skip to content

FRAUDSkill: Structured Frozen-Weight Skill Optimization for Audio Anti-Fraud Detection

Sep 2026 · 0 citations · 22 references
Computer Science

TL;DR

FRAUDSkill is proposed, a structured frozen-weight adaptation framework that leaves the underlying audio-language model unchanged while optimizing an external layer of skill programs, route-specific policies, and decision rules and combines structured output control with validation-guided multi-path inference to ensure protocol-compliant predictions.

Abstract

Large audio-language models have shown promise for anti-fraud detection by directly processing speech and reasoning over fraud-related evidence. Their deployment, however, requires predictions to follow a predefined label space and a structured decision protocol consisting of service-scenario identification, fraud detection, and conditional fraud-type classification. Existing fine-tuning and prompt-based approaches typically encode task knowledge, constraints, and decision rules into model parameters or manually maintained prompts, making them difficult to adapt as fraud patterns and labeling policies evolve. To this end, we propose FRAUDSkill, a structured frozen-weight adaptation framework that leaves the underlying audio-language model unchanged while optimizing an external layer of skill programs, route-specific policies, and decision rules. We further combine structured output control with validation-guided multi-path inference to ensure protocol-compliant predictions. On the TeleAntiFraud benchmark, FRAUDSkill achieves 73.50% Macro-F1, outperforming the shared frozen-model baseline by 31.96% while reducing invalid outputs to 1.94%. Extensive experiments demonstrate that external skill optimization provides an effective and adaptable solution for structured audio anti-fraud detection without modifying the underlying model. The source code is available at https://anonymous.4open.science/r/FRAUDSKILL-114514.

View source

Similar papers

Conference Aug 2026

AI-Driven Real-Time Financial Fraud Detection with Explainable Machine Learning

Digital financial fraud has intensified in step with the global proliferation of online payment channels, mobile banking, and contactless transactions. Conventional rule-based detection engines reliant on static thresholds and hand-coded heuristics cannot adapt quickly enough to the pace at which fraud patterns evolve,...

B. S, Alimabeevi A, Lok Ranjan Y. R et al. · 0 citations
Preprint Aug 2026

MINT: A Universal Zero-Shot Predictor for Transaction Data

The Multimodal Instruction Network for Transactions (MINT), a framework that connects a pretrained transaction sequence encoder to a decoder-only LLM through lightweight embedding injection, transaction-language alignment, and instruction tuning, achieves state-of-the-art predictive question-answering performance in bo...

Parameswaran Kamalaruban, Viktor Drobnyi, Maeve Madigan et al. · 0 citations
Conference Aug 2026

Multi-Adapter Qlora for Cross-Domain Financial Fraud Detection

Most fraud detection systems are developed for a single domain at a time: a bank uses one model for credit card fraud, an insurer uses another for claims, and each requires its own large labelled dataset while offering no explanation for its decisions. This paper tests whether a single lightweight language model can in...

Kondapaka Arun, Mursubai Sandhya Rani, K. Shailaja et al. · 0 citations
Review Open access Aug 2026

DKFraudNet: a knowledge-guided adversarial learning framework for fraud user detection

Introduction Fraud user identification in telecommunications is hindered by scarce, noisy, and imbalanced labels, while expert rules may provide ambiguous or contradictory evidence. Methods We propose DKFraudNet, a knowledge-guided framework that integrates domain knowledge regularization, an attention-adaptive conditi...

Ying-Jun Shen, Ren-Da Shi, Kaixi Song et al. · 0 citations
#natural language process... Preprint Sep 2026

Data-Centric Post-Training for Financial Reasoning: Mining, Distillation, and Verifiable Learning

Financial text, textbooks, and question-answer pairs are abundant, but only a small fraction is directly usable for reasoning-focused post-training. Existing QA pairs often lack explicit reasoning, sufficient context, or reliably verifiable answers, while textbooks must first be transformed into synthetic training exam...

Zhirayr Hayrapetyan, Andrei Kalmykov, Denis V. Kokosinskii et al. · 0 citations
Open access Aug 2026

StageGuard: A hierarchical BERT-BiLSTM framework with stage transition anomaly scoring for fraudulent conversation detection

Fraud conducted through chat applications has grown into a serious global threat affecting millions of people each year. Most detection systems still classify each message in isolation rather than reasoning over the conversation as a whole, which limits their ability to recognise the gradual manipulation that character...

Karrar M. Khudhair, Bareq M. Khudhair · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.