Skip to content
#small language model Open access

Construction of a Small Model Based on Large Model Knowledge Distillation in Anomaly Behaviour Recognition for Intelligent Connected Vehicles

Aug 2026 · 電腦學刊 · 0 citations · 13 references

TL;DR

Experimental results validate the feasibility of transferring knowledge from large models under low-computational-power constraints and provide a new technical pathway and engineering reference for recognising anomalies at the edge of intelligent connected vehicles.

Abstract

This paper proposes an end-to-end “large model annotation—small model distillation—onboard inference” framework to address a critical engineering bottleneck in the field of anomaly detection for intelligent connected vehicles: the tension between the strong reasoning capabilities of large language models and the severe resource constraints of edge computing. Specifically, the LongCat-Flash-Lite model (68.5 billion total parameters, approximately 3 billion activated parameters) is used as the teacher model, which performs three-level semantic annotation (Level 0: normal driving, Level 1: suspicious behaviour, Level 2: anomalous behaviour) on a balanced sample set pre-filtered by a TabNet coarse classifier, using few-shot prompting. Subsequently, the annotations serve as supervised signals to efficiently fine-tune the lightweight Qwen3-1.7B student model via Low-Rank Adaptation (LoRA), thereby transferring the large model’s anomaly-discrimination knowledge to the compact student model. Ultimately, the distilled student model performs inference independently on the onboard side without invoking a cloud-based large model API. In simulation experiments based on real vehicle telemetry data (originally 82.82 million records, downsampled to 1 million records), the proposed method achieved the following key results: (1) In the TabNet coarse-filtering stage, the binary classification accuracy reached 98.62% with a macro-average F1 score of 97.15%; (2) The LongCat-Flash-Lite teacher model annotated 120,000 balanced samples with an average confidence of 0.9174, and the distribution of the three annotation levels was 26.3%, 39.2%, and 34.6%, respectively; (3) After LoRA distillation, the student model’s accuracy improved from 27.50% to 32.10% (an absolute improvement of 4.60 percentage points), the macro-average F1 score increased from 0.1624 to 0.2468 (a relative improvement of 52.0%), and precision and recall improved by 7.49 and 2.83 percentage points, respectively. These experimental results validate the feasibility of transferring knowledge from large models under low-computational-power constraints and provide a new technical pathway and engineering reference for recognising anomalies at the edge of intelligent connected vehicles.

Read PDF

Similar papers

Open access Aug 2026

Myriad: a large multimodal model applying vision experts for industrial anomaly detection

A novel large multimodal model applying vision experts for industrial anomaly detection (abbreviated as Myriad), which treats conventional IAD models as VEs and converts their anomaly maps into lightweight prompts that steer a frozen Q-Former toward suspicious regions, while a compact low-rank adapter shapes features f...

Yuanze Li, Haolin Wang, Shihao Yuan et al. · 0 citations
Preprint Aug 2026

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classification, and Mahalanobis-based novelty detection on the public 3W dataset. These pipelines answer \textit{what happened}, but they do not explain \textit{why the model be...

L. G. O. Lopes, Thales Miranda de Almeida Vieira, E. T. de Lima et al. · 0 citations
Conference Jul 2026

A Self-Contained Traffic Anomaly Detection System based on Distillation Learning for Resource Constrained Edge Deployment

Road traffic anomalies pose a critical threat to public safety, with fatality risk increasing by approximately 2.6% per minute of delayed medical response. While deep learning-based detection systems have shown promise, prevailing approaches either depend on computationally intensive architectures or cloud-based infere...

A. V, M. Subash, Maheshwaran Athirstakumar et al. · 0 citations
Open access Aug 2026

Multi domain neural fusion for adaptive anomaly detection in connected and automated vehicles

Connected and Automated Vehicles (CAVs) rely on high-dimensional, multimodal sensor data and Vehicle-to-Everything (V2X) communication to support autonomous driving functions. This strong dependence exposes CAV systems to anomalies arising from sensor faults, environmental disturbances, and coordinated cyber–physical...

Kumar Dorthi, Ravi Kanth Kotha, Neelima Bayyapu · 0 citations
#computer vision Jul 2026

OPD-IAD: From Language Judgment to Industrial Anomaly Detection via On-Policy Self-Distillation

Large vision-language models (LVLMs) have recently shown strong potential for industrial anomaly detection (IAD) by providing image-level anomaly judgments and interpretable defect reasoning. However, current LVLM-based IAD methods still struggle to produce precise pixel-level anomaly maps from generated language judgm...

Shuimu Chen, Jing Jin, Nan Su et al. · 1 citation

Related blog posts

MIT News · Artificial Intelligence Sep 30, 2026

This game-playing AI is the new champ at Stratego

Able to defeat top-ranked human players and more efficient than other models, the new system could help decision-makers in military maneuvers or business negotiations.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.