Skip to content
Open access

PGA-LLM: A Probability-Guided Alignment Large Language Model Framework for Fault Diagnosis

Aug 2026 · Technologies · Vol 14, pp. 494 · 0 citations · 27 references

TL;DR

PGA-LLM, a novel fault diagnosis framework for industrial equipment that leverages large language models via probability-guided alignment and a progressive three-stage training scheme, encompassing encoder pre-training, interface optimization, and low-rank adaptation of Qwen2.

Abstract

Fault diagnosis for complex industrial equipment plays a crucial role in safeguarding production safety and advancing the capabilities of intelligent operation and maintenance. Current deep learning approaches have demonstrated promising accuracy in fault classification tasks; however, their signal representations alone cannot provide a transparent interface for embedded large language models. To tackle the aforementioned challenges, we propose PGA-LLM, a novel fault diagnosis framework for industrial equipment that leverages large language models via probability-guided alignment. First, a variational autoencoder (VAE)-based signal encoder embedded with reconstruction constraints is established. Joint reconstruction and classification objectives balance discriminative representation learning and signal reconstruction. Second, the probability-guided alignment (PGA) module combines fault-class probability guidance with a residual feature path; a learned gate fuses both paths before continuous soft-prompt projection. Furthermore, a progressive three-stage training scheme is adopted, encompassing encoder pre-training, interface optimization, and low-rank adaptation (LoRA) of Qwen2.5-1.5B. Extensive experiments are carried out on four standard datasets, CWRU, Gear, Mixed, and MBHM, and the Stage 2 signal-side output achieves classification accuracies of 97.1%, 99.0%, 93.4%, and 96.3%, respectively. The report-generation branch provides a schema-constrained signal-to-language interface for maintenance-oriented reporting.

Read PDF

Similar papers

Open access Sep 2026

A Multi-Stage Post-Training Framework for Domain-Specific Language Models in Fault Diagnosis

The rapid advancement of large language models (LLMs) has created new opportunities for intelligent fault diagnosis, particularly in complex industrial systems, such as heating, ventilation, and air conditioning (HVAC) in urban rail transit. Although LLMs have shown strong general reasoning capabilities, adapting them...

Wei Zhang, Hui Fang, Tong-Le Wu et al. · 0 citations
Open access Sep 2026

FD-SE-LLM: A Semantic-Enhanced Large Language Model Framework for Fault Diagnosis of Hydropower Carbon Brush Bearings

Cross-condition fault diagnosis remains a fundamental challenge in intelligent operation and maintenance of hydropower equipment, where deep learning methods suffer 20%−30% accuracy degradation under operating condition drifts. Large language models (LLMs), through massive multi-domain pre-training, offer cross-domain...

Li-Yan Yang, Li-Qing Zhu, Cheng-Lin Yang et al. · 0 citations
Open access Aug 2026

A vision–language-guided multimodal framework with dynamic prototype memory for small-sample fault diagnosis

Rotating machinery fault diagnosis under limited training samples and varying operating conditions remains challenging, as data scarcity and domain shift often weaken the generalisation capability of deep learning models. Recent advances in large-scale pretrained foundation models have provided transferable priors for...

Jin-Kai Fan, Fu Zhao, Yan Gong et al. · 0 citations
Conference Open access 2026

Large Language Model Technologies: Progress, Problems and Prospects

Large language models (LLMs) are built on the classic Transformer architecture and have become a core driving force for the rapid development of modern artificial intelligence. This paper presents a systematic review of LLMs, elaborating on their fundamental working principles, mainstream open-source models, effective...

Siyi Fan · 0 citations
Open access Sep 2026

A novel source-free domain adaptation with high-confidence sample selection and feature disentanglement for machinery fault diagnosis

Motivated by privacy concerns and the high cost of measured data transmission, source-free domain adaptation (SFDA) has attracted increasing attention for intelligent fault diagnosis. Instead of accessing raw measured source-domain data, SFDA transfers knowledge from pre-trained source models to target domains. However...

Yi-Ming Yuan, Kang Wu, Xing-Xing Jiang et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.