Skip to content
Preprint

Large reasoning models for abnormal situation management in safety-critical industrial processes

Aug 2026 · 0 citations · 53 references
Engineering Computer Science

TL;DR

It is shown that a general-purpose large reasoning model, with no task-specific training and only the information available to an operator, manages abnormal situations at run time through a bounded, programmatically verified action interface.

Abstract

Automation operates safety-critical processes inside their design envelope and leaves abnormal situations to human operators. Mismanagement of these situations is a leading contributor to process-safety incidents and a hindrance to achieving autonomy. Here we show that a general-purpose large reasoning model, with no task-specific training and only the information available to an operator, manages abnormal situations at run time through a bounded, programmatically verified action interface. Across 39 abnormal situations and operating-point changes on a plant-wide industrial benchmark process, the reasoning model maintained the plant within all hard constraints in all 39, while basic regulatory control failed in 15. It matched the plant's expert-engineered advanced control and diagnosed the root-cause fault in 15 of 15 safety-critical situations. Three independently developed models spanning a thirty-fold cost range exceeded the baseline. In a fully auditable evaluation, these results demonstrate run-time abnormal situation management without a human in the loop.

View source

Similar papers

Open access Sep 2026

Hierarchical Temporal Runtime Assurance for Controlled Agentic AI Systems: Safety Shielding, Auditable Action Repair, and Bounded Recovery

Agentic artificial intelligence requires safeguards that remain effective across trajectories rather than only at individual decisions. This study introduces MARIS-TRA, a hierarchical temporal runtime assurance extension of Controlled Agentic AI Systems. A compact, auditable fragment of Signal Temporal Logic (STL) prov...

Tymoteusz Miller · 0 citations
Preprint Aug 2026

AeroCopilotBench: Safety-Gated Evaluation of LLM Agents on Aircraft Emergency Procedures in an Executable Cockpit

Aviation knowledge question answering cannot directly assess the operational effectiveness and safety compliance of large language models throughout aircraft emergency procedures. We introduce AeroCopilotBench and its executable cockpit environment, ACOE, which define state-transition rules, task goals, and trajectory-...

Yu-Chen Yuan, Zheng-Huang Wu, Yuan-Gan Li et al. · 0 citations
Open access 2026

Enhancing Industrial Fault Diagnosis: A Probabilistic Expert System With LLM-Augmented Validation

Keeping industrial utility equipment reliable is a central problem in industrial predictive maintenance, especially for assets whose faults are reflected in coupled pressure, temperature, flow, power, and efficiency signals. Pure data-driven models can learn complex patterns but are difficult to audit, while convention...

Yunqi Li, Xin Tang, Yin-Bo Dai et al. · 0 citations
Conference Aug 2026

PrincipiaBlastFoam: Knowledge-Guided Multi-Agent Automation for Safety-Critical Blast Simulation

Simulation workflows support system safety analysis, infrastructure protection, and risk-informed engineering decisions. In blast engineering, however, a case can run to completion while still violating physical assumptions through unsafe model choices, inconsistent boundary conditions, or broken cross-file dependencie...

Shi-Hao Lin, Pei-Jie Yang, Zi-Cheng Xu et al. · 0 citations
2026

Automated Discovery of Finite State Automata for the Control of Discrete Event Systems

In many industrial systems, control-related information is often opaque, undermining the ability to formally ensure safety, correctness, and performance. Representing such systems as Discrete Event Systems (DES) models enables rigorous analysis, synthesis, and correct-by-construction control, but modeling is typically...

Jhonnatan R. Semler, Rosaine F. Semler, M. Wehrmeister et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.