Skip to content
Conference

An Agentic LLM-based Architecture for Automated Anomaly Detection and Adaptive Remediation

Jul 2026 · International Conference on Control, Decision and Information Technologies · pp. 532-537 · 0 citations · 14 references

Abstract

Traditional cloud monitoring often relies on static remediation procedures that are difficult to adapt to dynamic and heterogeneous infrastructures. This paper proposes an agentic LLM-based architecture for adaptive remediation planning from confirmed cloud anomalies. The goal is not to replace anomaly detectors, but to transform confirmed anomaly events into structured, policy-constrained remediation artifacts suitable for human-supervised operational workflows.The proposed workflow combines anomaly intake, contextual validation, playbook retrieval, and constrained playbook generation through a message-driven Multi-Agent System. Retrieval-Augmented Generation is used to correlate current incidents with historical knowledge and existing procedures, while deterministic guardrails enforce schema validation, policy constraints, command allow/deny lists, critical-resource checks, and human approval for high-impact or previously unseen actions.A containerised Proof of Concept demonstrates that confirmed anomalies can be transformed into CACAO-compatible remediation drafts within operationally reasonable time bounds. The evaluation focuses on generating and validating remediation plans under explicit operational constraints, rather than on anomaly detection benchmarking or on production-scale autonomous execution.

View source

Similar papers

Preprint Aug 2026

AgenticTwin: An Agentic LLM Framework Integrated with Digital Twin for Anomaly Detection

AgenticTwin is proposed, an agentic framework that integrates LLM-driven reasoning with a digital twin-based anomaly detection pipeline and enables human operators to ask relevant natural-language questions about the system.

Touseef Hasan, Mounika Ghanta, Souvika Sarkar et al. · 0 citations
Review Open access Aug 2026

From Detection to Verified Action: Operational Readiness for AI-Enabled Cloud Failure Management

The proposed Operational Decision-Readiness and Verification framework represents a deployment claim through analytical capability, operational authority, and assurance maturity and is conceptual and requires prospective and inter-rater validation.

Adepegba Akindayomi Akintade · 0 citations
Open access Sep 2026

From DevOps to XOps: an agent-driven reference architecture for autonomous enterprise operations

XOps is proposed, a five-layer reference architecture integrating PlatformOps, DataOps, MLOps and AIOps beneath an Agentic Orchestration layer with Policy-as-Code governance, together with a continuous-time Markov chain model quantifying the availability effect of agent-driven remediation.

Mete Köse, E. Küçüksille · 0 citations
Preprint Aug 2026

CyberLLM: A Multi-Agent LLM Framework for Autonomous Detection and Guarded Response in Automotive Cybersecurity

CyberLLM is presented, a multi-agent, LLM-orchestrated framework that autonomously detects vulnerabilities and executes remediations under a formal, runtime safety guard, and indicates that LLM agents can perform useful autonomous cyber-defense when wrapped in a deterministic, auditable safety envelope.

Nenad Petrovic, Oussama Jeddou, Feres Ben Fraj et al. · 0 citations
Review Open access 2026

A Multi-Agent DevSecOps Framework for Intelligent Vulnerability Detection and Auto-Remediation

This paper presents a multi-agent DevSecOps framework that integrates static code scanning, large language model (LLM) based security reasoning, automated repair generation, policy-as-code enforcement, and runtime monitoring into a unified event-driven pipeline. Five specialized agents collaborate through LangGraph sha...

Hai-Ning Fan, Li-Jie Zheng, Chen-Hao Han et al. · 0 citations
#artificial intelligence Review Aug 2026

Forward-Deployed Full-Stack Engineering for Autonomous Cloud MLOps

This work presents an evidence-gated multi-agent framework for transforming a natural-language MLOps cloud engineering task into a verified repository and operational cloud deployment and results show that the framework prevents unsupported lifecycle transitions and drives each run toward either a verified operational...

Sagar Srinivas Sakhinana, Venkataramana Runkana · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.