Skip to content
Preprint

Verication-driven closed-loop multi-agent large language modelframework for code-compliant structural design

Aug 2026 · 0 citations · 42 references
Computer Science

TL;DR

The framework couples a three-layernite-element verication system with a dual-node loop and injects feedback from an external physics-based verier into a closedrepair loop, indicating that compliance does not change detectably across the two backbone LLMs tested, indicating that it ishere attributed to the external verier rather than the model.

Abstract

Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving themill-suited to safety-critical tasks.Rather than trusting LLM self-correction,thisframework injects feedback from an external physics-based verier into a closedrepair loop.The framework couples a three-layernite-element verication systemwith a dual-node loop.Node 1 turns code violations into hard repair constraints,Node 2 turns a four-dimensional quality score into safety-rst soft constraints,and a retrieval-augmented code base makes every violation traceable to a clause.Overve structure types and 44 cases,code compliance rises from 56.8%to 98.6%and the composite score from 63.8 to 71.4(p<0.000001),using about 5.8%lessmaterial.Removing either node degrades performance,and compliance does notchange detectably across the two backbone LLMs tested,indicating that it ishere attributed to the external verier rather than the model.The framework,the 44-case benchmark and all experiment scripts are released as open source forreplicability.

View source

Similar papers

Conference Jul 2026

Metamorphic Testing of Multi-Agent LLM Systems: A Trace-Based Behavioral Oracle Framework

Multi-agent systems built on large language models (LLMs) are increasingly deployed for complex tasks requiring autonomous planning, tool use, and inter-agent coordination. However, the non-deterministic nature of LLM outputs and the emergent behavior arising from agent interactions render traditional test oracles inef...

Gopalakrishnan Marimuthu · 0 citations
Conference Aug 2026

Dual-Side Verification for Trustworthy LLM-Based Constrained Resource-Task Assignment Modeling

Constrained resource-task assignment(RTA) is a foundational problem in operations research and engineering decision support. Turning a natural-language assignment description into a correct optimization model and executable solver code requires expertise in both application semantics and mathematical programming. Large...

Bing-Xu Zhang, Tian-Le Pu, Long-Fei Zhang et al. · 0 citations
Jul 2026

VISA: A Structured Description Protocol for Agent-Based Simulation Models Towards Machine Reproducibility

VISA is presented, a structured, symbol-based description protocol that specifies a model in eight interconnected tables---four at the agent level (Agent, Variable, Sensing, Internal Function) and four at the model level (Associated Data, Input/Output, Schedule, Validation)---under the principle of minimality with comp...

Zhou He · 0 citations
Jul 2026

Open-Source LLM-Driven Formal Verification: A Multi-Agent Pipeline for RTL Repair

A multi-agent pipeline that couples an LLM with an open-source formal backend to repair RTL through counterexample-guided iteration and additionally reports a practical limitation of the Yosys bind directive relevant to the open-source formal verification community.

Hailey Tran · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.