The framework couples a three-layernite-element verication system with a dual-node loop and injects feedback from an external physics-based verier into a closedrepair loop, indicating that compliance does not change detectably across the two backbone LLMs tested, indicating that it ishere attributed to the external verier rather than the model.
Abstract
Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving themill-suited to safety-critical tasks.Rather than trusting LLM self-correction,thisframework injects feedback from an external physics-based verier into a closedrepair loop.The framework couples a three-layernite-element verication systemwith a dual-node loop.Node 1 turns code violations into hard repair constraints,Node 2 turns a four-dimensional quality score into safety-rst soft constraints,and a retrieval-augmented code base makes every violation traceable to a clause.Overve structure types and 44 cases,code compliance rises from 56.8%to 98.6%and the composite score from 63.8 to 71.4(p<0.000001),using about 5.8%lessmaterial.Removing either node degrades performance,and compliance does notchange detectably across the two backbone LLMs tested,indicating that it ishere attributed to the external verier rather than the model.The framework,the 44-case benchmark and all experiment scripts are released as open source forreplicability.
Multi-agent systems built on large language models (LLMs) are increasingly deployed for complex tasks requiring autonomous planning, tool use, and inter-agent coordination. However, the non-deterministic nature of LLM outputs and the emergent behavior arising from agent interactions render traditional test oracles inef...
Gopalakrishnan Marimuthu· International Conference on...· 0 citations
This work introduces a unified multi-agent framework, MAGS, that generates executable programs with formal safety guarantees, using Dafny as a verification-aware intermediate representation where safety properties can be mechanically checked.
Albert Wu, N. Roberts, Tzu-Heng Huang et al.· 0 citations
This work proposes a stateful, multi-agent validation pipeline that eradicated cross-phase hallucinations and proves adversarial auditing enables LLMs to reliably synthesize zero-error MBSE architectures.
Aleksei Velsh, Nenad Petrovic, Alois Knoll· 0 citations
Constrained resource-task assignment(RTA) is a foundational problem in operations research and engineering decision support. Turning a natural-language assignment description into a correct optimization model and executable solver code requires expertise in both application semantics and mathematical programming. Large...
Bing-Xu Zhang, Tian-Le Pu, Long-Fei Zhang et al.· 2026 12th International Conf...· 0 citations
VISA is presented, a structured, symbol-based description protocol that specifies a model in eight interconnected tables---four at the agent level (Agent, Variable, Sensing, Internal Function) and four at the model level (Associated Data, Input/Output, Schedule, Validation)---under the principle of minimality with comp...
A multi-agent pipeline that couples an LLM with an open-source formal backend to repair RTL through counterexample-guided iteration and additionally reports a practical limitation of the Yosys bind directive relevant to the open-source formal verification community.
Hailey Tran· arXiv.org· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.