Skip to content
Review

Using large language models to generate executable BPMN models based on text descriptions: an overview of approaches, limitations, and validation methods

2026 · SOFT MEASUREMENTS AND COMPUTING · 0 citations

TL;DR

It is argued that edge (sequence flow) generation is the weakest link once nodes are fixed, and typical structural failure modes (dangling nodes, disconnects, gateway violations, etc.) and causes tied to autoregressive generation are summarized.

Abstract

The paper addresses the use of large language models (LLMs) to automatically generate executable business processes in BPMN 2.0 from unstructured natural-language descriptions, with deployment to process engines such as Camunda Platform in mind. The text-to-BPMN XML mapping task is stated and decomposed into subproblems: extracting activities, events, and gateways; recovering control flow and branching; ensuring valid sequence flows and conformance to the BPMN specification. We survey process representations (BPMN XML, JSON as an intermediate format, graph-based models) and LLM adaptation methods: prompt engineering, instruction tuning, and fine-tuning. We argue that edge (sequence flow) generation is the weakest link once nodes are fixed, and summarize typical structural failure modes (dangling nodes, disconnects, gateway violations, etc.) and causes tied to autoregressive generation. A staged pipeline is proposed—separate generation of node set V and edge set E followed by post-validation — together with a three-level validation scheme: syntactic (BPMN XSD), structural (graph invariants), and executable (Camunda deploy and run). The article outlines a feedback-enabled pipeline architecture and discusses applicability and limitations.

View source

Similar papers

Conference Jul 2026

Operationalizing Large Language Models for Automated Software Requirement Interpretation and Change Impact Analysis

In fast-evolving software systems, effective 'natural language requirements parsing' and downstream change effect analysis capability across a multitude of codes represents low-hanging-fruit in this regard. We present a structured framework to deploy Large Language Models (LLMs) for automating two essential software en...

Nithya Krishnan, Kumaran Ramanujam, Suresh Babu Narra et al. · 0 citations
#large language models Book Open access Oct 2026

Modular Meta-Languages for Structured Instructions: A Novel Approach for LLM Integration into Evolving Engineering Toolchains

Large Language Models are increasingly used to generate structured engineering artifacts, yet the instruction artifacts that govern this generation are rarely treated as modeling artifacts in their own right. They typically appear as monolithic prompt blocks, schemas, or informal examples. When tools, metamodels, APIs,...

Louis Burk, Alexander Fischer, Christoph Scharnagl et al. · 0 citations
Open access 2026

Models, Prompts, and Code! A Semi-Formal State Machine Language for Multi-Paradigmatic Software Development

Compared to both classical UML tooling and fully LLM-based generation, the approach offers stronger determinism, better traceability, lower cognitive modeling effort, and reduced computational cost, while retaining the flexibility to express complex action behavior in natural language where formal specification would b...

Oliver Engling, Felix Schwägerl, Thomas Buchmann · 0 citations
#large language models Book Open access Oct 2026

Towards LLM-Assisted Business Process Modeling in an Industrial Modeling Tool: An Experience Report

Business Process Modeling (BPM) is a key and complex activity in Software and Systems Engineering. As a consequence, any relevant assistance integrated into the BPM tooling can be highly beneficial. Historically, Natural Language Processing (NLP) techniques have already been used in this context. However, the recent de...

Andrey Sadovykh, Loay Chlih, Bilal Said et al. · 0 citations
Open access Aug 2026

Analyzing structural and semantic similarities between formal business process models using ChatGPT-5.1: a test report

The results show that the applied LLM can reliably detect structural and semantic differences between formal business process models using Business Process Model and Notation, while distinguishing them from acceptable variations, demonstrating strong potential for automated model validation.

Christian Bennoit, S. Zamani, Tobias Greff · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.