Skip to content
Preprint

CADEngBench: It Looks Like CAD, but Does It Work? Evaluating Parametric Design, Assembly Reasoning, and Physics Simulation

Aug 2026 · 3 citations · 32 references
Computer Science

TL;DR

The results show that CAD evaluation must test engineering behavior rather than appearance alone, while editing supplied CAD is substantially easier than generating it, while complex edits and matched FEA remain difficult.

Abstract

A CAD model is not engineering-grade merely because it looks correct. It must satisfy design requirements, respond predictably to parameter changes, support controlled edits, match a reference structural response under a declared analysis, and connect to other parts through valid joints. We present CADEngBench, a two-track benchmark for these capabilities. CADEngBench-P evaluates 300 parametric parts, each used for one zero-to-CAD task and one functional-editing task (600 tasks in total), through boundary-representation (B-Rep) validity, engineering and DFM checks, parameter-family perturbations, functional editing, and matched linear-static FEA in CalculiX. CADEngBench-A evaluates 150 body pairs through ranked joint retrieval, exact face-and-edge grounding, joint-frame prediction, and kinematic verification. Across eight multimodal, code-capable models, editing supplied CAD is substantially easier than generating it, while complex edits and matched FEA remain difficult. Assembly predictions often locate the relevant region but fail to recover the recorded joint or mating entities. These results show that CAD evaluation must test engineering behavior rather than appearance alone.

View source

Similar papers

#computer vision Preprint Sep 2026

RealCADBench: Benchmarking Parametric CAD Modeling from Industrial Design Intents

Parametric computer-aided design (CAD) modeling is difficult to evaluate with a single metric. Existing CAD benchmarks often emphasize synthetic or CAD-native settings, limited input modalities, or executability and IoUs alone. We introduce RealCADBench, a benchmark for intent-to-program CAD modeling from real industri...

Guan-Ling Li, Zhi-Chao Huang, Hui-Mu Yu et al. · 0 citations
Preprint Aug 2026

CADENA: Stepwise CAD Reverse Engineering

Computer-Aided Design (CAD) underpins modern engineering, yet converting existing shapes into editable models still demands substantial expert effort. Most AI systems emit the entire CAD program in a single pass, never inspecting the intermediate geometry. In contrast, human engineers build a part feature by feature, c...

S. Kabisov, Gennadiy Savrasov, Maksim Elistratov et al. · 0 citations
Review Open access Aug 2026

CAD-to-VR framework for interactive assembly simulation in product design review

The results indicate that CAD-driven and no-code VR authoring can make assembly DRs more repeatable and more readily integrated into industrial product-development workflows.

Fabio Grandi, Rocco De Ciantis, Tommaso Romoli et al. · 0 citations
Preprint Aug 2026

OmniMech: All-in-one Multimodal Mechanical Benchmark for 3D Reconstruction

OmniMech is introduced, the first million-scale benchmark for evaluating VLMs on executable CAD generation from industrial manufacturing data, and experiments show that current VLMs and CAD-specialized models still struggle with executable program synthesis, fine-grained 3D reconstruction, and reliable enforcement of d...

Tai-Ting Lu, Run-Ze Liu, Zi-Wei Dong et al. · 0 citations
Jul 2026

Nova3D: Code-Native Generation of Programmable 3D Assets

Current 3D generative models mostly produce a final surface: a visually strong but largely opaque mesh. Interactive 3D worlds need more than a surface. They need named parts, an assembly hierarchy, measurable constraints, local edit handles, and joints for articulation. We present Nova3D, a system that generates 3D ass...

Nimra Noor, Muhammad Bilal, A. Hussain et al. · 1 citation
Preprint Aug 2026

Task-Driven 3D Printability Assistance via Geometry- and Knowledge-Grounded LLM Reasoning

The proposed framework leverages the reasoning and language-understanding capabilities of large language models (LLMs), while grounding the reasoning with geometry evidence and structured material/printer knowledge to generate reliable pre-print recommendations.

Zhao-Da Du, Qiaojie Zheng, Xiaoli Zhang · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.