Skip to content
Book Open access

Natural-Language to Geometry Diagrams: A Constraint-Based Pipeline for Precise Visual Reasoning

Jul 2026 · SIGGRAPH Posters · pp. 1-2 · 0 citations · 1 references
Computer Science

TL;DR

GenGX is a system that generates precise geometric diagrams from natural-language descriptions by combining large language model (LLM) interpretation with symbolic constraint solving by combining large language model (LLM) interpretation with symbolic constraint solving.

Abstract

We present GenGX, a system that generates precise geometric diagrams from natural-language descriptions by combining large language model (LLM) interpretation with symbolic constraint solving. User prompts are translated by an LLM autoformalizer into a structured intermediate representation (IR) encoding geometric entities, relationships, and construction semantics. The IR is passed to CoreGX, a constraint solver that synthesizes a deterministic construction sequence — operating above classical Euclidean primitives — that provably realizes the specified figure without numerical optimization. The system handles classical constructions, conics, curves, and transformations, and resolves both discrete and continuous ambiguity through explicit IR specifiers and a numeric clarity optimizer that selects a visually canonical representative from any underdetermined family of valid diagrams. This hybrid architecture avoids the spatial inaccuracies endemic to purely generative text-to-image approaches, produces reproducible results, and allows users to inspect and correct the IR directly.

Read PDF

Similar papers

#artificial intelligence Preprint Sep 2026

From Symbolic Perception to Logical Deduction: A Framework for Guiding Language Models in Geometric Reasoning

Plane geometry remains a significant challenge in AI, requiring the integration of visual perception and mathematical reasoning. While Large Multimodal Models (LMMs) naturally handle visuo-linguistic inputs, they are often computationally intensive and opaque. We demonstrate that a pure Large Language Model (LLM), when...

Wei-Chen Dai, Rafael Cabral, Ziyi Shou et al. · 0 citations
Conference Aug 2026

LVR-Draw: A Language-Vision Pipeline for Robotic Drawing with Interactive Scene Verification and Correction

This paper presents LVR-Draw, a fully local language–vision pipeline for robotic drawing that integrates structured scene generation, multimodal verification and correction, and deterministic execution within a unified Human–AI–Robot loop. Given a natural language prompt, a Large Language Model (LLM) generates a struct...

Nuttasorn Aiemsetthee, Renke Liu, Kave Salamatian et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Beyond Surface Forms: Symbolic Edits as a Test for Logical Reasoning with LLMs

Logical reasoning with large language models (LLMs) is a critical capability, as it reflects a system's ability to correctly deduce hypotheses from a given context using faithful deductive processes. However, LLM reasoning has often been shown to be sensitive to small surface-level variations in problem formulation, ra...

Ramya Keerthy Thatikonda, W. Buntine, Ehsan Shareghi · 0 citations
#artificial intelligence Preprint Sep 2026

From Terminology to Diagrams: Visual-Instruction Generation for Scientific Diagram Understanding

A framework for generating large-scale diagram-grounded instruction data by leveraging terminology derived from scientific curricula is introduced, and augmenting existing models such as LLaVA OneVision with SciGram establishes new state-of-the-art performance on diagram question answering.

Raúl Ortega, José Manuél Gómez-Pérez · 0 citations
Preprint Aug 2026

ExpConCAD: Experience-Guided Text-to-CAD Generation from Shape Descriptions with Implicit Spatial Constraints

It is argued that missing spatial constraints should be inferred with respect to the underlying construction structure and informed by reusable design experience and informed by reusable design experience in ExpConCAD, an experience-enhanced framework for implicit spatial constraint completion.

Jingyao Liu, Jin Tang, Chen Huang et al. · 0 citations
Conference Open access Sep 2026

GeoMind: Explicit Spatial Reasoning via Dual-Reference Geometric Modeling

GeoMind, a model-then-reason framework that employs a single LLM to autoregressively generate an explicit Geometric Description Language (GDL) map, serving as a grounded context to derive the final answer, suggests that explicit geometric grounding enables robust spatial reasoning without human annotation.

Xing Wei, Ao-Xiang Tian, Shaofan Liu et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.