Skip to content

Beyond Static Charts: Can Language and Vision Language Models Generate Interactive Data Visualization Interfaces?

Aug 2026 · 0 citations · 45 references
Computer Science

TL;DR

A structured multi stage interface generation framework that decomposes the task into visualization design representation, generation of multiple interface candidates, constraint-aware critique, and self-refinement is proposed, demonstrating a practical path toward more reliable language-driven interactive visualization systems.

Abstract

Data visualization is central to analytical reasoning, but real-world analysis increasingly requires language-driven interactive interfaces rather than static charts. Although recent large language and vision language models (LLMs/VLMs) have shown promise in generating static charts from natural language, their ability to generate interactive data visualization interfaces remains largely unexplored due to the lack of benchmarks. We introduce VIS-GEN, a benchmark for evaluating how well LLMs/VLMs can generate interactive visualization interfaces from natural language queries. VIS-GEN comprises 3,042 samples covering diverse analytical intents, including data filtering, temporal analysis, and visualization editing, each paired with dataset metadata and natural language queries that are designed to reflect realistic, goal driven data exploration scenarios. We benchmark 14 state-of-the-art open-source and closed-source LLMs/VLMs, revealing large performance gaps and frequent failures on queries involving implicit intent, multiple interaction alternatives, and complex editing operations, highlighting interactive interface generation as a key open challenge beyond static chart synthesis. To address this, we propose a structured multi stage interface generation framework that decomposes the task into visualization design representation, generation of multiple interface candidates, constraint-aware critique, and self-refinement. This approach improves the best models pass rate by 15.9 percentage points, demonstrating a practical path toward more reliable language-driven interactive visualization systems. We release VIS-GEN at https://github.com/vis-nlp/VIS-GEN.

View source

Similar papers

Preprint Aug 2026

VisEditBench: Can Vision-Language Models Edit Visualization Code from Multimodal Feedback?

This work introduces VisEditBench, a benchmark of 1,395 human-annotated visualization code-editing tasks grounded in realistic visualization workflows and failure cases, and proposes VisEditAgent, a render-grounded editing framework that iteratively generates, executes, validates, and refines candidate edits.

Mizanur Rahman, Arshia Azimlu, Shadikur Rahman et al. · 3 citations · ⚡1
Sep 2026

Toward a Grammar-Based Foundation of Visual Analytics

Visual analytics has produced numerous innovative solutions for helping humans gain insight from complex data, yet fundamental questions remain unanswered, including how interactive visualizations support insight generation. We propose grammars as a framework for reasoning about complex, often under-specified concepts...

A. Scott, Man-Ling Yang, Ashley Suh et al. · 0 citations
Preprint Aug 2026

QUARTZ: Qualitative Understanding via Accessible Representation and Visualization

Qualitative data visualizations -- concept maps, network graphs, Sankey diagrams, and coding stripes -- are integral to research practice, yet remain entirely inaccessible to blind and low-vision (BLV) researchers. While visualization has seen advanced multimodal solutions for quantitative charts, qualitative visualiza...

Omar Khan, JooYoung Seo · 0 citations
#artificial intelligence Open access Sep 2026

Vibe Analysis: Exploring LLM Adoption by Data Visualization Practitioners.

These findings show that Vis designers actively use LLMs for both creative and technical aspects of the visualization process, and opens up opportunities for research combining LLM-mediated work with Vis tools that incorporate data visualization guidance, constraints, and best practices.

S. C. Spivak, Aditi Krishna, Mahsan Nourani et al. · 0 citations

Related blog posts

MIT News · Artificial Intelligence Sep 24, 2026

Estimating suicide risk from text

A new language-processing tool could help identify the highest-risk individuals from natural language, enabling swifter interventions.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.