Raiven: LLM-Based Visualization Authoring via Domain-Specific Language Mediation

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Irger, Alexandra, Hugie, Ella, Guo, Minghao, Warchol, Simon, Moreland, Kenneth, Pugmire, David, Matusik, Wojciech, Pfister, Hanspeter
Natura: Preprint
Pubblicazione: 2026
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866911584933117952
author Irger, Alexandra
Hugie, Ella
Guo, Minghao
Warchol, Simon
Moreland, Kenneth
Pugmire, David
Matusik, Wojciech
Pfister, Hanspeter
author_facet Irger, Alexandra
Hugie, Ella
Guo, Minghao
Warchol, Simon
Moreland, Kenneth
Pugmire, David
Matusik, Wojciech
Pfister, Hanspeter
contents Visualization is central to scientific discovery, yet authoring tools remain split between information and scientific visualization, and expertise in one rarely transfers to the other. Large Language Model (LLM) based systems promise to bridge this gap through natural language, but current approaches generate code non-deterministically, with no guarantee of correctness and no protection against silent data fabrication. We present Raiven, a conversational system that mediates visualization authoring through a formally defined domain-specific language. RaivenDSL unifies scientific and information visualization in a single representation spanning 2D, 3D, and tabular data. The LLM produces a compact RaivenDSL specification under schema-guided constraints, and a deterministic compiler translates it to executable D3 or VTK.js code. Because the LLM operates only on dataset metadata, outputs are deterministic, specifications are verifiable before execution, and data fabrication is impossible by construction. In a 100-task benchmark, Raiven achieves 100% compilation, is up to six times faster and six times cheaper than state-of-the-art LLMs, while improving interaction quality, correctness, and data faithfulness. An expert user study shows that Raiven significantly reduces debugging effort and makes it easier to produce correct visualizations.
format Preprint
id arxiv_https___arxiv_org_abs_2604_10008
institution arXiv
publishDate 2026
record_format arxiv
spellingShingle Raiven: LLM-Based Visualization Authoring via Domain-Specific Language Mediation
Irger, Alexandra
Hugie, Ella
Guo, Minghao
Warchol, Simon
Moreland, Kenneth
Pugmire, David
Matusik, Wojciech
Pfister, Hanspeter
Human-Computer Interaction
Visualization is central to scientific discovery, yet authoring tools remain split between information and scientific visualization, and expertise in one rarely transfers to the other. Large Language Model (LLM) based systems promise to bridge this gap through natural language, but current approaches generate code non-deterministically, with no guarantee of correctness and no protection against silent data fabrication. We present Raiven, a conversational system that mediates visualization authoring through a formally defined domain-specific language. RaivenDSL unifies scientific and information visualization in a single representation spanning 2D, 3D, and tabular data. The LLM produces a compact RaivenDSL specification under schema-guided constraints, and a deterministic compiler translates it to executable D3 or VTK.js code. Because the LLM operates only on dataset metadata, outputs are deterministic, specifications are verifiable before execution, and data fabrication is impossible by construction. In a 100-task benchmark, Raiven achieves 100% compilation, is up to six times faster and six times cheaper than state-of-the-art LLMs, while improving interaction quality, correctness, and data faithfulness. An expert user study shows that Raiven significantly reduces debugging effort and makes it easier to produce correct visualizations.
title Raiven: LLM-Based Visualization Authoring via Domain-Specific Language Mediation
topic Human-Computer Interaction
url https://arxiv.org/abs/2604.10008