Do Text-to-Vis Benchmarks Test Real Use of Visualisations?
Fuente:
arXiv
Saved in:
| Main Authors: | Nguyen, Hy, He, Xuefei, Reeson, Andrew, Paris, Cecile, Poon, Josiah, Kummerfeld, Jonathan K. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
by: Miranda-Pena, Clarissa, et al.
Published: (2026)
by: Miranda-Pena, Clarissa, et al.
Published: (2026)
Your Students Don't Use LLMs Like You Wish They Did
by: Kobler, Sebastian, et al.
Published: (2026)
by: Kobler, Sebastian, et al.
Published: (2026)
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
by: Tian, Yuan, et al.
Published: (2024)
by: Tian, Yuan, et al.
Published: (2024)
VisEval: A Benchmark for Data Visualization in the Era of Large Language Models
by: Chen, Nan, et al.
Published: (2024)
by: Chen, Nan, et al.
Published: (2024)
A Directed Graph Model and Experimental Framework for Design and Study of Time-Dependent Text Visualisation
by: Fan, Songhai, et al.
Published: (2026)
by: Fan, Songhai, et al.
Published: (2026)
An AI-Resilient Text Rendering Technique for Reading and Skimming Documents
by: Gu, Ziwei, et al.
Published: (2024)
by: Gu, Ziwei, et al.
Published: (2024)
ReasonGraph: Visualisation of Reasoning Paths
by: Li, Zongqian, et al.
Published: (2025)
by: Li, Zongqian, et al.
Published: (2025)
AI-Resilient Interfaces
by: Glassman, Elena L., et al.
Published: (2024)
by: Glassman, Elena L., et al.
Published: (2024)
Data Verbalisation: What is Text Doing in a Data Visualisation?
by: Murrell, Paul
Published: (2025)
by: Murrell, Paul
Published: (2025)
Towards Metrics for Evaluating Creativity in Visualisation Design
by: Owen, Aron E, et al.
Published: (2024)
by: Owen, Aron E, et al.
Published: (2024)
Bridging Dictionary: AI-Generated Dictionary of Partisan Language Use
by: Jiang, Hang, et al.
Published: (2024)
by: Jiang, Hang, et al.
Published: (2024)
"Newspaper Eat" Means "Not Tasty": A Taxonomy and Benchmark for Coded Language in Real-World Chinese Online Reviews
by: Wan, Ruyuan, et al.
Published: (2026)
by: Wan, Ruyuan, et al.
Published: (2026)
Benchmarking LLM Tool-Use in the Wild
by: Yu, Peijie, et al.
Published: (2026)
by: Yu, Peijie, et al.
Published: (2026)
Large Language Models Pass the Turing Test
by: Jones, Cameron R., et al.
Published: (2025)
by: Jones, Cameron R., et al.
Published: (2025)
ChatVis: Automating Scientific Visualization with a Large Language Model
by: Mallick, Tanwi, et al.
Published: (2024)
by: Mallick, Tanwi, et al.
Published: (2024)
MASCOT: Towards Multi-Agent Socio-Collaborative Companion Systems
by: Wang, Yiyang, et al.
Published: (2026)
by: Wang, Yiyang, et al.
Published: (2026)
Survey Insights on M365 Copilot Adoption
by: Bano, Muneera, et al.
Published: (2024)
by: Bano, Muneera, et al.
Published: (2024)
UXAgent: An LLM Agent-Based Usability Testing Framework for Web Design
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
Enhancing Tennis Training with Real-Time Swing Data Visualisation in Immersive Virtual Reality
by: Najami, Ryan, et al.
Published: (2025)
by: Najami, Ryan, et al.
Published: (2025)
UXAgent: A System for Simulating Usability Testing of Web Design with LLM Agents
by: Lu, Yuxuan, et al.
Published: (2025)
by: Lu, Yuxuan, et al.
Published: (2025)
K-QA: A Real-World Medical Q&A Benchmark
by: Manes, Itay, et al.
Published: (2024)
by: Manes, Itay, et al.
Published: (2024)
Emojinize: Enriching Any Text with Emoji Translations
by: Klein, Lars Henning, et al.
Published: (2024)
by: Klein, Lars Henning, et al.
Published: (2024)
Understanding Expert Exploration in EHR Visualization Tools: The ParcoursVis Use Case
by: Assor, Ambre, et al.
Published: (2025)
by: Assor, Ambre, et al.
Published: (2025)
Investigating User Perspectives on Differentially Private Text Privatization
by: Meisenbacher, Stephen, et al.
Published: (2025)
by: Meisenbacher, Stephen, et al.
Published: (2025)
"What Are You Really Trying to Do?": Co-Creating Life Goals from Everyday Computer Use
by: Sapkota, Shardul, et al.
Published: (2026)
by: Sapkota, Shardul, et al.
Published: (2026)
Do LLMs Truly Benefit from Longer Context in Automatic Post-Editing?
by: Kim, Ahrii, et al.
Published: (2026)
by: Kim, Ahrii, et al.
Published: (2026)
Exploring the Efficacy of Large Language Models in Summarizing Mental Health Counseling Sessions: A Benchmark Study
by: Adhikary, Prottay Kumar, et al.
Published: (2024)
by: Adhikary, Prottay Kumar, et al.
Published: (2024)
UI-Evol: Automatic Knowledge Evolving for Computer Use Agents
by: Zhang, Ziyun, et al.
Published: (2025)
by: Zhang, Ziyun, et al.
Published: (2025)
Navigating Rifts in Human-LLM Grounding: Study and Benchmark
by: Shaikh, Omar, et al.
Published: (2025)
by: Shaikh, Omar, et al.
Published: (2025)
LLM-Augmented Semantic Steering of Text Embedding Projection Spaces
by: Liu, Wei, et al.
Published: (2026)
by: Liu, Wei, et al.
Published: (2026)
Robust Multilingual Text-to-Pictogram Mapping for Scalable Reading Rehabilitation
by: Jhilal, Soufiane, et al.
Published: (2026)
by: Jhilal, Soufiane, et al.
Published: (2026)
Personalized Real-time Jargon Support for Online Meetings
by: Song, Yifan, et al.
Published: (2025)
by: Song, Yifan, et al.
Published: (2025)
Advancing STT for Low-Resource Real-World Speech
by: D'Intino, Flavio, et al.
Published: (2025)
by: D'Intino, Flavio, et al.
Published: (2025)
HoT: Highlighted Chain of Thought for Referencing Supporting Facts from Inputs
by: Nguyen, Tin, et al.
Published: (2025)
by: Nguyen, Tin, et al.
Published: (2025)
Interactive Concept Learning for Uncovering Latent Themes in Large Text Collections
by: Pacheco, Maria Leonor, et al.
Published: (2023)
by: Pacheco, Maria Leonor, et al.
Published: (2023)
Digital Comprehensibility Assessment of Simplified Texts among Persons with Intellectual Disabilities
by: Säuberli, Andreas, et al.
Published: (2024)
by: Säuberli, Andreas, et al.
Published: (2024)
Virtual Agents for Alcohol Use Counseling: Exploring LLM-Powered Motivational Interviewing
by: Steenstra, Ian, et al.
Published: (2024)
by: Steenstra, Ian, et al.
Published: (2024)
Chart-to-Experience: Benchmarking Multimodal LLMs for Predicting Experiential Impact of Charts
by: Kim, Seon Gyeom, et al.
Published: (2025)
by: Kim, Seon Gyeom, et al.
Published: (2025)
Beyond Turn-taking: Introducing Text-based Overlap into Human-LLM Interactions
by: Kim, JiWoo, et al.
Published: (2025)
by: Kim, JiWoo, et al.
Published: (2025)
Granuscore: A Reference-Free Measure of Granularity for Text Analysis and Question Answering
by: Ellinger, Lukas, et al.
Published: (2026)
by: Ellinger, Lukas, et al.
Published: (2026)
Similar Items
-
An Empirical Analysis of Static Analysis Methods for Detection and Mitigation of Code Library Hallucinations
by: Miranda-Pena, Clarissa, et al.
Published: (2026) -
Your Students Don't Use LLMs Like You Wish They Did
by: Kobler, Sebastian, et al.
Published: (2026) -
SQLucid: Grounding Natural Language Database Queries with Interactive Explanations
by: Tian, Yuan, et al.
Published: (2024) -
VisEval: A Benchmark for Data Visualization in the Era of Large Language Models
by: Chen, Nan, et al.
Published: (2024) -
A Directed Graph Model and Experimental Framework for Design and Study of Time-Dependent Text Visualisation
by: Fan, Songhai, et al.
Published: (2026)