CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Verma, Arnav, Mukherjee, Kushin, Potts, Christopher, Kreiss, Elisa, Fan, Judith E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Measuring and predicting variation in the difficulty of questions about data visualizations
by: Verma, Arnav, et al.
Published: (2025)
by: Verma, Arnav, et al.
Published: (2025)
Affective Color Scales for Colormap Data Visualizations
by: Braun, Halle C., et al.
Published: (2025)
by: Braun, Halle C., et al.
Published: (2025)
Presenting Large Language Models as Companions Affects What Mental Capacities People Attribute to Them
by: Chen, Allison, et al.
Published: (2025)
by: Chen, Allison, et al.
Published: (2025)
AI-enhanced semantic feature norms for 786 concepts
by: Suresh, Siddharth, et al.
Published: (2025)
by: Suresh, Siddharth, et al.
Published: (2025)
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
by: Mena, Omar, et al.
Published: (2025)
by: Mena, Omar, et al.
Published: (2025)
Large Language Models estimate fine-grained human color-concept associations
by: Mukherjee, Kushin, et al.
Published: (2024)
by: Mukherjee, Kushin, et al.
Published: (2024)
PhysicsSolutionAgent: Towards Multimodal Explanations for Numerical Physics Problem Solving
by: Thole, Aditya, et al.
Published: (2026)
by: Thole, Aditya, et al.
Published: (2026)
HEDS 3.0: The Human Evaluation Data Sheet Version 3.0
by: Belz, Anya, et al.
Published: (2024)
by: Belz, Anya, et al.
Published: (2024)
VisEval: A Benchmark for Data Visualization in the Era of Large Language Models
by: Chen, Nan, et al.
Published: (2024)
by: Chen, Nan, et al.
Published: (2024)
Conversational DNA: A New Visual Language for Understanding Dialogue Structure in Human and AI
by: Lin, Baihan
Published: (2025)
by: Lin, Baihan
Published: (2025)
Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations
by: Liang, Chen, et al.
Published: (2026)
by: Liang, Chen, et al.
Published: (2026)
IMBUE: Improving Interpersonal Effectiveness through Simulation and Just-in-time Feedback with Human-Language Model Interaction
by: Lin, Inna Wanyin, et al.
Published: (2024)
by: Lin, Inna Wanyin, et al.
Published: (2024)
Evaluating LLM-Generated Q&A Test: a Student-Centered Study
by: Wróblewska, Anna, et al.
Published: (2025)
by: Wróblewska, Anna, et al.
Published: (2025)
Human-AI Narrative Synthesis to Foster Shared Understanding in Civic Decision-Making
by: Overney, Cassandra, et al.
Published: (2025)
by: Overney, Cassandra, et al.
Published: (2025)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
by: Vo, Truong, et al.
Published: (2025)
by: Vo, Truong, et al.
Published: (2025)
Pearmut: Human Evaluation of Translation Made Trivial
by: Zouhar, Vilém, et al.
Published: (2026)
by: Zouhar, Vilém, et al.
Published: (2026)
Are Humans as Brittle as Large Language Models?
by: Li, Jiahui, et al.
Published: (2025)
by: Li, Jiahui, et al.
Published: (2025)
Low-Resource Counterspeech Generation for Indic Languages: The Case of Bengali and Hindi
by: Das, Mithun, et al.
Published: (2024)
by: Das, Mithun, et al.
Published: (2024)
Human-AI Interaction Alignment: Designing, Evaluating, and Evolving Value-Centered AI For Reciprocal Human-AI Futures
by: Shen, Hua, et al.
Published: (2025)
by: Shen, Hua, et al.
Published: (2025)
Context-Aware Monolingual Human Evaluation of Machine Translation
by: Picinini, Silvio, et al.
Published: (2025)
by: Picinini, Silvio, et al.
Published: (2025)
DaKultur: Evaluating the Cultural Awareness of Language Models for Danish with Native Speakers
by: Müller-Eberstein, Max, et al.
Published: (2025)
by: Müller-Eberstein, Max, et al.
Published: (2025)
Toward a Human-Centered Evaluation Framework for Trustworthy LLM-Powered GUI Agents
by: Chen, Chaoran, et al.
Published: (2025)
by: Chen, Chaoran, et al.
Published: (2025)
PALLM: Evaluating and Enhancing PALLiative Care Conversations with Large Language Models
by: Wang, Zhiyuan, et al.
Published: (2024)
by: Wang, Zhiyuan, et al.
Published: (2024)
Large Language Models for Virtual Human Gesture Selection
by: Torshizi, Parisa Ghanad, et al.
Published: (2025)
by: Torshizi, Parisa Ghanad, et al.
Published: (2025)
Practicing with Language Models Cultivates Human Empathic Communication
by: Kumar, Aakriti, et al.
Published: (2026)
by: Kumar, Aakriti, et al.
Published: (2026)
VeriLA: A Human-Centered Evaluation Framework for Interpretable Verification of LLM Agent Failures
by: Sung, Yoo Yeon, et al.
Published: (2025)
by: Sung, Yoo Yeon, et al.
Published: (2025)
Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations
by: He, Keyu, et al.
Published: (2025)
by: He, Keyu, et al.
Published: (2025)
CloChat: Understanding How People Customize, Interact, and Experience Personas in Large Language Models
by: Ha, Juhye, et al.
Published: (2024)
by: Ha, Juhye, et al.
Published: (2024)
Human-Centered AI in Multidisciplinary Medical Discussions: Evaluating the Feasibility of a Chat-Based Approach to Case Assessment
by: Sawano, Shinnosuke, et al.
Published: (2025)
by: Sawano, Shinnosuke, et al.
Published: (2025)
Large Language Model-based Human-Agent Collaboration for Complex Task Solving
by: Feng, Xueyang, et al.
Published: (2024)
by: Feng, Xueyang, et al.
Published: (2024)
VeriLLMed: Interactive Visual Debugging of Medical Large Language Models with Knowledge Graphs
by: Xiang, Yurui, et al.
Published: (2026)
by: Xiang, Yurui, et al.
Published: (2026)
Sniff AI: Is My 'Spicy' Your 'Spicy'? Exploring LLM's Perceptual Alignment with Human Smell Experiences
by: Zhong, Shu, et al.
Published: (2024)
by: Zhong, Shu, et al.
Published: (2024)
Human Capital Visualization using Speech Amount during Meetings
by: Hashimoto, Ekai, et al.
Published: (2025)
by: Hashimoto, Ekai, et al.
Published: (2025)
From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction
by: Satish, Shree Harsha Bokkahalli, et al.
Published: (2026)
by: Satish, Shree Harsha Bokkahalli, et al.
Published: (2026)
An Evaluation-Centric Paradigm for Scientific Visualization Agents
by: Ai, Kuangshi, et al.
Published: (2025)
by: Ai, Kuangshi, et al.
Published: (2025)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
by: Joshi, Brihi, et al.
Published: (2025)
by: Joshi, Brihi, et al.
Published: (2025)
DICE: A Framework for Dimensional and Contextual Evaluation of Language Models
by: Shrivastava, Aryan, et al.
Published: (2025)
by: Shrivastava, Aryan, et al.
Published: (2025)
PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users
by: Kirk, Hannah Rose, et al.
Published: (2026)
by: Kirk, Hannah Rose, et al.
Published: (2026)
SymbolicThought: Integrating Language Models and Symbolic Reasoning for Consistent and Interpretable Human Relationship Understanding
by: Zhao, Runcong, et al.
Published: (2025)
by: Zhao, Runcong, et al.
Published: (2025)
Large Language Models Help Humans Verify Truthfulness -- Except When They Are Convincingly Wrong
by: Si, Chenglei, et al.
Published: (2023)
by: Si, Chenglei, et al.
Published: (2023)
Similar Items
-
Measuring and predicting variation in the difficulty of questions about data visualizations
by: Verma, Arnav, et al.
Published: (2025) -
Affective Color Scales for Colormap Data Visualizations
by: Braun, Halle C., et al.
Published: (2025) -
Presenting Large Language Models as Companions Affects What Mental Capacities People Attribute to Them
by: Chen, Allison, et al.
Published: (2025) -
AI-enhanced semantic feature norms for 786 concepts
by: Suresh, Siddharth, et al.
Published: (2025) -
Augmenting a Large Language Model with a Combination of Text and Visual Data for Conversational Visualization of Global Geospatial Data
by: Mena, Omar, et al.
Published: (2025)