See or Recall: A Sanity Check for the Role of Vision in Solving Visualization Question Answer Tasks with Multimodal LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Zhimin, Miao, Haichao, Yan, Xinyuan, Pascucci, Valerio, Berger, Matthew, Liu, Shusen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Visualization Literacy of Multimodal Large Language Models: A Comparative Study
by: Li, Zhimin, et al.
Published: (2024)
by: Li, Zhimin, et al.
Published: (2024)
The Visualization JUDGE : Can Multimodal Foundation Models Guide Visualization Design Through Visual Perception?
by: Berger, Matthew, et al.
Published: (2024)
by: Berger, Matthew, et al.
Published: (2024)
An Evaluation-Centric Paradigm for Scientific Visualization Agents
by: Ai, Kuangshi, et al.
Published: (2025)
by: Ai, Kuangshi, et al.
Published: (2025)
ParaView-MCP: An Autonomous Visualization Agent with Direct Tool Use
by: Liu, Shusen, et al.
Published: (2025)
by: Liu, Shusen, et al.
Published: (2025)
Concept Lens: Visually Analyzing the Consistency of Semantic Manipulation in GANs
by: Jeong, Sangwon, et al.
Published: (2024)
by: Jeong, Sangwon, et al.
Published: (2024)
Exploring Interaction Paradigms for LLM Agents in Scientific Visualization
by: Vonderhorst, Jackson, et al.
Published: (2026)
by: Vonderhorst, Jackson, et al.
Published: (2026)
Toward AI VIS Co-Scientists: A General and End-to-End Agent Harness for Solving Complex Data Visualization Tasks
by: Miao, Haichao, et al.
Published: (2026)
by: Miao, Haichao, et al.
Published: (2026)
Seeing the Many: Exploring Parameter Distributions Conditioned on Features in Surrogates
by: Wang, Xiaohan, et al.
Published: (2025)
by: Wang, Xiaohan, et al.
Published: (2025)
VMC: A Grammar for Visualizing Statistical Model Checks
by: Guo, Ziyang, et al.
Published: (2024)
by: Guo, Ziyang, et al.
Published: (2024)
Guided Reality: Generating Visually-Enriched AR Task Guidance with LLMs and Vision Models
by: Zhao, Ada Yi, et al.
Published: (2025)
by: Zhao, Ada Yi, et al.
Published: (2025)
How Do LLMs See Charts? A Comparative Study on High-Level Visualization Comprehension in Humans and LLMs
by: Jeon, Hyotaek, et al.
Published: (2026)
by: Jeon, Hyotaek, et al.
Published: (2026)
Do Vision-Language Models See Visualizations Like Humans? Alignment in Chart Categorization
by: Gyarmati, Péter Ferenc, et al.
Published: (2025)
by: Gyarmati, Péter Ferenc, et al.
Published: (2025)
Addressing and Visualizing Misalignments in Human Task-Solving Trajectories
by: Kim, Sejin, et al.
Published: (2024)
by: Kim, Sejin, et al.
Published: (2024)
DeepSee: Multidimensional Visualizations of Seabed Ecosystems
by: Coscia, Adam, et al.
Published: (2024)
by: Coscia, Adam, et al.
Published: (2024)
Web-based Visualization and Analytics of Petascale data: Equity as a Tide that Lifts All Boats
by: Panta, Aashish, et al.
Published: (2024)
by: Panta, Aashish, et al.
Published: (2024)
ScaleTrotter: Illustrative Visual Travels Across Negative Scales
by: Halladjian, Sarkis, et al.
Published: (2019)
by: Halladjian, Sarkis, et al.
Published: (2019)
Exploring the Capability of LLMs in Performing Low-Level Visual Analytic Tasks on SVG Data Visualizations
by: Xu, Zhongzheng, et al.
Published: (2024)
by: Xu, Zhongzheng, et al.
Published: (2024)
Exploring The Impact Of Proactive Generative AI Agent Roles In Time-Sensitive Collaborative Problem-Solving Tasks
by: Mukhopadhyay, Anirban, et al.
Published: (2026)
by: Mukhopadhyay, Anirban, et al.
Published: (2026)
From Perception to Decision: Assessing the Role of Chart Types Affordances in High-Level Decision Tasks
by: Li, Yixuan, et al.
Published: (2024)
by: Li, Yixuan, et al.
Published: (2024)
Invisible Saboteurs: Sycophantic LLMs Mislead Novices in Problem-Solving Tasks
by: Bo, Jessica Y., et al.
Published: (2025)
by: Bo, Jessica Y., et al.
Published: (2025)
The Impact of Elicitation and Contrasting Narratives on Engagement, Recall and Attitude Change with News Articles Containing Data Visualization
by: Rogha, Milad, et al.
Published: (2024)
by: Rogha, Milad, et al.
Published: (2024)
Do You See What I See? A Qualitative Study Eliciting High-Level Visualization Comprehension
by: Quadri, Ghulam Jilani, et al.
Published: (2024)
by: Quadri, Ghulam Jilani, et al.
Published: (2024)
LLM-Generated Tips Rival Expert-Created Tips in Helping Students Answer Quantum-Computing Questions
by: Krupp, Lars, et al.
Published: (2024)
by: Krupp, Lars, et al.
Published: (2024)
Decomposed Prompting to Answer Questions on a Course Discussion Board
by: Jaipersaud, Brandon, et al.
Published: (2024)
by: Jaipersaud, Brandon, et al.
Published: (2024)
Simulating Vision Impairment in Virtual Reality -- A Comparison of Visual Task Performance with Real and Simulated Tunnel Vision
by: Neugebauer, Alexander, et al.
Published: (2023)
by: Neugebauer, Alexander, et al.
Published: (2023)
From Following to Understanding: Investigating the Role of Reflective Prompts in AR-Guided Tasks to Promote Task Understanding
by: Zhang, Nandi, et al.
Published: (2025)
by: Zhang, Nandi, et al.
Published: (2025)
Personagram: Bridging Personas and Product Design for Creative Ideation with Multimodal LLMs
by: Kim, Taewook, et al.
Published: (2026)
by: Kim, Taewook, et al.
Published: (2026)
Seeing is Believing: The Role of Scatterplots in Recommender System Trust and Decision-Making
by: Doppalapudi, Bhavana, et al.
Published: (2024)
by: Doppalapudi, Bhavana, et al.
Published: (2024)
Do MLLMs See What We See? Analyzing Visualization Literacy Barriers in AI Systems
by: Mengli, et al.
Published: (2026)
by: Mengli, et al.
Published: (2026)
Cognitive Prosthetic: An AI-Enabled Multimodal System for Episodic Recall in Knowledge Work
by: Obiuwevwi, Lawrence, et al.
Published: (2026)
by: Obiuwevwi, Lawrence, et al.
Published: (2026)
CPS-TaskForge: Generating Collaborative Problem Solving Environments for Diverse Communication Tasks
by: Haduong, Nikita, et al.
Published: (2024)
by: Haduong, Nikita, et al.
Published: (2024)
Scalable Climate Data Analysis: Balancing Petascale Fidelity and Computational Cost
by: Panta, Aashish, et al.
Published: (2025)
by: Panta, Aashish, et al.
Published: (2025)
Interactive Sketchpad: A Multimodal Tutoring System for Collaborative, Visual Problem-Solving
by: Chen, Steven-Shine, et al.
Published: (2025)
by: Chen, Steven-Shine, et al.
Published: (2025)
Exploring Multiscale Navigation of Homogeneous and Dense Objects with Progressive Refinement in Virtual Reality
by: Pavanatto, Leonardo, et al.
Published: (2025)
by: Pavanatto, Leonardo, et al.
Published: (2025)
Do LLMs Provide Consistent Answers to Health-Related Questions across Languages?
by: Schlicht, Ipek Baris, et al.
Published: (2025)
by: Schlicht, Ipek Baris, et al.
Published: (2025)
Evaluating judgment of spatial correlation in visual displays of scalar field distributions
by: Zhao, Yayan, et al.
Published: (2025)
by: Zhao, Yayan, et al.
Published: (2025)
From Checking to Sensemaking: A Caregiver-in-the-Loop Framework for AI-Assisted Task Verification in Dementia Care
by: Lai, Joy, et al.
Published: (2025)
by: Lai, Joy, et al.
Published: (2025)
When Should Users Check? Modeling Confirmation Frequency inMulti-Step Agentic AI Tasks
by: Zhou, Jieyu, et al.
Published: (2025)
by: Zhou, Jieyu, et al.
Published: (2025)
I See You: Teacher Analytics with GPT-4 Vision-Powered Observational Assessment
by: Lee, Unggi, et al.
Published: (2024)
by: Lee, Unggi, et al.
Published: (2024)
Are LLMs ready for Visualization?
by: Vázquez, Pere-Pau
Published: (2024)
by: Vázquez, Pere-Pau
Published: (2024)
Similar Items
-
Visualization Literacy of Multimodal Large Language Models: A Comparative Study
by: Li, Zhimin, et al.
Published: (2024) -
The Visualization JUDGE : Can Multimodal Foundation Models Guide Visualization Design Through Visual Perception?
by: Berger, Matthew, et al.
Published: (2024) -
An Evaluation-Centric Paradigm for Scientific Visualization Agents
by: Ai, Kuangshi, et al.
Published: (2025) -
ParaView-MCP: An Autonomous Visualization Agent with Direct Tool Use
by: Liu, Shusen, et al.
Published: (2025) -
Concept Lens: Visually Analyzing the Consistency of Semantic Manipulation in GANs
by: Jeong, Sangwon, et al.
Published: (2024)