SimVecVis: A Dataset for Enhancing MLLMs in Visualization Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Can, Da, Chunlin, Long, Xiaoxiao, Yang, Yuxiao, Zhang, Yu, Wang, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision
by: Li, Chentao, et al.
Published: (2026)
by: Li, Chentao, et al.
Published: (2026)
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
MP-GUI: Modality Perception with MLLMs for GUI Understanding
by: Wang, Ziwei, et al.
Published: (2025)
by: Wang, Ziwei, et al.
Published: (2025)
ReVis: Towards Reusable Image-Based Visualizations with MLLMs
by: Wen, Xiaolin, et al.
Published: (2026)
by: Wen, Xiaolin, et al.
Published: (2026)
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning
by: Yang, Jiaxi, et al.
Published: (2025)
by: Yang, Jiaxi, et al.
Published: (2025)
SpatialViz-Bench: A Cognitively-Grounded Benchmark for Diagnosing Spatial Visualization in MLLMs
by: Wang, Siting, et al.
Published: (2025)
by: Wang, Siting, et al.
Published: (2025)
Panda or not Panda? Understanding Adversarial Attacks with Interactive Visualization
by: You, Yuzhe, et al.
Published: (2023)
by: You, Yuzhe, et al.
Published: (2023)
Collection Space Navigator: An Interactive Visualization Interface for Multidimensional Datasets
by: Ohm, Tillmann, et al.
Published: (2023)
by: Ohm, Tillmann, et al.
Published: (2023)
Beyond Object Categories: Multi-Attribute Reference Understanding for Visual Grounding
by: Guo, Hao, et al.
Published: (2025)
by: Guo, Hao, et al.
Published: (2025)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
by: Chen, Hongzhou, et al.
Published: (2024)
by: Chen, Hongzhou, et al.
Published: (2024)
Enhancing Saliency Prediction in Monitoring Tasks: The Role of Visual Highlights
by: Wu, Zekun, et al.
Published: (2024)
by: Wu, Zekun, et al.
Published: (2024)
The Visual Experience Dataset: Over 200 Recorded Hours of Integrated Eye Movement, Odometry, and Egocentric Video
by: Greene, Michelle R., et al.
Published: (2024)
by: Greene, Michelle R., et al.
Published: (2024)
SuDA: Support-based Domain Adaptation for Sim2Real Motion Capture with Flexible Sensors
by: Fang, Jiawei, et al.
Published: (2024)
by: Fang, Jiawei, et al.
Published: (2024)
RISEE: A Highly Interactive Naturalistic Driving Trajectories Dataset with Human Subjective Risk Perception and Eye-tracking Information
by: Wu, Xinzheng, et al.
Published: (2025)
by: Wu, Xinzheng, et al.
Published: (2025)
Self-Supervised Continuous Colormap Recovery from a 2D Scalar Field Visualization without a Legend
by: Liu, Hongxu, et al.
Published: (2025)
by: Liu, Hongxu, et al.
Published: (2025)
When, Where, and What? A Novel Benchmark for Accident Anticipation and Localization with Large Language Models
by: Liao, Haicheng, et al.
Published: (2024)
by: Liao, Haicheng, et al.
Published: (2024)
GenColor: Generative Color-Concept Association in Visual Design
by: Hou, Yihan, et al.
Published: (2025)
by: Hou, Yihan, et al.
Published: (2025)
Visually Grounded Narratives: Reducing Cognitive Burden in Researcher-Participant Interaction
by: Wu, Runtong, et al.
Published: (2025)
by: Wu, Runtong, et al.
Published: (2025)
DeepSORT-Driven Visual Tracking Approach for Gesture Recognition in Interactive Systems
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
AttributionScanner: A Visual Analytics System for Model Validation with Metadata-Free Slice Finding
by: Xuan, Xiwei, et al.
Published: (2024)
by: Xuan, Xiwei, et al.
Published: (2024)
Motion Sickness Modeling with Visual Vertical Estimation and Its Application to Autonomous Personal Mobility Vehicles
by: Liu, Hailong, et al.
Published: (2022)
by: Liu, Hailong, et al.
Published: (2022)
EduGage: Methods and Dataset for Sensor-Based Momentary Assessment of Engagement in Self-Guided Video Learning
by: Leng, Zikang, et al.
Published: (2026)
by: Leng, Zikang, et al.
Published: (2026)
CoEditor++: Instruction-based Visual Editing via Cognitive Reasoning
by: Ni, Minheng, et al.
Published: (2026)
by: Ni, Minheng, et al.
Published: (2026)
CHART-6: Human-Centered Evaluation of Data Visualization Understanding in Vision-Language Models
by: Verma, Arnav, et al.
Published: (2025)
by: Verma, Arnav, et al.
Published: (2025)
Tell Me Without Telling Me: Two-Way Prediction of Visualization Literacy and Visual Attention
by: Chang, Minsuk, et al.
Published: (2025)
by: Chang, Minsuk, et al.
Published: (2025)
Weak-Annotation of HAR Datasets using Vision Foundation Models
by: Bock, Marius, et al.
Published: (2024)
by: Bock, Marius, et al.
Published: (2024)
WEAR: An Outdoor Sports Dataset for Wearable and Egocentric Activity Recognition
by: Bock, Marius, et al.
Published: (2023)
by: Bock, Marius, et al.
Published: (2023)
OW-CLIP: Data-Efficient Visual Supervision for Open-World Object Detection via Human-AI Collaboration
by: Duan, Junwen, et al.
Published: (2025)
by: Duan, Junwen, et al.
Published: (2025)
CLAS: A Machine Learning Enhanced Framework for Exploring Large 3D Design Datasets
by: Zhang, XiuYu, et al.
Published: (2024)
by: Zhang, XiuYu, et al.
Published: (2024)
VideoA11y: Method and Dataset for Accessible Video Description
by: Li, Chaoyu, et al.
Published: (2025)
by: Li, Chaoyu, et al.
Published: (2025)
A Multimodal Dataset of Student Oral Presentations with Sensors and Evaluation Data
by: Becerra, Alvaro, et al.
Published: (2026)
by: Becerra, Alvaro, et al.
Published: (2026)
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents
by: Liang, Hao, et al.
Published: (2025)
by: Liang, Hao, et al.
Published: (2025)
GLIMPSE : Real-Time Text Recognition and Contextual Understanding for VQA in Wearables
by: Ramachandran, Akhil, et al.
Published: (2026)
by: Ramachandran, Akhil, et al.
Published: (2026)
A Dataset for Crucial Object Recognition in Blind and Low-Vision Individuals' Navigation
by: Islam, Md Touhidul, et al.
Published: (2024)
by: Islam, Md Touhidul, et al.
Published: (2024)
EgoPressure: A Dataset for Hand Pressure and Pose Estimation in Egocentric Vision
by: Zhao, Yiming, et al.
Published: (2024)
by: Zhao, Yiming, et al.
Published: (2024)
Talk to Parallel LiDARs: A Human-LiDAR Interaction Method Based on 3D Visual Grounding
by: Liu, Yuhang, et al.
Published: (2024)
by: Liu, Yuhang, et al.
Published: (2024)
LocoVR: Multiuser Indoor Locomotion Dataset in Virtual Reality
by: Takeyama, Kojiro, et al.
Published: (2024)
by: Takeyama, Kojiro, et al.
Published: (2024)
Spot The Ball: A Benchmark for Visual Social Inference
by: Balamurugan, Neha, et al.
Published: (2025)
by: Balamurugan, Neha, et al.
Published: (2025)
Accelerating Physical Property Reasoning for Augmented Visual Cognition
by: Lan, Hongbo, et al.
Published: (2025)
by: Lan, Hongbo, et al.
Published: (2025)
A Survey of Body and Face Motion: Datasets, Performance Evaluation Metrics and Generative Techniques
by: Sookha, Lownish Rai, et al.
Published: (2025)
by: Sookha, Lownish Rai, et al.
Published: (2025)
Similar Items
-
Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision
by: Li, Chentao, et al.
Published: (2026) -
QueryCraft: Transformer-Guided Query Initialization for Enhanced Human-Object Interaction Detection
by: Wang, Yuxiao, et al.
Published: (2025) -
MP-GUI: Modality Perception with MLLMs for GUI Understanding
by: Wang, Ziwei, et al.
Published: (2025) -
ReVis: Towards Reusable Image-Based Visualizations with MLLMs
by: Wen, Xiaolin, et al.
Published: (2026) -
RWKV-UI: UI Understanding with Enhanced Perception and Reasoning
by: Yang, Jiaxi, et al.
Published: (2025)