Beyond Single Plots: A Benchmark for Question Answering on Multi-Charts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Efat, Azher Ahmed, Song, Seok Hwan, Tavanapong, Wallapak |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Assessing Y-Axis Influence: Bias in Multimodal Language Models on Chart-to-Table Translation
von: Song, Seok Hwan, et al.
Veröffentlicht: (2026)
von: Song, Seok Hwan, et al.
Veröffentlicht: (2026)
Multi-Agent VQA: Exploring Multi-Agent Foundation Models in Zero-Shot Visual Question Answering
von: Jiang, Bowen, et al.
Veröffentlicht: (2024)
von: Jiang, Bowen, et al.
Veröffentlicht: (2024)
TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
von: Long, Qian, et al.
Veröffentlicht: (2024)
von: Long, Qian, et al.
Veröffentlicht: (2024)
EH-Benchmark Ophthalmic Hallucination Benchmark and Agent-Driven Top-Down Traceable Reasoning Workflow
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025)
PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology
von: Ghezloo, Fatemeh, et al.
Veröffentlicht: (2025)
von: Ghezloo, Fatemeh, et al.
Veröffentlicht: (2025)
MARIC: Multi-Agent Reasoning for Image Classification
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
von: Seo, Wonduk, et al.
Veröffentlicht: (2025)
Curriculum Guided Massive Multi Agent System Solving For Robust Long Horizon Tasks
von: Kar, Indrajit, et al.
Veröffentlicht: (2025)
von: Kar, Indrajit, et al.
Veröffentlicht: (2025)
PreMind: Multi-Agent Video Understanding for Advanced Indexing of Presentation-style Videos
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
von: Wei, Kangda, et al.
Veröffentlicht: (2025)
Is Large Language Model Performance on Reasoning Tasks Impacted by Different Ways Questions Are Asked?
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
von: Song, Seok Hwan, et al.
Veröffentlicht: (2025)
AdaReasoner: Dynamic Tool Orchestration for Iterative Visual Reasoning
von: Song, Mingyang, et al.
Veröffentlicht: (2026)
von: Song, Mingyang, et al.
Veröffentlicht: (2026)
Chain of Questions: Guiding Multimodal Curiosity in Language Models
von: Iji, Nima, et al.
Veröffentlicht: (2025)
von: Iji, Nima, et al.
Veröffentlicht: (2025)
RoboTwin 2.0: A Scalable Data Generator and Benchmark with Strong Domain Randomization for Robust Bimanual Robotic Manipulation
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
von: Chen, Tianxing, et al.
Veröffentlicht: (2025)
Judge Model for Large-scale Multimodality Benchmarks
von: Shih, Min-Han, et al.
Veröffentlicht: (2026)
von: Shih, Min-Han, et al.
Veröffentlicht: (2026)
SWIRL: A Staged Workflow for Interleaved Reinforcement Learning in Mobile GUI Control
von: Lu, Quanfeng, et al.
Veröffentlicht: (2025)
von: Lu, Quanfeng, et al.
Veröffentlicht: (2025)
Enhancing Agentic Autonomous Scientific Discovery with Vision-Language Model Capabilities
von: Gandhi, Kahaan, et al.
Veröffentlicht: (2025)
von: Gandhi, Kahaan, et al.
Veröffentlicht: (2025)
Towards Rationality in Language and Multimodal Agents: A Survey
von: Jiang, Bowen, et al.
Veröffentlicht: (2024)
von: Jiang, Bowen, et al.
Veröffentlicht: (2024)
Paper2Poster: Towards Multimodal Poster Automation from Scientific Papers
von: Pang, Wei, et al.
Veröffentlicht: (2025)
von: Pang, Wei, et al.
Veröffentlicht: (2025)
An Agentic System for Rare Disease Diagnosis with Traceable Reasoning
von: Zhao, Weike, et al.
Veröffentlicht: (2025)
von: Zhao, Weike, et al.
Veröffentlicht: (2025)
Metropolis-Hastings Captioning Game: Knowledge Fusion of Vision Language Models via Decentralized Bayesian Inference
von: Matsui, Yuta, et al.
Veröffentlicht: (2025)
von: Matsui, Yuta, et al.
Veröffentlicht: (2025)
REVECA: Adaptive Planning and Trajectory-based Validation in Cooperative Language Agents using Information Relevance and Relative Proximity
von: Seo, SeungWon, et al.
Veröffentlicht: (2024)
von: Seo, SeungWon, et al.
Veröffentlicht: (2024)
Transductive Visual Programming: Evolving Tool Libraries from Experience for Spatial Reasoning
von: Wu, Shengguang, et al.
Veröffentlicht: (2025)
von: Wu, Shengguang, et al.
Veröffentlicht: (2025)
InnoGym: Benchmarking the Innovation Potential of AI Agents
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
VideoMultiAgents: A Multi-Agent Framework for Video Question Answering
von: Kugo, Noriyuki, et al.
Veröffentlicht: (2025)
von: Kugo, Noriyuki, et al.
Veröffentlicht: (2025)
Automated Vehicles Should be Connected with Natural Language
von: Gao, Xiangbo, et al.
Veröffentlicht: (2025)
von: Gao, Xiangbo, et al.
Veröffentlicht: (2025)
Paper2Video: Automatic Video Generation from Scientific Papers
von: Zhu, Zeyu, et al.
Veröffentlicht: (2025)
von: Zhu, Zeyu, et al.
Veröffentlicht: (2025)
Many Heads Are Better Than One: Improved Scientific Idea Generation by A LLM-Based Multi-Agent System
von: Su, Haoyang, et al.
Veröffentlicht: (2024)
von: Su, Haoyang, et al.
Veröffentlicht: (2024)
MAP: Evaluation and Multi-Agent Enhancement of Large Language Models for Inpatient Pathways
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
von: Chen, Zhen, et al.
Veröffentlicht: (2025)
Talk to Right Specialists: Iterative Routing in Multi-agent Systems for Question Answering
von: Wu, Feijie, et al.
Veröffentlicht: (2025)
von: Wu, Feijie, et al.
Veröffentlicht: (2025)
BattleAgent: Multi-modal Dynamic Emulation on Historical Battles to Complement Historical Analysis
von: Lin, Shuhang, et al.
Veröffentlicht: (2024)
von: Lin, Shuhang, et al.
Veröffentlicht: (2024)
AutoGen Driven Multi Agent Framework for Iterative Crime Data Analysis and Prediction
von: Fatima, Syeda Kisaa, et al.
Veröffentlicht: (2025)
von: Fatima, Syeda Kisaa, et al.
Veröffentlicht: (2025)
COMBO: Compositional World Models for Embodied Multi-Agent Cooperation
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
ARGOS: Who, Where, and When in Agentic Multi-Camera Person Search
von: Kim, Myungchul, et al.
Veröffentlicht: (2026)
von: Kim, Myungchul, et al.
Veröffentlicht: (2026)
MAViS: A Multi-Agent Framework for Long-Sequence Video Storytelling
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
StoryAgent: Customized Storytelling Video Generation via Multi-Agent Collaboration
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
von: Hu, Panwen, et al.
Veröffentlicht: (2024)
Local Prompt Adaptation for Style-Consistent Multi-Object Generation in Diffusion Models
von: Sanjyal, Ankit
Veröffentlicht: (2025)
von: Sanjyal, Ankit
Veröffentlicht: (2025)
CaPo: Cooperative Plan Optimization for Efficient Embodied Multi-Agent Cooperation
von: Liu, Jie, et al.
Veröffentlicht: (2024)
von: Liu, Jie, et al.
Veröffentlicht: (2024)
Concept-RuleNet: Grounded Multi-Agent Neurosymbolic Reasoning in Vision Language Models
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
von: Sinha, Sanchit, et al.
Veröffentlicht: (2025)
A Multi-Agent System Enables Versatile Information Extraction from the Chemical Literature
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
StarCraftImage: A Dataset For Prototyping Spatial Reasoning Methods For Multi-Agent Environments
von: Kulinski, Sean, et al.
Veröffentlicht: (2024)
von: Kulinski, Sean, et al.
Veröffentlicht: (2024)
SIMPLOT: Enhancing Chart Question Answering by Distilling Essentials
von: Kim, Wonjoong, et al.
Veröffentlicht: (2024)
von: Kim, Wonjoong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Assessing Y-Axis Influence: Bias in Multimodal Language Models on Chart-to-Table Translation
von: Song, Seok Hwan, et al.
Veröffentlicht: (2026) -
Multi-Agent VQA: Exploring Multi-Agent Foundation Models in Zero-Shot Visual Question Answering
von: Jiang, Bowen, et al.
Veröffentlicht: (2024) -
TeamCraft: A Benchmark for Multi-Modal Multi-Agent Systems in Minecraft
von: Long, Qian, et al.
Veröffentlicht: (2024) -
EH-Benchmark Ophthalmic Hallucination Benchmark and Agent-Driven Top-Down Traceable Reasoning Workflow
von: Pan, Xiaoyu, et al.
Veröffentlicht: (2025) -
PathFinder: A Multi-Modal Multi-Agent System for Medical Diagnostic Decision-Making Applied to Histopathology
von: Ghezloo, Fatemeh, et al.
Veröffentlicht: (2025)