FlowVQA: Mapping Multimodal Logic in Visual Question Answering with Flowcharts
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Shubhankar, Chaurasia, Purvi, Varun, Yerram, Pandya, Pranshu, Gupta, Vatsal, Gupta, Vivek, Roth, Dan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
von: Gupta, Vatsal, et al.
Veröffentlicht: (2023)
von: Gupta, Vatsal, et al.
Veröffentlicht: (2023)
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024)
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024)
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025)
CORE-T: COherent REtrieval of Tables for Text-to-SQL
von: Soliman, Hassan, et al.
Veröffentlicht: (2026)
von: Soliman, Hassan, et al.
Veröffentlicht: (2026)
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)
Federated Retrieval-Augmented Generation: A Systematic Mapping Study
von: Chakraborty, Abhijit, et al.
Veröffentlicht: (2025)
von: Chakraborty, Abhijit, et al.
Veröffentlicht: (2025)
TransientTables: Evaluating LLMs' Reasoning on Temporally Evolving Semi-structured Tables
von: Shankarampeta, Abhilash, et al.
Veröffentlicht: (2025)
von: Shankarampeta, Abhilash, et al.
Veröffentlicht: (2025)
CRAFT: Training-Free Cascaded Retrieval for Tabular QA
von: Singh, Adarsh, et al.
Veröffentlicht: (2025)
von: Singh, Adarsh, et al.
Veröffentlicht: (2025)
TalentMine: LLM-Based Extraction and Question-Answering from Multimodal Talent Tables
von: Mannam, Varun, et al.
Veröffentlicht: (2025)
von: Mannam, Varun, et al.
Veröffentlicht: (2025)
ZeShot-VQA: Zero-Shot Visual Question Answering Framework with Answer Mapping for Natural Disaster Damage Assessment
von: Karimi, Ehsan, et al.
Veröffentlicht: (2025)
von: Karimi, Ehsan, et al.
Veröffentlicht: (2025)
Cascading Adaptors to Leverage English Data to Improve Performance of Question Answering for Low-Resource Languages
von: Pandya, Hariom A., et al.
Veröffentlicht: (2021)
von: Pandya, Hariom A., et al.
Veröffentlicht: (2021)
Cross-modal Retrieval for Knowledge-based Visual Question Answering
von: Lerner, Paul, et al.
Veröffentlicht: (2024)
von: Lerner, Paul, et al.
Veröffentlicht: (2024)
MapQA: Open-domain Geospatial Question Answering on Map Data
von: Li, Zekun, et al.
Veröffentlicht: (2025)
von: Li, Zekun, et al.
Veröffentlicht: (2025)
Autofocus Retrieval: An Effective Pipeline for Multi-Hop Question Answering With Semi-Structured Knowledge
von: Boer, Derian, et al.
Veröffentlicht: (2025)
von: Boer, Derian, et al.
Veröffentlicht: (2025)
fact check AI at SemEval-2025 Task 7: Multilingual and Crosslingual Fact-checked Claim Retrieval
von: Rastogi, Pranshu
Veröffentlicht: (2025)
von: Rastogi, Pranshu
Veröffentlicht: (2025)
Inferential Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
Evaluating LLMs' Mathematical Reasoning in Financial Document Question Answering
von: Srivastava, Pragya, et al.
Veröffentlicht: (2024)
von: Srivastava, Pragya, et al.
Veröffentlicht: (2024)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
von: Dong, Kuicai, et al.
Veröffentlicht: (2025)
MuRAR: A Simple and Effective Multimodal Retrieval and Answer Refinement Framework for Multimodal Question Answering
von: Zhu, Zhengyuan, et al.
Veröffentlicht: (2024)
von: Zhu, Zhengyuan, et al.
Veröffentlicht: (2024)
Hierarchical Vision-Language Reasoning for Multimodal Multiple-Choice Question Answering
von: Zhou, Ao, et al.
Veröffentlicht: (2025)
von: Zhou, Ao, et al.
Veröffentlicht: (2025)
A Multimodal Dense Retrieval Approach for Speech-Based Open-Domain Question Answering
von: Sidiropoulos, Georgios, et al.
Veröffentlicht: (2024)
von: Sidiropoulos, Georgios, et al.
Veröffentlicht: (2024)
Improving Robustness of Tabular Retrieval via Representational Stability
von: Bhandari, Kushal Raj, et al.
Veröffentlicht: (2026)
von: Bhandari, Kushal Raj, et al.
Veröffentlicht: (2026)
State Space Models are Strong Text Rerankers
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
von: Xu, Zhichao, et al.
Veröffentlicht: (2024)
Index Light, Reason Deep: Deferred Visual Ingestion for Visual-Dense Document Question Answering
von: Xu, Tao
Veröffentlicht: (2026)
von: Xu, Tao
Veröffentlicht: (2026)
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
Retrieval-Augmented Visual Question Answering via Built-in Autoregressive Search Engines
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
von: Long, Xinwei, et al.
Veröffentlicht: (2025)
Weaver: Interweaving SQL and LLM for Table Reasoning
von: Khoja, Rohit, et al.
Veröffentlicht: (2025)
von: Khoja, Rohit, et al.
Veröffentlicht: (2025)
Retrieval Augmented Generation for Domain-specific Question Answering
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
von: Sharma, Sanat, et al.
Veröffentlicht: (2024)
Context Convergence Improves Answering Inferential Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2026)
Passage Segmentation of Documents for Extractive Question Answering
von: Liu, Zuhong, et al.
Veröffentlicht: (2025)
von: Liu, Zuhong, et al.
Veröffentlicht: (2025)
Entity Retrieval for Answering Entity-Centric Questions
von: Shavarani, Hassan S., et al.
Veröffentlicht: (2024)
von: Shavarani, Hassan S., et al.
Veröffentlicht: (2024)
1-800-SHARED-TASKS at RegNLP: Lexical Reranking of Semantic Retrieval (LeSeR) for Regulatory Question Answering
von: Purbey, Jebish, et al.
Veröffentlicht: (2024)
von: Purbey, Jebish, et al.
Veröffentlicht: (2024)
A Breadth-First Catalog of Text Processing, Speech Processing and Multimodal Research in South Asian Languages
von: Gupta, Pranav
Veröffentlicht: (2024)
von: Gupta, Pranav
Veröffentlicht: (2024)
Rethinking Information Synthesis in Multimodal Question Answering A Multi-Agent Perspective
von: Rajput, Krishna Singh, et al.
Veröffentlicht: (2025)
von: Rajput, Krishna Singh, et al.
Veröffentlicht: (2025)
Recursive Question Understanding for Complex Question Answering over Heterogeneous Personal Data
von: Christmann, Philipp, et al.
Veröffentlicht: (2025)
von: Christmann, Philipp, et al.
Veröffentlicht: (2025)
MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering
von: Wu, Hui, et al.
Veröffentlicht: (2026)
von: Wu, Hui, et al.
Veröffentlicht: (2026)
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
von: Qian, Deniz, et al.
Veröffentlicht: (2026)
von: Qian, Deniz, et al.
Veröffentlicht: (2026)
It's High Time: A Survey of Temporal Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
Multi-Document Financial Question Answering using LLMs
von: Shah, Shalin, et al.
Veröffentlicht: (2024)
von: Shah, Shalin, et al.
Veröffentlicht: (2024)
Faithful Temporal Question Answering over Heterogeneous Sources
von: Jia, Zhen, et al.
Veröffentlicht: (2024)
von: Jia, Zhen, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Evaluating Concurrent Robustness of Language Models Across Diverse Challenge Sets
von: Gupta, Vatsal, et al.
Veröffentlicht: (2023) -
NTSEBENCH: Cognitive Reasoning Benchmark for Vision Language Models
von: Pandya, Pranshu, et al.
Veröffentlicht: (2024) -
RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables
von: Abhyankar, Nikhil, et al.
Veröffentlicht: (2025) -
CORE-T: COherent REtrieval of Tables for Text-to-SQL
von: Soliman, Hassan, et al.
Veröffentlicht: (2026) -
MapIQ: Evaluating Multimodal Large Language Models for Map Question Answering
von: Srivastava, Varun, et al.
Veröffentlicht: (2025)