ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Shu, Zhou, Shansong, Wang, Xinyang, Wang, Shiwei, Wu, Hulong, Fang, Yixiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2024)
by: Abdallah, Abdelrahman, et al.
Published: (2024)
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
by: Salemi, Alireza, et al.
Published: (2025)
by: Salemi, Alireza, et al.
Published: (2025)
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
by: Bahaj, Adil, et al.
Published: (2025)
by: Bahaj, Adil, et al.
Published: (2025)
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
by: Yang, Hang, et al.
Published: (2024)
by: Yang, Hang, et al.
Published: (2024)
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
by: Bhattarai, Kriti, et al.
Published: (2026)
by: Bhattarai, Kriti, et al.
Published: (2026)
Benchmarking Retrieval-Augmented Multimodal Generation for Document Question Answering
by: Dong, Kuicai, et al.
Published: (2025)
by: Dong, Kuicai, et al.
Published: (2025)
DragonVerseQA: Open-Domain Long-Form Context-Aware Question-Answering
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
Passage Segmentation of Documents for Extractive Question Answering
by: Liu, Zuhong, et al.
Published: (2025)
by: Liu, Zuhong, et al.
Published: (2025)
MARA: A Multimodal Adaptive Retrieval-Augmented Framework for Document Question Answering
by: Wu, Hui, et al.
Published: (2026)
by: Wu, Hui, et al.
Published: (2026)
Faithful Temporal Question Answering over Heterogeneous Sources
by: Jia, Zhen, et al.
Published: (2024)
by: Jia, Zhen, et al.
Published: (2024)
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
by: Islamaj, Rezarta, et al.
Published: (2026)
by: Islamaj, Rezarta, et al.
Published: (2026)
Recursive Question Understanding for Complex Question Answering over Heterogeneous Personal Data
by: Christmann, Philipp, et al.
Published: (2025)
by: Christmann, Philipp, et al.
Published: (2025)
Multi-Document Financial Question Answering using LLMs
by: Shah, Shalin, et al.
Published: (2024)
by: Shah, Shalin, et al.
Published: (2024)
SyllabusQA: A Course Logistics Question Answering Dataset
by: Fernandez, Nigel, et al.
Published: (2024)
by: Fernandez, Nigel, et al.
Published: (2024)
MapQA: Open-domain Geospatial Question Answering on Map Data
by: Li, Zekun, et al.
Published: (2025)
by: Li, Zekun, et al.
Published: (2025)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025)
by: Baumgärtner, Tim, et al.
Published: (2025)
PARSE: An Open-Domain Reasoning Question Answering Benchmark for Persian
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
EnronQA: Towards Personalized RAG over Private Documents
by: Ryan, Michael J., et al.
Published: (2025)
by: Ryan, Michael J., et al.
Published: (2025)
RAG-based Question Answering over Heterogeneous Data and Text
by: Christmann, Philipp, et al.
Published: (2024)
by: Christmann, Philipp, et al.
Published: (2024)
The ReQAP System for Question Answering over Personal Information
by: Christmann, Philipp, et al.
Published: (2025)
by: Christmann, Philipp, et al.
Published: (2025)
Towards Reliable and Interpretable Document Question Answering via VLMs
by: Chen, Alessio, et al.
Published: (2025)
by: Chen, Alessio, et al.
Published: (2025)
DEXTER: A Benchmark for open-domain Complex Question Answering using LLMs
by: Prabhu, Venktesh V. Deepali, et al.
Published: (2024)
by: Prabhu, Venktesh V. Deepali, et al.
Published: (2024)
Decomposing Retrieval Failures in RAG for Long-Document Financial Question Answering
by: Kobeissi, Amine, et al.
Published: (2026)
by: Kobeissi, Amine, et al.
Published: (2026)
Passage-specific Prompt Tuning for Passage Reranking in Question Answering with Large Language Models
by: Wu, Xuyang, et al.
Published: (2024)
by: Wu, Xuyang, et al.
Published: (2024)
Inferential Question Answering
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
Knowledge-Aware Diverse Reranking for Cross-Source Question Answering
by: Zhou, Tong
Published: (2025)
by: Zhou, Tong
Published: (2025)
MedCoT-RAG: Causal Chain-of-Thought RAG for Medical Question Answering
by: Wang, Ziyu, et al.
Published: (2025)
by: Wang, Ziyu, et al.
Published: (2025)
Benchmarking Real-Time Question Answering via Executable Code Workflows
by: Zhou, Wenjie, et al.
Published: (2026)
by: Zhou, Wenjie, et al.
Published: (2026)
IRPAPERS: A Visual Document Benchmark for Scientific Retrieval and Question Answering
by: Shorten, Connor, et al.
Published: (2026)
by: Shorten, Connor, et al.
Published: (2026)
Enhancing Complex Question Answering over Knowledge Graphs through Evidence Pattern Retrieval
by: Ding, Wentao, et al.
Published: (2024)
by: Ding, Wentao, et al.
Published: (2024)
ViHERMES: A Graph-Grounded Multihop Question Answering Benchmark and System for Vietnamese Healthcare Regulations
by: Nguyen, Long S. T., et al.
Published: (2026)
by: Nguyen, Long S. T., et al.
Published: (2026)
REAR: A Relevance-Aware Retrieval-Augmented Framework for Open-Domain Question Answering
by: Wang, Yuhao, et al.
Published: (2024)
by: Wang, Yuhao, et al.
Published: (2024)
Evaluating Large Language Models in Semantic Parsing for Conversational Question Answering over Knowledge Graphs
by: Schneider, Phillip, et al.
Published: (2024)
by: Schneider, Phillip, et al.
Published: (2024)
SRAG: Structured Retrieval-Augmented Generation for Multi-Entity Question Answering over Wikipedia Graph
by: Lin, Teng, et al.
Published: (2025)
by: Lin, Teng, et al.
Published: (2025)
JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability
by: Wang, Junda, et al.
Published: (2024)
by: Wang, Junda, et al.
Published: (2024)
TTQA-RS- A break-down prompting approach for Multi-hop Table-Text Question Answering with Reasoning and Summarization
by: Bardhan, Jayetri, et al.
Published: (2024)
by: Bardhan, Jayetri, et al.
Published: (2024)
FinCARDS: Card-Based Analyst Reranking for Financial Document Question Answering
by: Zhou, Yixi, et al.
Published: (2026)
by: Zhou, Yixi, et al.
Published: (2026)
It's High Time: A Survey of Temporal Question Answering
by: Piryani, Bhawna, et al.
Published: (2025)
by: Piryani, Bhawna, et al.
Published: (2025)
Context Convergence Improves Answering Inferential Questions
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
Entity Retrieval for Answering Entity-Centric Questions
by: Shavarani, Hassan S., et al.
Published: (2024)
by: Shavarani, Hassan S., et al.
Published: (2024)
Similar Items
-
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2024) -
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
by: Salemi, Alireza, et al.
Published: (2025) -
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
by: Bahaj, Adil, et al.
Published: (2025) -
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
by: Yang, Hang, et al.
Published: (2024) -
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
by: Bhattarai, Kriti, et al.
Published: (2026)