Saved in:
| Main Authors: | Bhattarai, Kriti, Keloth, Vipina K., Wright, Donald, Loza, Andrew, Ren, Yang, Xu, Hua |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2601.12632 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
by: Islamaj, Rezarta, et al.
Published: (2026)
by: Islamaj, Rezarta, et al.
Published: (2026)
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
by: Bahaj, Adil, et al.
Published: (2025)
by: Bahaj, Adil, et al.
Published: (2025)
ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
by: Wang, Shu, et al.
Published: (2026)
by: Wang, Shu, et al.
Published: (2026)
Overview of BioASQ 2025: The Thirteenth BioASQ Challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
by: Nentidis, Anastasios, et al.
Published: (2025)
by: Nentidis, Anastasios, et al.
Published: (2025)
Overview of BioASQ 2024: The twelfth BioASQ challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
by: Nentidis, Anastasios, et al.
Published: (2025)
by: Nentidis, Anastasios, et al.
Published: (2025)
Overview of BioASQ 2022: The tenth BioASQ challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
by: Nentidis, Anastasios, et al.
Published: (2022)
by: Nentidis, Anastasios, et al.
Published: (2022)
HCT-QA: A Benchmark for Question Answering on Human-Centric Tables
by: Ahmad, Mohammad S., et al.
Published: (2025)
by: Ahmad, Mohammad S., et al.
Published: (2025)
Jamendo-MT-QA: A Benchmark for Multi-Track Comparative Music Question Answering
by: Koh, Junyoung, et al.
Published: (2026)
by: Koh, Junyoung, et al.
Published: (2026)
SyllabusQA: A Course Logistics Question Answering Dataset
by: Fernandez, Nigel, et al.
Published: (2024)
by: Fernandez, Nigel, et al.
Published: (2024)
LaMP-QA: A Benchmark for Personalized Long-form Question Answering
by: Salemi, Alireza, et al.
Published: (2025)
by: Salemi, Alireza, et al.
Published: (2025)
Evaluating Robustness of Generative Search Engine on Adversarial Factual Questions
by: Hu, Xuming, et al.
Published: (2024)
by: Hu, Xuming, et al.
Published: (2024)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
by: Abdallah, Abdelrahman, et al.
Published: (2024)
by: Abdallah, Abdelrahman, et al.
Published: (2024)
KG20C & KG20C-QA: Scholarly Knowledge Graph Benchmarks for Link Prediction and Question Answering
by: Tran, Hung-Nghiep, et al.
Published: (2025)
by: Tran, Hung-Nghiep, et al.
Published: (2025)
IRB: Automated Generation of Robust Factuality Benchmarks
by: Do, Lam Thanh, et al.
Published: (2026)
by: Do, Lam Thanh, et al.
Published: (2026)
FabricQA-Extractor: A Question Answering System to Extract Information from Documents using Natural Language Questions
by: Wang, Qiming, et al.
Published: (2024)
by: Wang, Qiming, et al.
Published: (2024)
OpenLifelogQA: An Open-Ended Multi-Modal Lifelog Question-Answering Dataset
by: Tran, Quang-Linh, et al.
Published: (2025)
by: Tran, Quang-Linh, et al.
Published: (2025)
SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting
by: Ali, Mohammed, et al.
Published: (2025)
by: Ali, Mohammed, et al.
Published: (2025)
MapQA: Open-domain Geospatial Question Answering on Map Data
by: Li, Zekun, et al.
Published: (2025)
by: Li, Zekun, et al.
Published: (2025)
NeuroSym-BioCAT: Leveraging Neuro-Symbolic Methods for Biomedical Scholarly Document Categorization and Question Answering
by: Zamil, Parvez, et al.
Published: (2024)
by: Zamil, Parvez, et al.
Published: (2024)
LLM-MedQA: Enhancing Medical Question Answering through Case Studies in Large Language Models
by: Yang, Hang, et al.
Published: (2024)
by: Yang, Hang, et al.
Published: (2024)
DragonVerseQA: Open-Domain Long-Form Context-Aware Question-Answering
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
by: Lahiri, Aritra Kumar, et al.
Published: (2024)
Fact or Facsimile? Evaluating the Factual Robustness of Modern Retrievers
by: Wu, Haoyu, et al.
Published: (2025)
by: Wu, Haoyu, et al.
Published: (2025)
Answer Retrieval in Legal Community Question Answering
by: Askari, Arian, et al.
Published: (2024)
by: Askari, Arian, et al.
Published: (2024)
PeerQA: A Scientific Question Answering Dataset from Peer Reviews
by: Baumgärtner, Tim, et al.
Published: (2025)
by: Baumgärtner, Tim, et al.
Published: (2025)
Evaluating Search Engines and Large Language Models for Answering Health Questions
by: Fernández-Pichel, Marcos, et al.
Published: (2024)
by: Fernández-Pichel, Marcos, et al.
Published: (2024)
Efficient and Reproducible Biomedical Question Answering using Retrieval Augmented Generation
by: Stuhlmann, Linus, et al.
Published: (2025)
by: Stuhlmann, Linus, et al.
Published: (2025)
Towards Robust Expert Finding in Community Question Answering Platforms
by: Amendola, Maddalena, et al.
Published: (2025)
by: Amendola, Maddalena, et al.
Published: (2025)
MedTrust-RAG: Evidence Verification and Trust Alignment for Biomedical Question Answering
by: Ning, Yingpeng, et al.
Published: (2025)
by: Ning, Yingpeng, et al.
Published: (2025)
An Empirical Study of Evaluating Long-form Question Answering
by: Xian, Ning, et al.
Published: (2025)
by: Xian, Ning, et al.
Published: (2025)
Evaluating Position Bias in Large Language Model Recommendations
by: Bito, Ethan, et al.
Published: (2025)
by: Bito, Ethan, et al.
Published: (2025)
Optimizing Question Semantic Space for Dynamic Retrieval-Augmented Multi-hop Question Answering
by: Ye, Linhao, et al.
Published: (2025)
by: Ye, Linhao, et al.
Published: (2025)
Are Smaller Open-Weight LLMs Closing the Gap to Proprietary Models for Biomedical Question Answering?
by: Stachura, Damian, et al.
Published: (2025)
by: Stachura, Damian, et al.
Published: (2025)
VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering
by: Košprdić, Miloš, et al.
Published: (2026)
by: Košprdić, Miloš, et al.
Published: (2026)
Biomedical Question Answering via Multi-Level Summarization on a Local Knowledge Graph
by: Guan, Lingxiao, et al.
Published: (2025)
by: Guan, Lingxiao, et al.
Published: (2025)
Investigating the Factual Knowledge Boundary of Large Language Models with Retrieval Augmentation
by: Ren, Ruiyang, et al.
Published: (2023)
by: Ren, Ruiyang, et al.
Published: (2023)
PARSE: An Open-Domain Reasoning Question Answering Benchmark for Persian
by: Mozafari, Jamshid, et al.
Published: (2026)
by: Mozafari, Jamshid, et al.
Published: (2026)
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
by: Alawwad, Hessa A., et al.
Published: (2025)
by: Alawwad, Hessa A., et al.
Published: (2025)
Generate-then-Ground in Retrieval-Augmented Generation for Multi-hop Question Answering
by: Shi, Zhengliang, et al.
Published: (2024)
by: Shi, Zhengliang, et al.
Published: (2024)
UQABench: Evaluating User Embedding for Prompting LLMs in Personalized Question Answering
by: Liu, Langming, et al.
Published: (2025)
by: Liu, Langming, et al.
Published: (2025)
Evaluating Large Language Models in Semantic Parsing for Conversational Question Answering over Knowledge Graphs
by: Schneider, Phillip, et al.
Published: (2024)
by: Schneider, Phillip, et al.
Published: (2024)
Similar Items
-
MedHopQA: A Disease-Centered Multi-Hop Reasoning Benchmark and Evaluation Framework for LLM-Based Biomedical Question Answering
by: Islamaj, Rezarta, et al.
Published: (2026) -
MizanQA: Benchmarking Large Language Models on Moroccan Legal Question Answering
by: Bahaj, Adil, et al.
Published: (2025) -
ASTRA-QA: A Benchmark for Abstract Question Answering over Documents
by: Wang, Shu, et al.
Published: (2026) -
Overview of BioASQ 2025: The Thirteenth BioASQ Challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
by: Nentidis, Anastasios, et al.
Published: (2025) -
Overview of BioASQ 2024: The twelfth BioASQ challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
by: Nentidis, Anastasios, et al.
Published: (2025)