BioGraphletQA: Knowledge-Anchored Generation of Complex QA Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jonker, Richard A. A., Martins, Bárbara Maria Ribeiro de Abreu, Matos, Sérgio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA
von: Jonker, Richard A. A., et al.
Veröffentlicht: (2026)
von: Jonker, Richard A. A., et al.
Veröffentlicht: (2026)
JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge
von: Cao, Zhihan, et al.
Veröffentlicht: (2025)
von: Cao, Zhihan, et al.
Veröffentlicht: (2025)
RJUA-QA: A Comprehensive QA Dataset for Urology
von: Lyu, Shiwei, et al.
Veröffentlicht: (2023)
von: Lyu, Shiwei, et al.
Veröffentlicht: (2023)
KET-QA: A Dataset for Knowledge Enhanced Table Question Answering
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
FictionalQA: A Dataset for Studying Memorization and Knowledge Acquisition
von: Kirchenbauer, John, et al.
Veröffentlicht: (2025)
von: Kirchenbauer, John, et al.
Veröffentlicht: (2025)
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
HumMusQA: A Human-written Music Understanding QA Benchmark Dataset
von: Weck, Benno, et al.
Veröffentlicht: (2026)
von: Weck, Benno, et al.
Veröffentlicht: (2026)
Syn-QA2: Evaluating False Assumptions in Long-tail Questions with Synthetic QA Datasets
von: Daswani, Ashwin, et al.
Veröffentlicht: (2024)
von: Daswani, Ashwin, et al.
Veröffentlicht: (2024)
LinkQA: Synthesizing Diverse QA from Multiple Seeds Strongly Linked by Knowledge Points
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
von: Zhang, Xuemiao, et al.
Veröffentlicht: (2025)
CUS-QA: Local-Knowledge-Oriented Open-Ended Question Answering Dataset
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
von: Libovický, Jindřich, et al.
Veröffentlicht: (2025)
AirQA: A Comprehensive QA Dataset for AI Research with Instance-Level Evaluation
von: Huang, Tiancheng, et al.
Veröffentlicht: (2025)
von: Huang, Tiancheng, et al.
Veröffentlicht: (2025)
PolQA: Polish Question Answering Dataset
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
von: Rybak, Piotr, et al.
Veröffentlicht: (2022)
UETQuintet at BioCreative IX -- MedHopQA: Enhancing Biomedical QA with Selective Multi-hop Reasoning and Contextual Retrieval
von: Nguyen, Quoc-An, et al.
Veröffentlicht: (2026)
von: Nguyen, Quoc-An, et al.
Veröffentlicht: (2026)
StorySparkQA: Expert-Annotated QA Pairs with Real-World Knowledge for Children's Story-Based Learning
von: Chen, Jiaju, et al.
Veröffentlicht: (2023)
von: Chen, Jiaju, et al.
Veröffentlicht: (2023)
GRS-QA -- Graph Reasoning-Structured Question Answering Dataset
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
von: Pahilajani, Anish, et al.
Veröffentlicht: (2024)
Can LLMs Evaluate Complex Attribution in QA? Automatic Benchmarking using Knowledge Graphs
von: Hu, Nan, et al.
Veröffentlicht: (2024)
von: Hu, Nan, et al.
Veröffentlicht: (2024)
SEC-QA: A Systematic Evaluation Corpus for Financial QA
von: Lai, Viet Dac, et al.
Veröffentlicht: (2024)
von: Lai, Viet Dac, et al.
Veröffentlicht: (2024)
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
von: Dineen, Jacob, et al.
Veröffentlicht: (2025)
von: Dineen, Jacob, et al.
Veröffentlicht: (2025)
JDocQA: Japanese Document Question Answering Dataset for Generative Language Models
von: Onami, Eri, et al.
Veröffentlicht: (2024)
von: Onami, Eri, et al.
Veröffentlicht: (2024)
DebateQA: Evaluating Question Answering on Debatable Knowledge
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
von: Xu, Rongwu, et al.
Veröffentlicht: (2024)
NeoQA: Evidence-based Question Answering with Generated News Events
von: Glockner, Max, et al.
Veröffentlicht: (2025)
von: Glockner, Max, et al.
Veröffentlicht: (2025)
CaresAI at BioCreative IX Track 1 -- LLM for Biomedical QA
von: Abdel-Salam, Reem, et al.
Veröffentlicht: (2025)
von: Abdel-Salam, Reem, et al.
Veröffentlicht: (2025)
FoQA: A Faroese Question-Answering Dataset
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
von: Simonsen, Annika, et al.
Veröffentlicht: (2025)
Augmenting LLM Reasoning with Dynamic Notes Writing for Complex QA
von: Maheshwary, Rishabh, et al.
Veröffentlicht: (2025)
von: Maheshwary, Rishabh, et al.
Veröffentlicht: (2025)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
Suvach -- Generated Hindi QA benchmark
von: Narayanan, Vaishak, et al.
Veröffentlicht: (2024)
von: Narayanan, Vaishak, et al.
Veröffentlicht: (2024)
emrQA-msquad: A Medical Dataset Structured with the SQuAD V2.0 Framework, Enriched with emrQA Medical Information
von: Eladio, Jimenez, et al.
Veröffentlicht: (2024)
von: Eladio, Jimenez, et al.
Veröffentlicht: (2024)
LiteraryQA: Towards Effective Evaluation of Long-document Narrative QA
von: Bonomo, Tommaso, et al.
Veröffentlicht: (2025)
von: Bonomo, Tommaso, et al.
Veröffentlicht: (2025)
ExpertGenQA: Open-ended QA generation in Specialized Domains
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2025)
von: Shahgir, Haz Sameen, et al.
Veröffentlicht: (2025)
MedConceptsQA: Open Source Medical Concepts QA Benchmark
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2024)
von: Shoham, Ofir Ben, et al.
Veröffentlicht: (2024)
ProMQA-Assembly: Multimodal Procedural QA Dataset on Assembly
von: Hasegawa, Kimihiro, et al.
Veröffentlicht: (2025)
von: Hasegawa, Kimihiro, et al.
Veröffentlicht: (2025)
PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant
von: Yin, Congrui, et al.
Veröffentlicht: (2025)
von: Yin, Congrui, et al.
Veröffentlicht: (2025)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
SwaQuAD-24: QA Benchmark Dataset in Swahili
von: Kondoro, Alfred Malengo
Veröffentlicht: (2024)
von: Kondoro, Alfred Malengo
Veröffentlicht: (2024)
RAG-BioQA: A Retrieval-Augmented Generation Framework for Long-Form Biomedical Question Answering
von: Panchumarthi, Lovely Yeswanth, et al.
Veröffentlicht: (2025)
von: Panchumarthi, Lovely Yeswanth, et al.
Veröffentlicht: (2025)
MoleculeQA: A Dataset to Evaluate Factual Accuracy in Molecular Comprehension
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
von: Lu, Xingyu, et al.
Veröffentlicht: (2024)
Reassessing Extractive QA Datasets at Scale: LLM-as-a-Judge and In-Depth Analyses
von: Ho, Xanh, et al.
Veröffentlicht: (2025)
von: Ho, Xanh, et al.
Veröffentlicht: (2025)
HypoTermQA: Hypothetical Terms Dataset for Benchmarking Hallucination Tendency of LLMs
von: Uluoglakci, Cem, et al.
Veröffentlicht: (2024)
von: Uluoglakci, Cem, et al.
Veröffentlicht: (2024)
CondAmbigQA: A Benchmark and Dataset for Conditional Ambiguous Question Answering
von: Li, Zongxi, et al.
Veröffentlicht: (2025)
von: Li, Zongxi, et al.
Veröffentlicht: (2025)
KGQuest: Template-Driven QA Generation from Knowledge Graphs with LLM-Based Refinement
von: Nayab, Sania, et al.
Veröffentlicht: (2025)
von: Nayab, Sania, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA
von: Jonker, Richard A. A., et al.
Veröffentlicht: (2026) -
JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge
von: Cao, Zhihan, et al.
Veröffentlicht: (2025) -
RJUA-QA: A Comprehensive QA Dataset for Urology
von: Lyu, Shiwei, et al.
Veröffentlicht: (2023) -
KET-QA: A Dataset for Knowledge Enhanced Table Question Answering
von: Hu, Mengkang, et al.
Veröffentlicht: (2024) -
FictionalQA: A Dataset for Studying Memorization and Knowledge Acquisition
von: Kirchenbauer, John, et al.
Veröffentlicht: (2025)