BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA
Fuente:
arXiv
Saved in:
| Main Authors: | Jonker, Richard A. A., Christiansen, Alexander, Maniatis, Alexandros, Garrido, Rúben, Lima, Rogério Braunschweiger de Freitas, Jurowetzki, Roman, Matos, Sérgio |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
by: Majeedi, Abrar, et al.
Published: (2026)
by: Majeedi, Abrar, et al.
Published: (2026)
sebis at ArchEHR-QA 2026: How Much Can You Do Locally? Evaluating Grounded EHR QA on a Single Notebook
by: Yurt, Ibrahim Ebrar, et al.
Published: (2026)
by: Yurt, Ibrahim Ebrar, et al.
Published: (2026)
UTSA-NLP at ArchEHR-QA 2025: Improving EHR Question Answering via Self-Consistency Prompting
by: Shields-Menard, Sara, et al.
Published: (2025)
by: Shields-Menard, Sara, et al.
Published: (2025)
Neural at ArchEHR-QA 2025: Agentic Prompt Optimization for Evidence-Grounded Clinical Question Answering
by: Bogireddy, Sai Prasanna Teja Reddy, et al.
Published: (2025)
by: Bogireddy, Sai Prasanna Teja Reddy, et al.
Published: (2025)
Yale-DM-Lab at ArchEHR-QA 2026: Deterministic Grounding and Multi-Pass Evidence Alignment for EHR Question Answering
by: Irankhah, Elyas, et al.
Published: (2026)
by: Irankhah, Elyas, et al.
Published: (2026)
HealthNLP_Retrievers at ArchEHR-QA 2026: Cascaded LLM Pipeline for Grounded Clinical Question Answering
by: Hosen, Md Biplob, et al.
Published: (2026)
by: Hosen, Md Biplob, et al.
Published: (2026)
heiDS at ArchEHR-QA 2025: From Fixed-k to Query-dependent-k for Retrieval Augmented Generation
by: Chouhan, Ashish, et al.
Published: (2025)
by: Chouhan, Ashish, et al.
Published: (2025)
ArgHiTZ at ArchEHR-QA 2025: A Two-Step Divide and Conquer Approach to Patient Question Answering for Top Factuality
by: Cuadrón, Adrián, et al.
Published: (2025)
by: Cuadrón, Adrián, et al.
Published: (2025)
QU-NLP at ArchEHR-QA 2026: Two-Stage QLoRA Fine-Tuning of Qwen3-4B for Patient-Oriented Clinical Question Answering and Evidence Sentence Alignment
by: AL-Smadi, Mohammad
Published: (2026)
by: AL-Smadi, Mohammad
Published: (2026)
BioGraphletQA: Knowledge-Anchored Generation of Complex QA Datasets
by: Jonker, Richard A. A., et al.
Published: (2026)
by: Jonker, Richard A. A., et al.
Published: (2026)
ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities
by: Xu, Peng, et al.
Published: (2024)
by: Xu, Peng, et al.
Published: (2024)
Team Anotheroption at SemEval-2025 Task 8: Bridging the Gap Between Open-Source and Proprietary LLMs in Table QA
by: Evkarpidi, Nikolas, et al.
Published: (2025)
by: Evkarpidi, Nikolas, et al.
Published: (2025)
MedConceptsQA: Open Source Medical Concepts QA Benchmark
by: Shoham, Ofir Ben, et al.
Published: (2024)
by: Shoham, Ofir Ben, et al.
Published: (2024)
Continuous QA Learning with Structured Prompts
by: Zheng, Yinhe
Published: (2022)
by: Zheng, Yinhe
Published: (2022)
Answering Questions in Stages: Prompt Chaining for Contract QA
by: Roegiest, Adam, et al.
Published: (2024)
by: Roegiest, Adam, et al.
Published: (2024)
RJUA-QA: A Comprehensive QA Dataset for Urology
by: Lyu, Shiwei, et al.
Published: (2023)
by: Lyu, Shiwei, et al.
Published: (2023)
SEC-QA: A Systematic Evaluation Corpus for Financial QA
by: Lai, Viet Dac, et al.
Published: (2024)
by: Lai, Viet Dac, et al.
Published: (2024)
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
by: Dineen, Jacob, et al.
Published: (2025)
by: Dineen, Jacob, et al.
Published: (2025)
ChatQA: Surpassing GPT-4 on Conversational QA and RAG
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
by: Le, Chenqian, et al.
Published: (2025)
by: Le, Chenqian, et al.
Published: (2025)
LiteraryQA: Towards Effective Evaluation of Long-document Narrative QA
by: Bonomo, Tommaso, et al.
Published: (2025)
by: Bonomo, Tommaso, et al.
Published: (2025)
ExpertGenQA: Open-ended QA generation in Specialized Domains
by: Shahgir, Haz Sameen, et al.
Published: (2025)
by: Shahgir, Haz Sameen, et al.
Published: (2025)
Harnessing Collective Intelligence of LLMs for Robust Biomedical QA: A Multi-Model Approach
by: Panou, Dimitra, et al.
Published: (2025)
by: Panou, Dimitra, et al.
Published: (2025)
Explainablity QA dataset
by: Anonymous
Published: (2026)
by: Anonymous
Published: (2026)
AgriQA Dataset
by: Eldem, Ayşe
Published: (2026)
by: Eldem, Ayşe
Published: (2026)
Few-Shot Prompting for Extractive Quranic QA with Instruction-Tuned LLMs
by: Basem, Mohamed, et al.
Published: (2025)
by: Basem, Mohamed, et al.
Published: (2025)
GenQA: Generating Millions of Instructions from a Handful of Prompts
by: Chen, Jiuhai, et al.
Published: (2024)
by: Chen, Jiuhai, et al.
Published: (2024)
HiQA: A Hierarchical Contextual Augmentation RAG for Multi-Documents QA
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
SimpleDevQA: Benchmarking Large Language Models on Development Knowledge QA
by: Zhang, Jing, et al.
Published: (2025)
by: Zhang, Jing, et al.
Published: (2025)
CON-QA: Privacy-Preserving QA using cloud LLMs in Contract Domain
by: Singh, Ajeet Kumar, et al.
Published: (2025)
by: Singh, Ajeet Kumar, et al.
Published: (2025)
AirQA: A Comprehensive QA Dataset for AI Research with Instance-Level Evaluation
by: Huang, Tiancheng, et al.
Published: (2025)
by: Huang, Tiancheng, et al.
Published: (2025)
DeQA-Doc: Adapting DeQA-Score to Document Image Quality Assessment
by: Gao, Junjie, et al.
Published: (2025)
by: Gao, Junjie, et al.
Published: (2025)
JBE-QA: Japanese Bar Exam QA Dataset for Assessing Legal Domain Knowledge
by: Cao, Zhihan, et al.
Published: (2025)
by: Cao, Zhihan, et al.
Published: (2025)
HumMusQA: A Human-written Music Understanding QA Benchmark Dataset
by: Weck, Benno, et al.
Published: (2026)
by: Weck, Benno, et al.
Published: (2026)
Efficient Table QA via TableGrid Navigation and Progressive Inference Prompting
by: Maurya, Amritansh, et al.
Published: (2026)
by: Maurya, Amritansh, et al.
Published: (2026)
Synthetic Data-Driven Prompt Tuning for Financial QA over Tables and Documents
by: Yu, Yaoning, et al.
Published: (2025)
by: Yu, Yaoning, et al.
Published: (2025)
Robust Driving QA through Metadata-Grounded Context and Task-Specific Prompts
by: Yu, Seungjun, et al.
Published: (2025)
by: Yu, Seungjun, et al.
Published: (2025)
Self-Prompting Large Language Models for Zero-Shot Open-Domain QA
by: Li, Junlong, et al.
Published: (2022)
by: Li, Junlong, et al.
Published: (2022)
Audiopedia: Audio QA with Knowledge
by: Penamakuri, Abhirama Subramanyam, et al.
Published: (2024)
by: Penamakuri, Abhirama Subramanyam, et al.
Published: (2024)
Analysis of the QC/QA survey
by: Rojas, Ricardo
Published: (2007)
by: Rojas, Ricardo
Published: (2007)
Similar Items
-
Neural at ArchEHR-QA 2026: One Method Fits All: Unified Prompt Optimization for Clinical QA over EHRs
by: Majeedi, Abrar, et al.
Published: (2026) -
sebis at ArchEHR-QA 2026: How Much Can You Do Locally? Evaluating Grounded EHR QA on a Single Notebook
by: Yurt, Ibrahim Ebrar, et al.
Published: (2026) -
UTSA-NLP at ArchEHR-QA 2025: Improving EHR Question Answering via Self-Consistency Prompting
by: Shields-Menard, Sara, et al.
Published: (2025) -
Neural at ArchEHR-QA 2025: Agentic Prompt Optimization for Evidence-Grounded Clinical Question Answering
by: Bogireddy, Sai Prasanna Teja Reddy, et al.
Published: (2025) -
Yale-DM-Lab at ArchEHR-QA 2026: Deterministic Grounding and Multi-Pass Evidence Alignment for EHR Question Answering
by: Irankhah, Elyas, et al.
Published: (2026)