Harnessing Collective Intelligence of LLMs for Robust Biomedical QA: A Multi-Model Approach
Fuente:
arXiv
Saved in:
| Main Authors: | Panou, Dimitra, Dimopoulos, Alexandros C., Koubarakis, Manolis, Reczko, Martin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Transformer-based Language Models for Reasoning in the Description Logic ALCQ
by: Poulis, Angelos, et al.
Published: (2024)
by: Poulis, Angelos, et al.
Published: (2024)
The Large Language Model GreekLegalRoBERTa
by: Saketos, Vasileios, et al.
Published: (2024)
by: Saketos, Vasileios, et al.
Published: (2024)
Transformers in the Service of Description Logic-based Contexts
by: Poulis, Angelos, et al.
Published: (2023)
by: Poulis, Angelos, et al.
Published: (2023)
Συμμόρφωση με τον GDPR και το Ομοσπονδιακό EGA
by: Savakis, Charalambos, et al.
Published: (2026)
by: Savakis, Charalambos, et al.
Published: (2026)
DistillER: Knowledge Distillation in Entity Resolution with Large Language Models
by: Zeakis, Alexandros, et al.
Published: (2026)
by: Zeakis, Alexandros, et al.
Published: (2026)
UETQuintet at BioCreative IX -- MedHopQA: Enhancing Biomedical QA with Selective Multi-hop Reasoning and Contextual Retrieval
by: Nguyen, Quoc-An, et al.
Published: (2026)
by: Nguyen, Quoc-An, et al.
Published: (2026)
DeepRAG: Integrating Hierarchical Reasoning and Process Supervision for Biomedical Multi-Hop QA
by: Ji, Yuelyu, et al.
Published: (2025)
by: Ji, Yuelyu, et al.
Published: (2025)
Beyond Retrieval: Ensembling Cross-Encoders and GPT Rerankers with LLMs for Biomedical QA
by: Verma, Shashank, et al.
Published: (2025)
by: Verma, Shashank, et al.
Published: (2025)
BIT.UA-AAUBS at ArchEHR-QA 2026: Evaluating Open-Source and Proprietary LLMs via Prompting in Low-Resource QA
by: Jonker, Richard A. A., et al.
Published: (2026)
by: Jonker, Richard A. A., et al.
Published: (2026)
AdvisorQA: Towards Helpful and Harmless Advice-seeking Question Answering with Collective Intelligence
by: Kim, Minbeom, et al.
Published: (2024)
by: Kim, Minbeom, et al.
Published: (2024)
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
by: Bhattarai, Kriti, et al.
Published: (2026)
by: Bhattarai, Kriti, et al.
Published: (2026)
Integrating Domain Knowledge for Financial QA: A Multi-Retriever RAG Approach with LLMs
by: Zhang, Yukun, et al.
Published: (2025)
by: Zhang, Yukun, et al.
Published: (2025)
QA-LIGN: Aligning LLMs through Constitutionally Decomposed QA
by: Dineen, Jacob, et al.
Published: (2025)
by: Dineen, Jacob, et al.
Published: (2025)
TerraQ: Spatiotemporal Question-Answering on Satellite Image Archives
by: Kefalidis, Sergios-Anestis, et al.
Published: (2025)
by: Kefalidis, Sergios-Anestis, et al.
Published: (2025)
Harnessing Large Language Models for Biomedical Named Entity Recognition
by: Chen, Jian, et al.
Published: (2025)
by: Chen, Jian, et al.
Published: (2025)
ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
by: Zong, Qing, et al.
Published: (2024)
by: Zong, Qing, et al.
Published: (2024)
Enhancing Robustness in Biomedical NLI Models: A Probing Approach for Clinical Trials
by: Mustafa, Ata
Published: (2024)
by: Mustafa, Ata
Published: (2024)
Harnessing RLHF for Robust Unanswerability Recognition and Trustworthy Response Generation in LLMs
by: Lin, Shuyuan, et al.
Published: (2025)
by: Lin, Shuyuan, et al.
Published: (2025)
DocHop-QA: Towards Multi-Hop Reasoning over Multimodal Document Collections
by: Park, Jiwon, et al.
Published: (2025)
by: Park, Jiwon, et al.
Published: (2025)
Evaluating Prompt Engineering Techniques for RAG in Small Language Models: A Multi-Hop QA Approach
by: Mohammadi, Amir Hossein, et al.
Published: (2026)
by: Mohammadi, Amir Hossein, et al.
Published: (2026)
Contextual Breach: Assessing the Robustness of Transformer-based QA Models
by: Saadat, Asir, et al.
Published: (2024)
by: Saadat, Asir, et al.
Published: (2024)
On-the-fly Definition Augmentation of LLMs for Biomedical NER
by: Munnangi, Monica, et al.
Published: (2024)
by: Munnangi, Monica, et al.
Published: (2024)
P-RAG: Prompt-Enhanced Parametric RAG with LoRA and Selective CoT for Biomedical and Multi-Hop QA
by: Lyu, Xingda, et al.
Published: (2026)
by: Lyu, Xingda, et al.
Published: (2026)
BioMedSearch: A Multi-Source Biomedical Retrieval Framework Based on LLMs
by: Liu, Congying, et al.
Published: (2025)
by: Liu, Congying, et al.
Published: (2025)
A Bayesian Approach to Harnessing the Power of LLMs in Authorship Attribution
by: Hu, Zhengmian, et al.
Published: (2024)
by: Hu, Zhengmian, et al.
Published: (2024)
Compound-QA: A Benchmark for Evaluating LLMs on Compound Questions
by: Hou, Yutao, et al.
Published: (2024)
by: Hou, Yutao, et al.
Published: (2024)
Streamlining Biomedical Research with Specialized LLMs
by: Chen, Linqing, et al.
Published: (2025)
by: Chen, Linqing, et al.
Published: (2025)
M3SciQA: A Multi-Modal Multi-Document Scientific QA Benchmark for Evaluating Foundation Models
by: Li, Chuhan, et al.
Published: (2024)
by: Li, Chuhan, et al.
Published: (2024)
Navigating Large-Scale Document Collections: MuDABench for Multi-Document Analytical QA
by: Li, Zhanli, et al.
Published: (2026)
by: Li, Zhanli, et al.
Published: (2026)
DisasterQA: A Benchmark for Assessing the performance of LLMs in Disaster Response
by: Rawat, Rajat
Published: (2024)
by: Rawat, Rajat
Published: (2024)
Robust Knowledge Editing via Explicit Reasoning Chains for Distractor-Resilient Multi-Hop QA
by: Wu, Yuchen, et al.
Published: (2025)
by: Wu, Yuchen, et al.
Published: (2025)
RefactorCoderQA: Benchmarking LLMs for Multi-Domain Coding Question Solutions in Cloud and Edge Deployment
by: Rahman, Shadikur, et al.
Published: (2025)
by: Rahman, Shadikur, et al.
Published: (2025)
Do LLMs Surpass Encoders for Biomedical NER?
by: Obeidat, Motasem S, et al.
Published: (2025)
by: Obeidat, Motasem S, et al.
Published: (2025)
MimeQA: Towards Socially-Intelligent Nonverbal Foundation Models
by: Li, Hengzhi, et al.
Published: (2025)
by: Li, Hengzhi, et al.
Published: (2025)
Benchmarking LLMs for Pairwise Causal Discovery in Biomedical and Multi-Domain Contexts
by: Anuyah, Sydney, et al.
Published: (2026)
by: Anuyah, Sydney, et al.
Published: (2026)
How Robust are the Tabular QA Models for Scientific Tables? A Study using Customized Dataset
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
Instruction Tuning and CoT Prompting for Contextual Medical QA with LLMs
by: Le, Chenqian, et al.
Published: (2025)
by: Le, Chenqian, et al.
Published: (2025)
Factuality or Fiction? Benchmarking Modern LLMs on Ambiguous QA with Citations
by: Patel, Maya, et al.
Published: (2024)
by: Patel, Maya, et al.
Published: (2024)
RAG-BioQA: A Retrieval-Augmented Generation Framework for Long-Form Biomedical Question Answering
by: Panchumarthi, Lovely Yeswanth, et al.
Published: (2025)
by: Panchumarthi, Lovely Yeswanth, et al.
Published: (2025)
Improving Automatic Evaluation of Large Language Models (LLMs) in Biomedical Relation Extraction via LLMs-as-the-Judge
by: Laskar, Md Tahmid Rahman, et al.
Published: (2025)
by: Laskar, Md Tahmid Rahman, et al.
Published: (2025)
Similar Items
-
Transformer-based Language Models for Reasoning in the Description Logic ALCQ
by: Poulis, Angelos, et al.
Published: (2024) -
The Large Language Model GreekLegalRoBERTa
by: Saketos, Vasileios, et al.
Published: (2024) -
Transformers in the Service of Description Logic-based Contexts
by: Poulis, Angelos, et al.
Published: (2023) -
Συμμόρφωση με τον GDPR και το Ομοσπονδιακό EGA
by: Savakis, Charalambos, et al.
Published: (2026) -
DistillER: Knowledge Distillation in Entity Resolution with Large Language Models
by: Zeakis, Alexandros, et al.
Published: (2026)