Introducing Answered with Evidence -- a framework for evaluating whether LLM responses to biomedical questions are founded in evidence
Fuente:
arXiv
Saved in:
| Main Authors: | Baldwin, Julian D, Dinh, Christina, Mukerji, Arjun, Sanghavi, Neil, Gombar, Saurabh |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Answering real-world clinical questions using large language model based systems
by: Low, Yen Sia, et al.
Published: (2024)
by: Low, Yen Sia, et al.
Published: (2024)
LLM-PQA: LLM-enhanced Prediction Query Answering
by: Li, Ziyu, et al.
Published: (2024)
by: Li, Ziyu, et al.
Published: (2024)
RAG based Question-Answering for Contextual Response Prediction System
by: Veturi, Sriram, et al.
Published: (2024)
by: Veturi, Sriram, et al.
Published: (2024)
Generalized knowledge-enhanced framework for biomedical entity and relation extraction
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
LLM-Assisted Multi-Teacher Continual Learning for Visual Question Answering in Robotic Surgery
by: Du, Yuyang, et al.
Published: (2024)
by: Du, Yuyang, et al.
Published: (2024)
Augmenting Researchy Questions with Sub-question Judgments
by: Ju, Jia-Huei, et al.
Published: (2025)
by: Ju, Jia-Huei, et al.
Published: (2025)
Introducing Semantic Capability in LinkedIn's Content Search Engine
by: Yang, Xin, et al.
Published: (2024)
by: Yang, Xin, et al.
Published: (2024)
Rethinking Hybrid Retrieval: When Small Embeddings and LLM Re-ranking Beat Bigger Models
by: Rao, Arjun, et al.
Published: (2025)
by: Rao, Arjun, et al.
Published: (2025)
Evaluating the Retrieval Component in LLM-Based Question Answering Systems
by: Alinejad, Ashkan, et al.
Published: (2024)
by: Alinejad, Ashkan, et al.
Published: (2024)
EviRerank: Adaptive Evidence Construction for Long-Document LLM Reranking
by: Li, Minghan, et al.
Published: (2024)
by: Li, Minghan, et al.
Published: (2024)
Enhancing Complex Question Answering over Knowledge Graphs through Evidence Pattern Retrieval
by: Ding, Wentao, et al.
Published: (2024)
by: Ding, Wentao, et al.
Published: (2024)
SE-PQA: Personalized Community Question Answering
by: Kasela, Pranav, et al.
Published: (2023)
by: Kasela, Pranav, et al.
Published: (2023)
Answer Retrieval in Legal Community Question Answering
by: Askari, Arian, et al.
Published: (2024)
by: Askari, Arian, et al.
Published: (2024)
A Human-in-the-Loop, LLM-Centered Architecture for Knowledge-Graph Question Answering
by: Pusch, Larissa, et al.
Published: (2026)
by: Pusch, Larissa, et al.
Published: (2026)
ResumeFlow: An LLM-facilitated Pipeline for Personalized Resume Generation and Refinement
by: Zinjad, Saurabh Bhausaheb, et al.
Published: (2024)
by: Zinjad, Saurabh Bhausaheb, et al.
Published: (2024)
Can we repurpose multiple-choice question-answering models to rerank retrieved documents?
by: Catapang, Jasper Kyle
Published: (2025)
by: Catapang, Jasper Kyle
Published: (2025)
MLPs are Efficient Distilled Generative Recommenders
by: Guo, Zitian, et al.
Published: (2026)
by: Guo, Zitian, et al.
Published: (2026)
Open-Ended and Knowledge-Intensive Video Question Answering
by: Alam, Md Zarif Ul, et al.
Published: (2025)
by: Alam, Md Zarif Ul, et al.
Published: (2025)
Vietnamese Legal Information Retrieval in Question-Answering System
by: Ba, Thiem Nguyen, et al.
Published: (2024)
by: Ba, Thiem Nguyen, et al.
Published: (2024)
An Empirical Study of Evaluating Long-form Question Answering
by: Xian, Ning, et al.
Published: (2025)
by: Xian, Ning, et al.
Published: (2025)
ZSE-Cap: A Zero-Shot Ensemble for Image Retrieval and Prompt-Guided Captioning
by: Dinh, Duc-Tai, et al.
Published: (2025)
by: Dinh, Duc-Tai, et al.
Published: (2025)
Overview of the MedHopQA track at BioCreative IX: track description, participation and evaluation of systems for multi-hop medical question answering
by: Islamaj, Rezarta, et al.
Published: (2026)
by: Islamaj, Rezarta, et al.
Published: (2026)
RARe: Retrieval Augmented Retrieval with In-Context Examples
by: Tejaswi, Atula, et al.
Published: (2024)
by: Tejaswi, Atula, et al.
Published: (2024)
Towards Robust Expert Finding in Community Question Answering Platforms
by: Amendola, Maddalena, et al.
Published: (2025)
by: Amendola, Maddalena, et al.
Published: (2025)
Zero-Shot Complex Question-Answering on Long Scientific Documents
by: Wang, Wanting
Published: (2025)
by: Wang, Wanting
Published: (2025)
Answering Multimodal Exclusion Queries with Lightweight Sparse Disentangled Representations
by: J, Prachi, et al.
Published: (2025)
by: J, Prachi, et al.
Published: (2025)
Options-Aware Dense Retrieval for Multiple-Choice query Answering
by: Singh, Manish, et al.
Published: (2025)
by: Singh, Manish, et al.
Published: (2025)
Unlocking the `Why' of Buying: Introducing a New Dataset and Benchmark for Purchase Reason and Post-Purchase Experience
by: Chen, Tao, et al.
Published: (2024)
by: Chen, Tao, et al.
Published: (2024)
Enhancing CTR Prediction through Sequential Recommendation Pre-training: Introducing the SRP4CTR Framework
by: Han, Ruidong, et al.
Published: (2024)
by: Han, Ruidong, et al.
Published: (2024)
JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability
by: Wang, Junda, et al.
Published: (2024)
by: Wang, Junda, et al.
Published: (2024)
ST-Raptor: LLM-Powered Semi-Structured Table Question Answering
by: Tang, Zirui, et al.
Published: (2025)
by: Tang, Zirui, et al.
Published: (2025)
Expressiveness Limits of Autoregressive Semantic ID Generation in Generative Recommendation
by: Hou, Yupeng, et al.
Published: (2026)
by: Hou, Yupeng, et al.
Published: (2026)
Évaluation des capacités de réponse de larges modèles de langage (LLM) pour des questions d'historiens
by: Chartier, Mathieu, et al.
Published: (2024)
by: Chartier, Mathieu, et al.
Published: (2024)
UQABench: Evaluating User Embedding for Prompting LLMs in Personalized Question Answering
by: Liu, Langming, et al.
Published: (2025)
by: Liu, Langming, et al.
Published: (2025)
Optimizing Retrieval-Augmented Generation with Elasticsearch for Enhanced Question-Answering Systems
by: Chen, Jiajing, et al.
Published: (2024)
by: Chen, Jiajing, et al.
Published: (2024)
Advancing continual lifelong learning in neural information retrieval: definition, dataset, framework, and empirical evaluation
by: Hou, Jingrui, et al.
Published: (2023)
by: Hou, Jingrui, et al.
Published: (2023)
Improving Health Question Answering with Reliable and Time-Aware Evidence Retrieval
by: Vladika, Juraj, et al.
Published: (2024)
by: Vladika, Juraj, et al.
Published: (2024)
NORMY: Non-Uniform History Modeling for Open Retrieval Conversational Question Answering
by: Rashid, Muhammad Shihab, et al.
Published: (2024)
by: Rashid, Muhammad Shihab, et al.
Published: (2024)
Synthetic Data Generation with Large Language Models for Personalized Community Question Answering
by: Braga, Marco, et al.
Published: (2024)
by: Braga, Marco, et al.
Published: (2024)
GeoRAG: A Question-Answering Approach from a Geographical Perspective
by: Wang, Jian, et al.
Published: (2025)
by: Wang, Jian, et al.
Published: (2025)
Similar Items
-
Answering real-world clinical questions using large language model based systems
by: Low, Yen Sia, et al.
Published: (2024) -
LLM-PQA: LLM-enhanced Prediction Query Answering
by: Li, Ziyu, et al.
Published: (2024) -
RAG based Question-Answering for Contextual Response Prediction System
by: Veturi, Sriram, et al.
Published: (2024) -
Generalized knowledge-enhanced framework for biomedical entity and relation extraction
by: Nguyen, Minh, et al.
Published: (2024) -
LLM-Assisted Multi-Teacher Continual Learning for Visual Question Answering in Robotic Surgery
by: Du, Yuyang, et al.
Published: (2024)