Towards Leveraging Large Language Models for Automated Medical Q&A Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Krolik, Jack, Mahal, Herprit, Ahmad, Feroz, Trivedi, Gaurav, Saket, Bahador |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Method Validation of Large Language Model Medical Translation Across High- and Low-Resource Languages
di: Anyaegbuna, Chukwuebuka, et al.
Pubblicazione: (2026)
di: Anyaegbuna, Chukwuebuka, et al.
Pubblicazione: (2026)
The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes
di: Alba, Charles, et al.
Pubblicazione: (2024)
di: Alba, Charles, et al.
Pubblicazione: (2024)
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
di: Ketir, Si-Belkacem Yamine, et al.
Pubblicazione: (2026)
di: Ketir, Si-Belkacem Yamine, et al.
Pubblicazione: (2026)
A Method for the Architecture of a Medical Vertical Large Language Model Based on Deepseek R1
di: Zhang, Mingda, et al.
Pubblicazione: (2025)
di: Zhang, Mingda, et al.
Pubblicazione: (2025)
Survey and Experiments on Mental Disorder Detection via Social Media: From Large Language Models and RAG to Agents
di: Ge, Zhuohan, et al.
Pubblicazione: (2025)
di: Ge, Zhuohan, et al.
Pubblicazione: (2025)
Evaluating Large Language Models for IUCN Red List Species Information
di: Uryu, Shinya
Pubblicazione: (2025)
di: Uryu, Shinya
Pubblicazione: (2025)
MedPI: Evaluating AI Systems in Medical Patient-facing Interactions
di: V., Diego Fajardo, et al.
Pubblicazione: (2025)
di: V., Diego Fajardo, et al.
Pubblicazione: (2025)
Improving Drug Identification in Overdose Death Surveillance using Large Language Models
di: Funnell, Arthur J., et al.
Pubblicazione: (2025)
di: Funnell, Arthur J., et al.
Pubblicazione: (2025)
Evaluating the Challenges of LLMs in Real-world Medical Follow-up: A Comparative Study and An Optimized Framework
di: Liu, Jinyan, et al.
Pubblicazione: (2025)
di: Liu, Jinyan, et al.
Pubblicazione: (2025)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
di: Sasse, Kuleen, et al.
Pubblicazione: (2024)
Shallow Robustness, Deep Vulnerabilities: Multi-Turn Evaluation of Medical LLMs
di: Manczak, Blazej, et al.
Pubblicazione: (2025)
di: Manczak, Blazej, et al.
Pubblicazione: (2025)
The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
di: Ahmed, Md Shamim, et al.
Pubblicazione: (2026)
Integrating clinical reasoning into large language model-based diagnosis through etiology-aware attention steering
di: Li, Peixian, et al.
Pubblicazione: (2025)
di: Li, Peixian, et al.
Pubblicazione: (2025)
Forgotten Words: Benchmarking NeoBERT for Dementia Detection in Low-Resource Conversational Filipino and English Speech
di: Floresca, Rez Samantha Z., et al.
Pubblicazione: (2026)
di: Floresca, Rez Samantha Z., et al.
Pubblicazione: (2026)
An Interpretable Ensemble of Graph and Language Models for Improving Search Relevance in E-Commerce
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
di: Choudhary, Nurendra, et al.
Pubblicazione: (2024)
Language Models and Retrieval Augmented Generation for Automated Structured Data Extraction from Diagnostic Reports
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2024)
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2024)
ECG-Byte: A Tokenizer for End-to-End Generative Electrocardiogram Language Modeling
di: Han, William, et al.
Pubblicazione: (2024)
di: Han, William, et al.
Pubblicazione: (2024)
Igea: a Decoder-Only Language Model for Biomedical Text Generation in Italian
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2024)
di: Buonocore, Tommaso Mario, et al.
Pubblicazione: (2024)
ELMTEX: Fine-Tuning Large Language Models for Structured Clinical Information Extraction. A Case Study on Clinical Reports
di: Guluzade, Aynur, et al.
Pubblicazione: (2025)
di: Guluzade, Aynur, et al.
Pubblicazione: (2025)
CuraView: A Multi-Agent Framework for Medical Hallucination Detection with GraphRAG-Enhanced Knowledge Verification
di: Ye, Severin, et al.
Pubblicazione: (2026)
di: Ye, Severin, et al.
Pubblicazione: (2026)
Transparency-First Medical Language Models: Datasheets, Model Cards, and End-to-End Data Provenance for Clinical NLP
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
di: Imanov, Olaf Yunus Laitinen, et al.
Pubblicazione: (2026)
AcuityBench: Evaluating Clinical Acuity Identification and Uncertainty Alignment
di: Linzmayer, Robin, et al.
Pubblicazione: (2026)
di: Linzmayer, Robin, et al.
Pubblicazione: (2026)
Counterfactual Causal Inference in Natural Language with Large Language Models
di: Gendron, Gaël, et al.
Pubblicazione: (2024)
di: Gendron, Gaël, et al.
Pubblicazione: (2024)
Performance of Large Language Models in Supporting Medical Diagnosis and Treatment
di: Sousa, Diogo, et al.
Pubblicazione: (2025)
di: Sousa, Diogo, et al.
Pubblicazione: (2025)
VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation
di: Kiet, Huynh Trung, et al.
Pubblicazione: (2026)
di: Kiet, Huynh Trung, et al.
Pubblicazione: (2026)
Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters
di: Shah, Aaryan, et al.
Pubblicazione: (2026)
di: Shah, Aaryan, et al.
Pubblicazione: (2026)
BALI: Enhancing Biomedical Language Representations through Knowledge Graph and Language Model Alignment
di: Sakhovskiy, Andrey, et al.
Pubblicazione: (2025)
di: Sakhovskiy, Andrey, et al.
Pubblicazione: (2025)
AI-MASLD Metabolic Dysfunction and Information Steatosis of Large Language Models in Unstructured Clinical Narratives
di: Shen, Yuan, et al.
Pubblicazione: (2025)
di: Shen, Yuan, et al.
Pubblicazione: (2025)
BLT: Can Large Language Models Handle Basic Legal Text?
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2023)
di: Blair-Stanek, Andrew, et al.
Pubblicazione: (2023)
Agentic Automation of BT-RADS Scoring: End-to-End Multi-Agent System for Standardized Brain Tumor Follow-up Assessment
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2026)
di: Jabal, Mohamed Sobhi, et al.
Pubblicazione: (2026)
Continuous Predictive Modeling of Clinical Notes and ICD Codes in Patient Health Records
di: Caralt, Mireia Hernandez, et al.
Pubblicazione: (2024)
di: Caralt, Mireia Hernandez, et al.
Pubblicazione: (2024)
Comparative Analysis of LoRA-Adapted Embedding Models for Clinical Cardiology Text Representation
di: Young, Richard J., et al.
Pubblicazione: (2025)
di: Young, Richard J., et al.
Pubblicazione: (2025)
Chronic pain patient narratives allow for the estimation of current pain intensity
di: Nunes, Diogo A. P., et al.
Pubblicazione: (2022)
di: Nunes, Diogo A. P., et al.
Pubblicazione: (2022)
Comparative Study of Large Language Models on Chinese Film Script Continuation: An Empirical Analysis Based on GPT-5.2 and Qwen-Max
di: Cao, Yuxuan, et al.
Pubblicazione: (2026)
di: Cao, Yuxuan, et al.
Pubblicazione: (2026)
ChemPro: A Progressive Chemistry Benchmark for Large Language Models
di: Baranwal, Aaditya, et al.
Pubblicazione: (2026)
di: Baranwal, Aaditya, et al.
Pubblicazione: (2026)
The Table of Media Bias Elements: A sentence-level taxonomy of media bias types and propaganda techniques
di: Menzner, Tim, et al.
Pubblicazione: (2026)
di: Menzner, Tim, et al.
Pubblicazione: (2026)
Towards Greater Leverage: Scaling Laws for Efficient Mixture-of-Experts Language Models
di: Tian, Changxin, et al.
Pubblicazione: (2025)
di: Tian, Changxin, et al.
Pubblicazione: (2025)
Leveraging Large Language Models for Semantic Query Processing in a Scholarly Knowledge Graph
di: Jia, Runsong, et al.
Pubblicazione: (2024)
di: Jia, Runsong, et al.
Pubblicazione: (2024)
The Representational Alignment between Humans and Language Models is implicitly driven by a Concreteness Effect
di: Iaia, Cosimo, et al.
Pubblicazione: (2025)
di: Iaia, Cosimo, et al.
Pubblicazione: (2025)
A Multimodal Pipeline for Clinical Data Extraction: Applying Vision-Language Models to Scans of Transfusion Reaction Reports
di: Schäfer, Henning, et al.
Pubblicazione: (2025)
di: Schäfer, Henning, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Multi-Method Validation of Large Language Model Medical Translation Across High- and Low-Resource Languages
di: Anyaegbuna, Chukwuebuka, et al.
Pubblicazione: (2026) -
The Foundational Capabilities of Large Language Models in Predicting Postoperative Risks Using Clinical Notes
di: Alba, Charles, et al.
Pubblicazione: (2024) -
Can Large Language Models Imitate Human Speech for Clinical Assessment? LLM-Driven Data Augmentation for Cognitive Score Prediction
di: Ketir, Si-Belkacem Yamine, et al.
Pubblicazione: (2026) -
A Method for the Architecture of a Medical Vertical Large Language Model Based on Deepseek R1
di: Zhang, Mingda, et al.
Pubblicazione: (2025) -
Survey and Experiments on Mental Disorder Detection via Social Media: From Large Language Models and RAG to Agents
di: Ge, Zhuohan, et al.
Pubblicazione: (2025)