Quantifying Hallucinations in Language Language Models on Medical Textbooks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Colelough, Brandon C., Bartels, Davis, Demner-Fushman, Dina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Overview of the ClinIQLink 2025 Shared Task on Medical Question-Answering
von: Colelough, Brandon, et al.
Veröffentlicht: (2025)
von: Colelough, Brandon, et al.
Veröffentlicht: (2025)
Automated Evaluation can Distinguish the Good and Bad AI Responses to Patient Questions about Hospitalization
von: Soni, Sarvesh, et al.
Veröffentlicht: (2025)
von: Soni, Sarvesh, et al.
Veröffentlicht: (2025)
Toward Relieving Clinician Burden by Automatically Generating Progress Notes using Interim Hospital Data
von: Soni, Sarvesh, et al.
Veröffentlicht: (2024)
von: Soni, Sarvesh, et al.
Veröffentlicht: (2024)
BioACE: An Automated Framework for Biomedical Answer and Citation Evaluations
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
von: Gupta, Deepak, et al.
Veröffentlicht: (2026)
Lessons from the TREC Plain Language Adaptation of Biomedical Abstracts (PLABA) track
von: Ondov, Brian, et al.
Veröffentlicht: (2025)
von: Ondov, Brian, et al.
Veröffentlicht: (2025)
A Dataset for Addressing Patient's Information Needs related to Clinical Course of Hospitalization
von: Soni, Sarvesh, et al.
Veröffentlicht: (2025)
von: Soni, Sarvesh, et al.
Veröffentlicht: (2025)
Overview of TREC 2024 Medical Video Question Answering (MedVidQA) Track
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
von: Gupta, Deepak, et al.
Veröffentlicht: (2024)
A Dataset and Benchmark for Consumer Healthcare Question Summarization
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
von: Basu, Abhishek, et al.
Veröffentlicht: (2025)
JEBS: A Fine-grained Biomedical Lexical Simplification Task
von: Xia, William, et al.
Veröffentlicht: (2025)
von: Xia, William, et al.
Veröffentlicht: (2025)
Credal Transformer: A Principled Approach for Quantifying and Mitigating Hallucinations in Large Language Models
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
von: Ji, Shihao, et al.
Veröffentlicht: (2025)
Reducing Hallucinations of Medical Multimodal Large Language Models with Visual Retrieval-Augmented Generation
von: Chu, Yun-Wei, et al.
Veröffentlicht: (2025)
von: Chu, Yun-Wei, et al.
Veröffentlicht: (2025)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
von: Zhang, Yue, et al.
Veröffentlicht: (2023)
A Dataset and Resources for Identifying Patient Health Literacy Information from Clinical Notes
von: Bittner, Madeline, et al.
Veröffentlicht: (2026)
von: Bittner, Madeline, et al.
Veröffentlicht: (2026)
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
von: Zuo, Kaiwen, et al.
Veröffentlicht: (2024)
von: Zuo, Kaiwen, et al.
Veröffentlicht: (2024)
Calibrated Language Models Must Hallucinate
von: Kalai, Adam Tauman, et al.
Veröffentlicht: (2023)
von: Kalai, Adam Tauman, et al.
Veröffentlicht: (2023)
Hallucination Detection with Small Language Models
von: Cheung, Ming
Veröffentlicht: (2025)
von: Cheung, Ming
Veröffentlicht: (2025)
Benchmarking Deflection and Hallucination in Large Vision-Language Models
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
von: Moratelli, Nicholas, et al.
Veröffentlicht: (2026)
Quantifying Semantic Emergence in Language Models
von: Chen, Hang, et al.
Veröffentlicht: (2024)
von: Chen, Hang, et al.
Veröffentlicht: (2024)
An Evolutionary Large Language Model for Hallucination Mitigation
von: Boulesnane, Abdennour, et al.
Veröffentlicht: (2024)
von: Boulesnane, Abdennour, et al.
Veröffentlicht: (2024)
ANAH: Analytical Annotation of Hallucinations in Large Language Models
von: Ji, Ziwei, et al.
Veröffentlicht: (2024)
von: Ji, Ziwei, et al.
Veröffentlicht: (2024)
Confabulation: The Surprising Value of Large Language Model Hallucinations
von: Sui, Peiqi, et al.
Veröffentlicht: (2024)
von: Sui, Peiqi, et al.
Veröffentlicht: (2024)
Copy-Paste to Mitigate Large Language Model Hallucinations
von: Long, Yongchao, et al.
Veröffentlicht: (2025)
von: Long, Yongchao, et al.
Veröffentlicht: (2025)
The Impact of Negated Text on Hallucination with Large Language Models
von: Seo, Jaehyung, et al.
Veröffentlicht: (2025)
von: Seo, Jaehyung, et al.
Veröffentlicht: (2025)
Theoretical Foundations and Mitigation of Hallucination in Large Language Models
von: Gumaan, Esmail
Veröffentlicht: (2025)
von: Gumaan, Esmail
Veröffentlicht: (2025)
Advancing Frontiers in SLAM: A Survey of Symbolic Representation and Human-Machine Teaming in Environmental Mapping
von: Colelough, Brandon Curtis
Veröffentlicht: (2024)
von: Colelough, Brandon Curtis
Veröffentlicht: (2024)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
von: Pandit, Shrey, et al.
Veröffentlicht: (2025)
Evaluating Multimodal Large Language Models on Educational Textbook Question Answering
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
von: Alawwad, Hessa A., et al.
Veröffentlicht: (2025)
Triggering Hallucinations in LLMs: A Quantitative Study of Prompt-Induced Hallucination in Large Language Models
von: Sato, Makoto
Veröffentlicht: (2025)
von: Sato, Makoto
Veröffentlicht: (2025)
How Large Language Models are Designed to Hallucinate
von: Ackermann, Richard, et al.
Veröffentlicht: (2025)
von: Ackermann, Richard, et al.
Veröffentlicht: (2025)
Augmenting Black-box LLMs with Medical Textbooks for Biomedical Question Answering
von: Wang, Yubo, et al.
Veröffentlicht: (2023)
von: Wang, Yubo, et al.
Veröffentlicht: (2023)
A Scoping Review of Natural Language Processing in Addressing Medically Inaccurate Information: Errors, Misinformation, and Hallucination
von: Sun, Zhaoyi, et al.
Veröffentlicht: (2025)
von: Sun, Zhaoyi, et al.
Veröffentlicht: (2025)
Beyond Facts: Evaluating Intent Hallucination in Large Language Models
von: Hao, Yijie, et al.
Veröffentlicht: (2025)
von: Hao, Yijie, et al.
Veröffentlicht: (2025)
HALO: An Ontology for Representing and Categorizing Hallucinations in Large Language Models
von: Nananukul, Navapat, et al.
Veröffentlicht: (2023)
von: Nananukul, Navapat, et al.
Veröffentlicht: (2023)
Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations
von: Yu, Lei, et al.
Veröffentlicht: (2024)
von: Yu, Lei, et al.
Veröffentlicht: (2024)
Reference-free Hallucination Detection for Large Vision-Language Models
von: Li, Qing, et al.
Veröffentlicht: (2024)
von: Li, Qing, et al.
Veröffentlicht: (2024)
Investigating Hallucinations in Pruned Large Language Models for Abstractive Summarization
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
von: Chrysostomou, George, et al.
Veröffentlicht: (2023)
Neural Probe-Based Hallucination Detection for Large Language Models
von: Liang, Shize, et al.
Veröffentlicht: (2025)
von: Liang, Shize, et al.
Veröffentlicht: (2025)
Counterfactual Probing for Hallucination Detection and Mitigation in Large Language Models
von: Feng, Yijun
Veröffentlicht: (2025)
von: Feng, Yijun
Veröffentlicht: (2025)
HICD: Hallucination-Inducing via Attention Dispersion for Contrastive Decoding to Mitigate Hallucinations in Large Language Models
von: Jiang, Xinyan, et al.
Veröffentlicht: (2025)
von: Jiang, Xinyan, et al.
Veröffentlicht: (2025)
Medical Hallucinations in Foundation Models and Their Impact on Healthcare
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
von: Kim, Yubin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Overview of the ClinIQLink 2025 Shared Task on Medical Question-Answering
von: Colelough, Brandon, et al.
Veröffentlicht: (2025) -
Automated Evaluation can Distinguish the Good and Bad AI Responses to Patient Questions about Hospitalization
von: Soni, Sarvesh, et al.
Veröffentlicht: (2025) -
Toward Relieving Clinician Burden by Automatically Generating Progress Notes using Interim Hospital Data
von: Soni, Sarvesh, et al.
Veröffentlicht: (2024) -
BioACE: An Automated Framework for Biomedical Answer and Citation Evaluations
von: Gupta, Deepak, et al.
Veröffentlicht: (2026) -
Lessons from the TREC Plain Language Adaptation of Biomedical Abstracts (PLABA) track
von: Ondov, Brian, et al.
Veröffentlicht: (2025)