MedHalu: Hallucinations in Responses to Healthcare Queries by Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Agarwal, Vibhor, Jin, Yiqiao, Chandra, Mohit, De Choudhury, Munmun, Kumar, Srijan, Sastry, Nishanth |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Conversation Kernels: A Flexible Mechanism to Learn Relevant Context for Online Conversation Understanding
by: Agarwal, Vibhor, et al.
Published: (2025)
by: Agarwal, Vibhor, et al.
Published: (2025)
A Framework for Situating Innovations, Opportunities, and Challenges in Advancing Vertical Systems with Large AI Models
by: Verma, Gaurav, et al.
Published: (2025)
by: Verma, Gaurav, et al.
Published: (2025)
CodeMirage: Hallucinations in Code Generated by Large Language Models
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
Reasoning Is Not All You Need: Examining LLMs for Multi-Turn Mental Health Conversations
by: Chandra, Mohit, et al.
Published: (2025)
by: Chandra, Mohit, et al.
Published: (2025)
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
by: Chen, Kedi, et al.
Published: (2024)
by: Chen, Kedi, et al.
Published: (2024)
Halu-J: Critique-Based Hallucination Judge
by: Wang, Binjie, et al.
Published: (2024)
by: Wang, Binjie, et al.
Published: (2024)
Understanding the Humans Behind Online Misinformation: An Observational Study Through the Lens of the COVID-19 Pandemic
by: Chandra, Mohit, et al.
Published: (2023)
by: Chandra, Mohit, et al.
Published: (2023)
Linguistic Comparison of AI- and Human-Written Responses to Online Mental Health Queries
by: Saha, Koustuv, et al.
Published: (2025)
by: Saha, Koustuv, et al.
Published: (2025)
Sysformer: Safeguarding Frozen Large Language Models with Adaptive System Prompts
by: Sharma, Kartik, et al.
Published: (2025)
by: Sharma, Kartik, et al.
Published: (2025)
UniGuard: Towards Universal Safety Guardrails for Jailbreak Attacks on Multimodal Large Language Models
by: Oh, Sejoon, et al.
Published: (2024)
by: Oh, Sejoon, et al.
Published: (2024)
Investigating Symbolic Triggers of Hallucination in Gemma Models Across HaluEval and TruthfulQA
by: Lamba, Naveen, et al.
Published: (2025)
by: Lamba, Naveen, et al.
Published: (2025)
Efficient Knowledge Probing of Large Language Models by Adapting Pre-trained Embeddings
by: Sharma, Kartik, et al.
Published: (2025)
by: Sharma, Kartik, et al.
Published: (2025)
Do Large Language Models Align with Core Mental Health Counseling Competencies?
by: Nguyen, Viet Cuong, et al.
Published: (2024)
by: Nguyen, Viet Cuong, et al.
Published: (2024)
SymLoc: Symbolic Localization of Hallucination across HaluEval and TruthfulQA
by: Lamba, Naveen, et al.
Published: (2025)
by: Lamba, Naveen, et al.
Published: (2025)
HaluEval-Wild: Evaluating Hallucinations of Language Models in the Wild
by: Zhu, Zhiying, et al.
Published: (2024)
by: Zhu, Zhiying, et al.
Published: (2024)
Decentralised Moderation for Interoperable Social Networks: A Conversation-based Approach for Pleroma and the Fediverse
by: Agarwal, Vibhor, et al.
Published: (2024)
by: Agarwal, Vibhor, et al.
Published: (2024)
MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning
by: Huang, Ruijun, et al.
Published: (2026)
by: Huang, Ruijun, et al.
Published: (2026)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
UniSD: Towards a Unified Self-Distillation Framework for Large Language Models
by: Jin, Yiqiao, et al.
Published: (2026)
by: Jin, Yiqiao, et al.
Published: (2026)
MM-Soc: Benchmarking Multimodal Large Language Models in Social Media Platforms
by: Jin, Yiqiao, et al.
Published: (2024)
by: Jin, Yiqiao, et al.
Published: (2024)
MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models
by: Zuo, Kaiwen, et al.
Published: (2024)
by: Zuo, Kaiwen, et al.
Published: (2024)
Lived Experience Not Found: LLMs Struggle to Align with Experts on Addressing Adverse Drug Reactions from Psychiatric Medication Use
by: Chandra, Mohit, et al.
Published: (2024)
by: Chandra, Mohit, et al.
Published: (2024)
MedHallu: A Comprehensive Benchmark for Detecting Medical Hallucinations in Large Language Models
by: Pandit, Shrey, et al.
Published: (2025)
by: Pandit, Shrey, et al.
Published: (2025)
HaluMem: Evaluating Hallucinations in Memory Systems of Agents
by: Chen, Ding, et al.
Published: (2025)
by: Chen, Ding, et al.
Published: (2025)
Query-Efficient Planning with Language Models
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2024)
by: Gonzalez-Pumariega, Gonzalo, et al.
Published: (2024)
Evaluating Large Language Models for Health-related Queries with Presuppositions
by: Kaur, Navreet, et al.
Published: (2023)
by: Kaur, Navreet, et al.
Published: (2023)
A Risk Taxonomy and Reflection Tool for Large Language Model Adoption in Public Health
by: Zhou, Jiawei, et al.
Published: (2024)
by: Zhou, Jiawei, et al.
Published: (2024)
The Typing Cure: Experiences with Large Language Model Chatbots for Mental Health Support
by: Song, Inhwa, et al.
Published: (2024)
by: Song, Inhwa, et al.
Published: (2024)
SARA: Selective and Adaptive Retrieval-augmented Generation with Context Compression
by: Jin, Yiqiao, et al.
Published: (2025)
by: Jin, Yiqiao, et al.
Published: (2025)
An Investigation on Group Query Hallucination Attacks
by: Miao, Kehao, et al.
Published: (2025)
by: Miao, Kehao, et al.
Published: (2025)
Ethical Reasoning and Moral Value Alignment of LLMs Depend on the Language we Prompt them in
by: Agarwal, Utkarsh, et al.
Published: (2024)
by: Agarwal, Utkarsh, et al.
Published: (2024)
Do Moral Judgment and Reasoning Capability of LLMs Change with Language? A Study using the Multilingual Defining Issues Test
by: Khandelwal, Aditi, et al.
Published: (2024)
by: Khandelwal, Aditi, et al.
Published: (2024)
Alleviating Hallucinations of Large Language Models through Induced Hallucinations
by: Zhang, Yue, et al.
Published: (2023)
by: Zhang, Yue, et al.
Published: (2023)
Towards Experience-Centered AI: A Framework for Integrating Lived Experience in Design and Development
by: Gautam, Sanjana, et al.
Published: (2025)
by: Gautam, Sanjana, et al.
Published: (2025)
Lightweight Large Language Model for Medication Enquiry: Med-Pal
by: Elangovan, Kabilan, et al.
Published: (2024)
by: Elangovan, Kabilan, et al.
Published: (2024)
A Community-Centric Perspective for Characterizing and Detecting Anti-Asian Violence-Provoking Speech
by: Verma, Gaurav, et al.
Published: (2024)
by: Verma, Gaurav, et al.
Published: (2024)
ExploreSelf: Fostering User-driven Exploration and Reflection on Personal Challenges with Adaptive Guidance by Large Language Models
by: Song, Inhwa, et al.
Published: (2024)
by: Song, Inhwa, et al.
Published: (2024)
An Evolutionary Large Language Model for Hallucination Mitigation
by: Boulesnane, Abdennour, et al.
Published: (2024)
by: Boulesnane, Abdennour, et al.
Published: (2024)
MedCalc-Bench: Evaluating Large Language Models for Medical Calculations
by: Khandekar, Nikhil, et al.
Published: (2024)
by: Khandekar, Nikhil, et al.
Published: (2024)
MedSumm: A Multimodal Approach to Summarizing Code-Mixed Hindi-English Clinical Queries
by: Ghosh, Akash, et al.
Published: (2024)
by: Ghosh, Akash, et al.
Published: (2024)
Similar Items
-
Conversation Kernels: A Flexible Mechanism to Learn Relevant Context for Online Conversation Understanding
by: Agarwal, Vibhor, et al.
Published: (2025) -
A Framework for Situating Innovations, Opportunities, and Challenges in Advancing Vertical Systems with Large AI Models
by: Verma, Gaurav, et al.
Published: (2025) -
CodeMirage: Hallucinations in Code Generated by Large Language Models
by: Agarwal, Vibhor, et al.
Published: (2024) -
Reasoning Is Not All You Need: Examining LLMs for Multi-Turn Mental Health Conversations
by: Chandra, Mohit, et al.
Published: (2025) -
DiaHalu: A Dialogue-level Hallucination Evaluation Benchmark for Large Language Models
by: Chen, Kedi, et al.
Published: (2024)