The Biased Oracle: Assessing LLMs' Understandability and Empathy in Medical Diagnoses
Fuente:
arXiv
Guardado en:
| Autores principales: | Yao, Jianzhou, Liu, Shunchang, Drui, Guillaume, Pettersson, Rikard, Blasimme, Alessandro, Kijewski, Sara |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs
por: Liu, Yang, et al.
Publicado: (2025)
por: Liu, Yang, et al.
Publicado: (2025)
DiversityMedQA: Assessing Demographic Biases in Medical Diagnosis using Large Language Models
por: Rawat, Rajat, et al.
Publicado: (2024)
por: Rawat, Rajat, et al.
Publicado: (2024)
LLMs are Biased Evaluators But Not Biased for Retrieval Augmented Generation
por: Chen, Yen-Shan, et al.
Publicado: (2024)
por: Chen, Yen-Shan, et al.
Publicado: (2024)
Flattery, Fluff, and Fog: Diagnosing and Mitigating Idiosyncratic Biases in Preference Models
por: Bharadwaj, Anirudh, et al.
Publicado: (2025)
por: Bharadwaj, Anirudh, et al.
Publicado: (2025)
The Emotional Spectrum of LLMs: Leveraging Empathy and Emotion-Based Markers for Mental Health Support
por: De Grandi, Alessandro, et al.
Publicado: (2024)
por: De Grandi, Alessandro, et al.
Publicado: (2024)
ProSA: Assessing and Understanding the Prompt Sensitivity of LLMs
por: Zhuo, Jingming, et al.
Publicado: (2024)
por: Zhuo, Jingming, et al.
Publicado: (2024)
Mixed Signals: Understanding Model Disagreement in Multimodal Empathy Detection
por: Srikanth, Maya, et al.
Publicado: (2025)
por: Srikanth, Maya, et al.
Publicado: (2025)
From Generation to Collaboration: Using LLMs to Edit for Empathy in Healthcare
por: Luo, Man, et al.
Publicado: (2026)
por: Luo, Man, et al.
Publicado: (2026)
Assessing Evaluation Metrics for Neural Test Oracle Generation
por: Shin, Jiho, et al.
Publicado: (2023)
por: Shin, Jiho, et al.
Publicado: (2023)
On the Risk of Misleading Reports: Diagnosing Textual Biases in Multimodal Clinical AI
por: Restrepo, David, et al.
Publicado: (2025)
por: Restrepo, David, et al.
Publicado: (2025)
Humans or LLMs as the Judge? A Study on Judgement Biases
por: Chen, Guiming Hardy, et al.
Publicado: (2024)
por: Chen, Guiming Hardy, et al.
Publicado: (2024)
LLMs Should Incorporate Explicit Mechanisms for Human Empathy
por: You, Xiaoxing, et al.
Publicado: (2026)
por: You, Xiaoxing, et al.
Publicado: (2026)
Detecting and Steering LLMs' Empathy in Action
por: Cadile, Juan P.
Publicado: (2025)
por: Cadile, Juan P.
Publicado: (2025)
The Gold Medals in an Empty Room: Diagnosing Metalinguistic Reasoning in LLMs with Camlang
por: Liu, Fenghua, et al.
Publicado: (2025)
por: Liu, Fenghua, et al.
Publicado: (2025)
PUB: A Pragmatics Understanding Benchmark for Assessing LLMs' Pragmatics Capabilities
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
Empathy and the Right to Be an Exception: What LLMs Can and Cannot Do
por: Kidder, William, et al.
Publicado: (2024)
por: Kidder, William, et al.
Publicado: (2024)
HEART-felt Narratives: Tracing Empathy and Narrative Style in Personal Stories with LLMs
por: Shen, Jocelyn, et al.
Publicado: (2024)
por: Shen, Jocelyn, et al.
Publicado: (2024)
Efficient-Empathy: Towards Efficient and Effective Selection of Empathy Data
por: Sun, Linzhuang, et al.
Publicado: (2024)
por: Sun, Linzhuang, et al.
Publicado: (2024)
Assessing Empathy in Large Language Models with Real-World Physician-Patient Interactions
por: Luo, Man, et al.
Publicado: (2024)
por: Luo, Man, et al.
Publicado: (2024)
Kardia-R1: Unleashing LLMs to Reason toward Understanding and Empathy for Emotional Support via Rubric-as-Judge Reinforcement Learning
por: Yuan, Jiahao, et al.
Publicado: (2025)
por: Yuan, Jiahao, et al.
Publicado: (2025)
Causal Understanding by LLMs: The Role of Uncertainty
por: Lithgow-Serrano, Oscar, et al.
Publicado: (2025)
por: Lithgow-Serrano, Oscar, et al.
Publicado: (2025)
From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users
por: Zheng, Wuqiang, et al.
Publicado: (2026)
por: Zheng, Wuqiang, et al.
Publicado: (2026)
Are LLMs Empathetic to All? Investigating the Influence of Multi-Demographic Personas on a Model's Empathy
por: Malik, Ananya, et al.
Publicado: (2025)
por: Malik, Ananya, et al.
Publicado: (2025)
SYNTHEMPATHY: A Scalable Empathy Corpus Generated Using LLMs Without Any Crowdsourcing
por: Chen, Run, et al.
Publicado: (2025)
por: Chen, Run, et al.
Publicado: (2025)
Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
por: Singh, Shivalika, et al.
Publicado: (2024)
por: Singh, Shivalika, et al.
Publicado: (2024)
GenderBench: Evaluation Suite for Gender Biases in LLMs
por: Pikuliak, Matúš
Publicado: (2025)
por: Pikuliak, Matúš
Publicado: (2025)
Analyzing Dialectical Biases in LLMs for Knowledge and Reasoning Benchmarks
por: Pan, Eileen, et al.
Publicado: (2025)
por: Pan, Eileen, et al.
Publicado: (2025)
Camellia: Benchmarking Cultural Biases in LLMs for Asian Languages
por: Naous, Tarek, et al.
Publicado: (2025)
por: Naous, Tarek, et al.
Publicado: (2025)
Quantifying Gender Biases Towards Politicians on Reddit
por: Marjanovic, Sara, et al.
Publicado: (2021)
por: Marjanovic, Sara, et al.
Publicado: (2021)
Psyche-R1: Towards Reliable Psychological LLMs through Unified Empathy, Expertise, and Reasoning
por: Dai, Chongyuan, et al.
Publicado: (2025)
por: Dai, Chongyuan, et al.
Publicado: (2025)
Leveraging LLMs for Predicting Unknown Diagnoses from Clinical Notes
por: Albassam, Dina, et al.
Publicado: (2025)
por: Albassam, Dina, et al.
Publicado: (2025)
C2PO: Diagnosing and Disentangling Bias Shortcuts in LLMs
por: Feng, Xuan, et al.
Publicado: (2025)
por: Feng, Xuan, et al.
Publicado: (2025)
Distilling Empathy from Large Language Models
por: Xie, Henry J., et al.
Publicado: (2025)
por: Xie, Henry J., et al.
Publicado: (2025)
Empathy-R1: A Chain-of-Empathy and Reinforcement Learning Framework for Long-Form Mental Health Support
por: Yao, Xianrong, et al.
Publicado: (2025)
por: Yao, Xianrong, et al.
Publicado: (2025)
Un-considering Contextual Information: Assessing LLMs' Understanding of Indexical Elements
por: Oguz, Metehan, et al.
Publicado: (2025)
por: Oguz, Metehan, et al.
Publicado: (2025)
Question Answering on Patient Medical Records with Private Fine-Tuned LLMs
por: Kothari, Sara, et al.
Publicado: (2025)
por: Kothari, Sara, et al.
Publicado: (2025)
Understanding GUI Agent Localization Biases through Logit Sharpness
por: Tao, Xingjian, et al.
Publicado: (2025)
por: Tao, Xingjian, et al.
Publicado: (2025)
MEDEQUALQA: Evaluating Biases in LLMs with Counterfactual Reasoning
por: Ghosh, Rajarshi, et al.
Publicado: (2025)
por: Ghosh, Rajarshi, et al.
Publicado: (2025)
Diagnosing Medical Datasets with Training Dynamics
por: Wenderoth, Laura
Publicado: (2024)
por: Wenderoth, Laura
Publicado: (2024)
To Lie or Not to Lie? Investigating The Biased Spread of Global Lies by LLMs
por: Khan, Zohaib, et al.
Publicado: (2026)
por: Khan, Zohaib, et al.
Publicado: (2026)
Ejemplares similares
-
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs
por: Liu, Yang, et al.
Publicado: (2025) -
DiversityMedQA: Assessing Demographic Biases in Medical Diagnosis using Large Language Models
por: Rawat, Rajat, et al.
Publicado: (2024) -
LLMs are Biased Evaluators But Not Biased for Retrieval Augmented Generation
por: Chen, Yen-Shan, et al.
Publicado: (2024) -
Flattery, Fluff, and Fog: Diagnosing and Mitigating Idiosyncratic Biases in Preference Models
por: Bharadwaj, Anirudh, et al.
Publicado: (2025) -
The Emotional Spectrum of LLMs: Leveraging Empathy and Emotion-Based Markers for Mental Health Support
por: De Grandi, Alessandro, et al.
Publicado: (2024)