Ensuring Safety and Trust: Analyzing the Risks of Large Language Models in Medicine
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Yifan, Jin, Qiao, Leaman, Robert, Liu, Xiaoyu, Xiong, Guangzhi, Sarfo-Gyamfi, Maame, Gong, Changlin, Ferrière-Steinert, Santiago, Wilbur, W. John, Li, Xiaojun, Yuan, Jiaxin, An, Bang, Castro, Kelvin S., Álvarez, Francisco Erramuspe, Stockle, Matías, Zhang, Aidong, Huang, Furong, Lu, Zhiyong |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Retrieval-Augmented Generation for Medicine
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Adversarial Attacks on Large Language Models in Medicine
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
Entry-level guide to the use of large language models for medical research
by: Jin, Qiao, et al.
Published: (2024)
by: Jin, Qiao, et al.
Published: (2024)
MedCite: Can Language Models Generate Verifiable Text for Medicine?
by: Wang, Xiao, et al.
Published: (2025)
by: Wang, Xiao, et al.
Published: (2025)
MedCalc-Bench: Evaluating Large Language Models for Medical Calculations
by: Khandekar, Nikhil, et al.
Published: (2024)
by: Khandekar, Nikhil, et al.
Published: (2024)
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
Structural Causality-based Generalizable Concept Discovery Models
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
ASCENT-ViT: Attention-based Scale-aware Concept Learning Framework for Enhanced Alignment in Vision Transformers
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
COCO-Tree: Compositional Hierarchical Concept Trees for Enhanced Reasoning in Vision Language Models
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
CoLiDR: Concept Learning using Aggregated Disentangled Representations
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
Neural Additive Experts: Context-Gated Experts for Controllable Model Additivity
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
A Self-explaining Neural Architecture for Generalizable Concept Learning
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
ProtoNAM: Prototypical Neural Additive Models for Interpretable Deep Tabular Learning
by: Xiong, Guangzhi, et al.
Published: (2024)
by: Xiong, Guangzhi, et al.
Published: (2024)
Unmasking and Quantifying Racial Bias of Large Language Models in Medical Report Generation
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
CT-Bench: A Benchmark for Multimodal Lesion Understanding in Computed Tomography
by: Zhu, Qingqing, et al.
Published: (2026)
by: Zhu, Qingqing, et al.
Published: (2026)
Zur Sprachdynamik des Konjunktivs im Bairischen in Österreich
by: Philipp Stöckle
Published: (2022)
by: Philipp Stöckle
Published: (2022)
GCAV: A Global Concept Activation Vector Framework for Cross-Layer Consistency in Interpretability
by: He, Zhenghao, et al.
Published: (2025)
by: He, Zhenghao, et al.
Published: (2025)
Concept-RuleNet: Grounded Multi-Agent Neurosymbolic Reasoning in Vision Language Models
by: Sinha, Sanchit, et al.
Published: (2025)
by: Sinha, Sanchit, et al.
Published: (2025)
Retrieving Counterfactuals Improves Visual In-Context Learning
by: Xiong, Guangzhi, et al.
Published: (2026)
by: Xiong, Guangzhi, et al.
Published: (2026)
Mare, fiume, ruscello
by: Stöckle, Susanne Magdalena
Published: (2022)
by: Stöckle, Susanne Magdalena
Published: (2022)
Med-V1: Small Language Models for Zero-shot and Scalable Biomedical Evidence Attribution
by: Jin, Qiao, et al.
Published: (2026)
by: Jin, Qiao, et al.
Published: (2026)
Beyond Multiple-Choice Accuracy: Real-World Challenges of Implementing Large Language Models in Healthcare
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
PubMed and Beyond: Biomedical Literature Search in the Age of Artificial Intelligence
by: Jin, Qiao, et al.
Published: (2023)
by: Jin, Qiao, et al.
Published: (2023)
What Do Biomedical NER and Entity Linking Benchmarks Measure? A Corpus-Centric Diagnostic Framework
by: Leaman, Robert, et al.
Published: (2026)
by: Leaman, Robert, et al.
Published: (2026)
Asplenium X kentuckiense on Granitic Gneiss in Georgia
by: Duncan, Wilbur H. (Wilbur Howard)
Published: (1966)
by: Duncan, Wilbur H. (Wilbur Howard)
Published: (1966)
Emerging Challenges in Ghana's HPV Vaccination Rollout: Trust, Consent, and Communication
by: Michael Sarfo, et al.
Published: (2026)
by: Michael Sarfo, et al.
Published: (2026)
Reasoning Beyond Chain-of-Thought: A Latent Computational Mode in Large Language Models
by: He, Zhenghao, et al.
Published: (2026)
by: He, Zhenghao, et al.
Published: (2026)
Attention-guided Fine-tuning of Multimodal Large Language Models Improves Chain-of-Thought Reasoning
by: Sinha, Sanchit, et al.
Published: (2026)
by: Sinha, Sanchit, et al.
Published: (2026)
Toward Faithful Retrieval-Augmented Generation with Sparse Autoencoders
by: Xiong, Guangzhi, et al.
Published: (2025)
by: Xiong, Guangzhi, et al.
Published: (2025)
CASL: Concept-Aligned Sparse Latents for Interpreting Diffusion Models
by: He, Zhenghao, et al.
Published: (2026)
by: He, Zhenghao, et al.
Published: (2026)
Chapter Uterus
by: Steinert, Ulrike
Published: (2019)
by: Steinert, Ulrike
Published: (2019)
Chapter 9 Maintenance of Value and the Value of Maintenance
by: Steinert, Steffen
Published: (2024)
by: Steinert, Steffen
Published: (2024)
Chapter ‘Tested’ Remedies in Mesopotamian Medical Texts
by: Steinert, Ulrike
Published: (2019)
by: Steinert, Ulrike
Published: (2019)
Mineral counts from Geschiebemergel of various localities in northern Germany
by: Steinert, Harald
Published: (2011)
by: Steinert, Harald
Published: (2011)
A survey of recent methods for addressing AI fairness and bias in biomedicine
by: Yang, Yifan, et al.
Published: (2024)
by: Yang, Yifan, et al.
Published: (2024)
Explore Spurious Correlations at the Concept Level in Language Models for Text Classification
by: Zhou, Yuhang, et al.
Published: (2023)
by: Zhou, Yuhang, et al.
Published: (2023)
Radio signal generation in milliseconds: enabling multi-parameter reconstruction of ultra-high-energy cosmic rays
by: Ferrière, Arsène
Published: (2026)
by: Ferrière, Arsène
Published: (2026)
A grammar of Pite Saami
by: Wilbur, Joshua
Published: (2015)
by: Wilbur, Joshua
Published: (2015)
The eye of the tiger / Wilbur Smith
by: Smith, Wilbur
by: Smith, Wilbur
Similar Items
-
Benchmarking Retrieval-Augmented Generation for Medicine
by: Xiong, Guangzhi, et al.
Published: (2024) -
Improving Retrieval-Augmented Generation in Medicine with Iterative Follow-up Questions
by: Xiong, Guangzhi, et al.
Published: (2024) -
Adversarial Attacks on Large Language Models in Medicine
by: Yang, Yifan, et al.
Published: (2024) -
Entry-level guide to the use of large language models for medical research
by: Jin, Qiao, et al.
Published: (2024) -
MedCite: Can Language Models Generate Verifiable Text for Medicine?
by: Wang, Xiao, et al.
Published: (2025)