Generalist Foundation Models Are Not Clinical Enough for Hospital Operations
Fuente:
arXiv
Saved in:
| Main Authors: | Jiang, Lavender Y., Chen, Angelica, Han, Xu, Liu, Xujin Chris, Dua, Radhika, Eaton, Kevin, Wolff, Frederick, Steele, Robert, Zhang, Jeff, Alyakin, Anton, Pan, Qingkai, Chen, Yanbing, Sangwon, Karl L., Alber, Daniel A., Stryker, Jaden, Lee, Jin Vivian, Aphinyanaphongs, Yindalon, Cho, Kyunghyun, Oermann, Eric Karl |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generalist Large Language Models Outperform Clinical Tools on Medical Benchmarks
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
MedMobile: A mobile-sized language model with clinical capabilities
by: Vishwanath, Krithik, et al.
Published: (2024)
by: Vishwanath, Krithik, et al.
Published: (2024)
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
by: Rahman, Salman, et al.
Published: (2024)
by: Rahman, Salman, et al.
Published: (2024)
Evaluating the performance and fragility of large language models on the self-assessment for neurological surgeons
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
BPQA Dataset: Evaluating How Well Language Models Leverage Blood Pressures to Answer Biomedical Questions
by: Hang, Chi, et al.
Published: (2025)
by: Hang, Chi, et al.
Published: (2025)
Paradox of De-identification: A Critique of HIPAA Safe Harbour in the Age of LLMs
by: Jiang, Lavender Y., et al.
Published: (2026)
by: Jiang, Lavender Y., et al.
Published: (2026)
Medical large language models are easily distracted
by: Vishwanath, Krithik, et al.
Published: (2025)
by: Vishwanath, Krithik, et al.
Published: (2025)
It is Too Many Options: Pitfalls of Multiple-Choice Questions in Generative AI and Medical Education
by: Singh, Shrutika, et al.
Published: (2025)
by: Singh, Shrutika, et al.
Published: (2025)
MedG-KRP: Medical Graph Knowledge Representation Probing
by: Rosenbaum, Gabriel R., et al.
Published: (2024)
by: Rosenbaum, Gabriel R., et al.
Published: (2024)
Refining Packing and Shuffling Strategies for Enhanced Performance in Generative Language Models
by: Chen, Yanbing, et al.
Published: (2024)
by: Chen, Yanbing, et al.
Published: (2024)
Large Language Models Predict Functional Outcomes after Acute Ischemic Stroke
by: Kapoor, Anjali K., et al.
Published: (2026)
by: Kapoor, Anjali K., et al.
Published: (2026)
On the Relationship Between the Choice of Representation and In-Context Learning
by: Marinescu, Ioana, et al.
Published: (2025)
by: Marinescu, Ioana, et al.
Published: (2025)
Clinically Grounded Agent-based Report Evaluation: An Interpretable Metric for Radiology Report Generation
by: Dua, Radhika, et al.
Published: (2025)
by: Dua, Radhika, et al.
Published: (2025)
CNS-Obsidian: A Neurosurgical Vision-Language Model Built From Scientific Publications
by: Alyakin, Anton, et al.
Published: (2025)
by: Alyakin, Anton, et al.
Published: (2025)
New-Onset Diabetes Assessment Using Artificial Intelligence-Enhanced Electrocardiography
by: Zhang, Hao, et al.
Published: (2022)
by: Zhang, Hao, et al.
Published: (2022)
"All You Need" is Not All You Need for a Paper Title: On the Origins of a Scientific Meme
by: Alyakin, Anton
Published: (2025)
by: Alyakin, Anton
Published: (2025)
A spinor proof of the classification of stable minimal surfaces in $\mathbb{R}^3$
by: Stryker, Douglas
Published: (2026)
by: Stryker, Douglas
Published: (2026)
Stable 2-systole bounds in positive scalar curvature
by: Stryker, Douglas
Published: (2026)
by: Stryker, Douglas
Published: (2026)
Min-max construction of prescribed mean curvature hypersurfaces in noncompact manifolds
by: Stryker, Douglas
Published: (2024)
by: Stryker, Douglas
Published: (2024)
Lebensgeschichten alter Eltern kognitiv beeinträchtigter Menschen
by: Oermann, Lisa
Published: (2023)
by: Oermann, Lisa
Published: (2023)
Germany's Foreign Policy Towards Poland and the Czech Republic
by: Cordell, Karl, et al.
Published: (2019)
by: Cordell, Karl, et al.
Published: (2019)
Germany's Foreign Policy Towards Poland and the Czech Republic
by: Cordell, Karl, et al.
Published: (2025)
by: Cordell, Karl, et al.
Published: (2025)
Shearing approach to gauge-invariant Trotterization
by: Stryker, Jesse R.
Published: (2021)
by: Stryker, Jesse R.
Published: (2021)
New Trends in English Education: Selected Addresses Delivered at the Conference on English Education (4th, Carnegie Institute of Technology, March 31, April 1, 2, 1966).
by: Stryker, David, Ed.
Published: (1966)
by: Stryker, David, Ed.
Published: (1966)
ChatGPT Solving Complex Kidney Transplant Cases: A Comparative Study With Human Respondents
by: Michal A. Mankowski, et al.
Published: (2024)
by: Michal A. Mankowski, et al.
Published: (2024)
Conditional Generative Modeling for Enhanced Credit Risk Management in Supply Chain Finance
by: Zhang, Qingkai, et al.
Published: (2025)
by: Zhang, Qingkai, et al.
Published: (2025)
Conditional Generative Modeling for Enhanced Credit Risk Management in Supply Chain Finance
by: Qingkai Zhang, et al.
Published: (2025)
by: Qingkai Zhang, et al.
Published: (2025)
A Real-Time Defense Against Object Vanishing Adversarial Patch Attacks for Object Detection in Autonomous Vehicles
by: Mu, Jaden
Published: (2024)
by: Mu, Jaden
Published: (2024)
Optimal regularity for minimizers of the prescribed mean curvature functional over isotopies
by: Sarnataro, Lorenzo, et al.
Published: (2023)
by: Sarnataro, Lorenzo, et al.
Published: (2023)
edwardlavender/Patter.jl: Patter.jl
by: Edward Lavender
Published: (2025)
by: Edward Lavender
Published: (2025)
Clinical Management and Outcomes of Isolated Fetal Pyelectasis Detected at Second‐Trimester Ultrasound: A Retrospective Cohort Study
by: Ilona Lavender
Published: (2026)
by: Ilona Lavender
Published: (2026)
Los Angeles : Two Hundred / David Lavender
by: Lavender, David
by: Lavender, David
The Immaterial Turn in Medieval Latin Theories of Sensation
by: Jordan Lavender
Published: (2026)
by: Jordan Lavender
Published: (2026)
Large-Scale Multi-omic Biosequence Transformers for Modeling Protein-Nucleic Acid Interactions
by: Chen, Sully F., et al.
Published: (2024)
by: Chen, Sully F., et al.
Published: (2024)
AutoEval: Autonomous Evaluation of Generalist Robot Manipulation Policies in the Real World
by: Zhou, Zhiyuan, et al.
Published: (2025)
by: Zhou, Zhiyuan, et al.
Published: (2025)
Perturbative hydrogenic Lamb shifts and radiative decay rates -- an so(4,2)-based algebraic approach
by: Alber, Gernot
Published: (2026)
by: Alber, Gernot
Published: (2026)
Scheitern als Performance
by: Alber, Nicole
Published: (2024)
by: Alber, Nicole
Published: (2024)
Settleable and Non-Settleable Suspended Sediments in the Ogeechee River Estuary, Georgia, U.S.A
by: Alber, M
Published: (1996)
by: Alber, M
Published: (1996)
The Predictive Value of Structural OCT in Preclinical AD
by: Jessica Alber
Published: (2024)
by: Jessica Alber
Published: (2024)
A Model-Independent Radio Telescope Dark Matter Search in the L and S Bands
by: Keller, Aya, et al.
Published: (2021)
by: Keller, Aya, et al.
Published: (2021)
Similar Items
-
Generalist Large Language Models Outperform Clinical Tools on Medical Benchmarks
by: Vishwanath, Krithik, et al.
Published: (2025) -
MedMobile: A mobile-sized language model with clinical capabilities
by: Vishwanath, Krithik, et al.
Published: (2024) -
Generalization in Healthcare AI: Evaluation of a Clinical Large Language Model
by: Rahman, Salman, et al.
Published: (2024) -
Evaluating the performance and fragility of large language models on the self-assessment for neurological surgeons
by: Vishwanath, Krithik, et al.
Published: (2025) -
BPQA Dataset: Evaluating How Well Language Models Leverage Blood Pressures to Answer Biomedical Questions
by: Hang, Chi, et al.
Published: (2025)