The challenge of uncertainty quantification of large language models in medicine
Fuente:
arXiv
Saved in:
| Main Authors: | Atf, Zahra, Safavi-Naini, Seyed Amir Ahmad, Lewis, Peter R., Mahjoubfar, Aref, Naderi, Nariman, Savage, Thomas R., Soroush, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
by: Naderi, Nariman, et al.
Published: (2025)
by: Naderi, Nariman, et al.
Published: (2025)
Self-Reported Confidence of Large Language Models in Gastroenterology: Analysis of Commercial, Open-Source, and Quantized Models
by: Naderi, Nariman, et al.
Published: (2025)
by: Naderi, Nariman, et al.
Published: (2025)
State of Abdominal CT Datasets: A Critical Review of Bias, Clinical Relevance, and Real-world Applicability
by: Danaei, Saeide, et al.
Published: (2025)
by: Danaei, Saeide, et al.
Published: (2025)
ScenarioBench: Trace-Grounded Compliance Evaluation for Text-to-SQL and RAG
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Is Trust Correlated With Explainability in AI? A Meta-Analysis
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Rule-Based Moral Principles for Explaining Uncertainty in Natural Language Generation
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Kantian-Utilitarian XAI: Meta-Explained
by: Atf, Zahra, et al.
Published: (2025)
by: Atf, Zahra, et al.
Published: (2025)
Grounding Clinical AI Competency in Human Cognition Through the Clinical World Model and Skill-Mix Framework
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2026)
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2026)
3DLAND: 3D Lesion Abdominal Anomaly Localization Dataset
by: Advand, Mehran, et al.
Published: (2026)
by: Advand, Mehran, et al.
Published: (2026)
Safety challenges of AI in medicine in the era of large language models
by: Wang, Xiaoye, et al.
Published: (2024)
by: Wang, Xiaoye, et al.
Published: (2024)
Hybrid Encryption with Certified Deletion in Preprocessing Model
by: Dey, Kunal, et al.
Published: (2026)
by: Dey, Kunal, et al.
Published: (2026)
Secure Composition of Quantum Key Distribution and Symmetric Key Encryption
by: Dey, Kunal, et al.
Published: (2025)
by: Dey, Kunal, et al.
Published: (2025)
Vision Language Models versus Machine Learning Models Performance on Polyp Detection and Classification in Colonoscopy Images
by: Khalafi, Mohammad Amin, et al.
Published: (2025)
by: Khalafi, Mohammad Amin, et al.
Published: (2025)
The evolution of systems biology and systems medicine: From mechanistic models to uncertainty quantification
by: Qiao, Lingxia, et al.
Published: (2024)
by: Qiao, Lingxia, et al.
Published: (2024)
Predicting Post‐ ERCP Pancreatitis Using Machine Learning: Risk Stratification and Feature Importance Analysis
by: Erfan Arabpour, et al.
Published: (2026)
by: Erfan Arabpour, et al.
Published: (2026)
Towards trustworthy artificial intelligence in musculoskeletal medicine: A narrative review on uncertainty quantification
by: Amir M. Vahdani, et al.
Published: (2025)
by: Amir M. Vahdani, et al.
Published: (2025)
AMIGO: Agentic Multi-Image Grounding Oracle Benchmark
by: Wang, Min, et al.
Published: (2026)
by: Wang, Min, et al.
Published: (2026)
Vision-Language and Large Language Model Performance in Gastroenterology: GPT, Claude, Llama, Phi, Mistral, Gemma, and Quantized Models
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2024)
by: Safavi-Naini, Seyed Amir Ahmad, et al.
Published: (2024)
FarsEval-PKBETS: A new diverse benchmark for evaluating Persian large language models
by: Shamsfard, Mehrnoush, et al.
Published: (2025)
by: Shamsfard, Mehrnoush, et al.
Published: (2025)
Robust and Reusable Fuzzy Extractors for Low-entropy Rate Randomness Sources
by: Panja, Somnath, et al.
Published: (2024)
by: Panja, Somnath, et al.
Published: (2024)
Fast quantum state preparation and bath dynamics using non-Gaussian variational ansatz and quantum optimal control
by: Bond, Liam J., et al.
Published: (2023)
by: Bond, Liam J., et al.
Published: (2023)
AxLLM: accelerator architecture for large language models with computation reuse capability
by: Ahadi, Soroush, et al.
Published: (2025)
by: Ahadi, Soroush, et al.
Published: (2025)
Efficient semantic uncertainty quantification in language models via diversity-steered sampling
by: Park, Ji Won, et al.
Published: (2025)
by: Park, Ji Won, et al.
Published: (2025)
Universal dynamics from a single-particle dark state
by: Daraban, Ruben, et al.
Published: (2026)
by: Daraban, Ruben, et al.
Published: (2026)
Phonon-mediated quantum gates in trapped ions coupled to an ultracold atomic gas
by: Oghittu, Lorenzo, et al.
Published: (2024)
by: Oghittu, Lorenzo, et al.
Published: (2024)
CCA-Secure Hybrid Encryption in Correlated Randomness Model and KEM Combiners
by: Panja, Somnath, et al.
Published: (2024)
by: Panja, Somnath, et al.
Published: (2024)
Emerging clinical applications of large language models in emergency medicine
by: Jon Herries
Published: (2024)
by: Jon Herries
Published: (2024)
Physical measurement with in-line fiber Mach-Zehnder interferometer using differential phase white light interferometry
by: Aref, Seyed Hashem
Published: (2017)
by: Aref, Seyed Hashem
Published: (2017)
R2D2 image reconstruction with model uncertainty quantification in radio astronomy
by: Aghabiglou, Amir, et al.
Published: (2024)
by: Aghabiglou, Amir, et al.
Published: (2024)
Towards interfacing large language models with ASR systems using confidence measures and prompting
by: Naderi, Maryam, et al.
Published: (2024)
by: Naderi, Maryam, et al.
Published: (2024)
Multilevel Electromagnetically Induced Transparency Cooling
by: Fouka, Katya, et al.
Published: (2025)
by: Fouka, Katya, et al.
Published: (2025)
A Physics-Informed Machine Learning Framework for Solid Boundary Treatment in Meshfree Particle Methods
by: Mehranfar, Nariman, et al.
Published: (2025)
by: Mehranfar, Nariman, et al.
Published: (2025)
Physics-guided impact localisation and force estimation in composite plates with uncertainty quantification
by: Xiao, Dong, et al.
Published: (2025)
by: Xiao, Dong, et al.
Published: (2025)
Transforming healthcare with large language models: Current applications, challenges, and future directions—a literature review
by: Muhammad Umar, et al.
Published: (2025)
by: Muhammad Umar, et al.
Published: (2025)
Surrogate modeling for uncertainty quantification in nonlinear dynamics
by: Marelli, S., et al.
Published: (2025)
by: Marelli, S., et al.
Published: (2025)
A critical review of methods and challenges in large language models
by: Moradi, Milad, et al.
Published: (2024)
by: Moradi, Milad, et al.
Published: (2024)
Calibrated uncertainty quantification for prosumer flexibility aggregation in ancillary service markets
by: Kumar, Yogesh Pipada Sunil, et al.
Published: (2026)
by: Kumar, Yogesh Pipada Sunil, et al.
Published: (2026)
Closing the gap between open-source and commercial large language models for medical evidence summarization
by: Zhang, Gongbo, et al.
Published: (2024)
by: Zhang, Gongbo, et al.
Published: (2024)
Bayesian neural network correction of RANS turbulence models with uncertainty quantification in separated flows
by: Buchanan, Tyler, et al.
Published: (2026)
by: Buchanan, Tyler, et al.
Published: (2026)
Leveraging large language models for automated performance appraisals: Opportunities and challenges
by: Kuchibhotla, Sri
Published: (2025)
by: Kuchibhotla, Sri
Published: (2025)
Similar Items
-
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
by: Naderi, Nariman, et al.
Published: (2025) -
Self-Reported Confidence of Large Language Models in Gastroenterology: Analysis of Commercial, Open-Source, and Quantized Models
by: Naderi, Nariman, et al.
Published: (2025) -
State of Abdominal CT Datasets: A Critical Review of Bias, Clinical Relevance, and Real-world Applicability
by: Danaei, Saeide, et al.
Published: (2025) -
ScenarioBench: Trace-Grounded Compliance Evaluation for Text-to-SQL and RAG
by: Atf, Zahra, et al.
Published: (2025) -
Is Trust Correlated With Explainability in AI? A Meta-Analysis
by: Atf, Zahra, et al.
Published: (2025)