How well do LLMs cite relevant medical references? An evaluation framework and analyses
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wu, Kevin, Wu, Eric, Cassasola, Ally, Zhang, Angela, Wei, Kevin, Nguyen, Teresa, Riantawan, Sith, Riantawan, Patricia Shi, Ho, Daniel E., Zou, James |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs?
par: Wu, Eric, et autres
Publié: (2024)
par: Wu, Eric, et autres
Publié: (2024)
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
par: Kwon, Yongchan, et autres
Publié: (2023)
par: Kwon, Yongchan, et autres
Publié: (2023)
ClashEval: Quantifying the tug-of-war between an LLM's internal prior and external evidence
par: Wu, Kevin, et autres
Publié: (2024)
par: Wu, Kevin, et autres
Publié: (2024)
Regulating AI Adaptation: An Analysis of AI Medical Device Updates
par: Wu, Kevin, et autres
Publié: (2024)
par: Wu, Kevin, et autres
Publié: (2024)
Dynamic disruption index across citation and cited references windows: Recommendations for thresholds in research evaluation
par: Chen, Hongkan, et autres
Publié: (2025)
par: Chen, Hongkan, et autres
Publié: (2025)
Dynamic disruption index across citation and cited references windows: Recommendations for thresholds in research evaluation
par: Hongkan Chen, et autres
Publié: (2026)
par: Hongkan Chen, et autres
Publié: (2026)
Which is the cited source? A new perspective on article evaluation based on semantic similarity—Citation contribution attribution
par: Siluo Yang, et autres
Publié: (2025)
par: Siluo Yang, et autres
Publié: (2025)
MedArena: Comparing LLMs for Medicine-in-the-Wild Clinician Preferences
par: Wu, Eric, et autres
Publié: (2026)
par: Wu, Eric, et autres
Publié: (2026)
Chapter 3 Droit de cité
par: Smithies, James, et autres
Publié: (2023)
par: Smithies, James, et autres
Publié: (2023)
Declaration 7 On the need for quoting bibliographical or other references for all names cited in zoological works
par: International Commission on Zoological Nomenclature, et autres
Publié: (1943)
par: International Commission on Zoological Nomenclature, et autres
Publié: (1943)
New paper-by-paper classification for Scopus based on references reclassified by the origin of the papers citing them
par: Alvarez-Llorente, Jesus M., et autres
Publié: (2024)
par: Alvarez-Llorente, Jesus M., et autres
Publié: (2024)
Boltzmann framework for polyatomic gases: review on well-posedness, higher integrability and physical relevance
par: Alonso, Ricardo, et autres
Publié: (2025)
par: Alonso, Ricardo, et autres
Publié: (2025)
MedCaseReasoning: Evaluating and learning diagnostic reasoning from clinical case reports
par: Wu, Kevin, et autres
Publié: (2025)
par: Wu, Kevin, et autres
Publié: (2025)
Hepatic fibrosis in patients diagnosed with rheumatoid arthritis and psoriatic arthritis receiving methotrexate: A cross‐sectional analysis using transient elastography
par: Ratchaya Lertnawapan, et autres
Publié: (2025)
par: Ratchaya Lertnawapan, et autres
Publié: (2025)
Compositionality in perception: A framework
par: Kevin J. Lande
Publié: (2024)
par: Kevin J. Lande
Publié: (2024)
On “ Colonial scientific-medical documentary films and the legitimization of an ideal state in post-war Spain,” or the importance of being cited
par: Francisco Javier Martínez
Publié: (2017)
par: Francisco Javier Martínez
Publié: (2017)
Carrollian conformal correlators and massless scattering amplitudes
par: Nguyen, Kevin
Publié: (2023)
par: Nguyen, Kevin
Publié: (2023)
Hydrodynamics of two-dimensional CFTs
par: Nguyen, Kevin
Publié: (2025)
par: Nguyen, Kevin
Publié: (2025)
Lectures on Carrollian Holography
par: Nguyen, Kevin
Publié: (2025)
par: Nguyen, Kevin
Publié: (2025)
Hear ye! Hear ye! A proclamation of deadlines and deeds done well
par: Kevin W. Eva
Publié: (2024)
par: Kevin W. Eva
Publié: (2024)
EthoCRED: a framework to guide reporting and evaluation of the relevance and reliability of behavioural ecotoxicity studies
par: Michael G. Bertram, et autres
Publié: (2024)
par: Michael G. Bertram, et autres
Publié: (2024)
EthoCRED: a framework to guide reporting and evaluation of the relevance and reliability of behavioural ecotoxicity studies.
par: Bertram, Michael G, et autres
Publié: (2025)
par: Bertram, Michael G, et autres
Publié: (2025)
The 50 most cited studies on trochleoplasty
par: Alexander Pfarrmaier, et autres
Publié: (2025)
par: Alexander Pfarrmaier, et autres
Publié: (2025)
« Camp » dira-t-on, une cité improvisée
par: Lilia Benbelaïd
Publié: (2022)
par: Lilia Benbelaïd
Publié: (2022)
How to select slices for annotation to train best-performing deep learning segmentation models for cross-sectional medical images?
par: Zhang, Yixin, et autres
Publié: (2024)
par: Zhang, Yixin, et autres
Publié: (2024)
allydetre/snotelprocessr: snotelprocessr 1.0.0
par: Ally Detre
Publié: (2025)
par: Ally Detre
Publié: (2025)
Mary and Early Christian Women
par: Kateusz, Ally
Publié: (2020)
par: Kateusz, Ally
Publié: (2020)
RETRACTED: The interplay of corporate governance, internal audit effectiveness, and sustainable financial reporting quality in Tanzanian commercial banks
par: Zawadi Ally
Publié: (2024)
par: Zawadi Ally
Publié: (2024)
Librarian Recommended: A Survey of Engineering Education Resources in Library Research Guides
par: Ally Wood
Publié: (2025)
par: Ally Wood
Publié: (2025)
Evaluating the Efficiency and Sustainability of Domestic and Foreign Banks in Tanzania: Insights From the Digital Transformation Era
par: Zawadi Ally
Publié: (2025)
par: Zawadi Ally
Publié: (2025)
What is the future of mobile learning in education?
par: Mohamed Ally
Publié: (2014)
par: Mohamed Ally
Publié: (2014)
A validity evaluation of lexicon‐based sentiment analysis of medical students' clinical performance from in‐training evaluation reports
par: Irene Ma, et autres
Publié: (2025)
par: Irene Ma, et autres
Publié: (2025)
Living Off the LLM: How LLMs Will Change Adversary Tactics
par: Oesch, Sean, et autres
Publié: (2025)
par: Oesch, Sean, et autres
Publié: (2025)
SPIN-Bench: How Well Do LLMs Plan Strategically and Reason Socially?
par: Yao, Jianzhu, et autres
Publié: (2025)
par: Yao, Jianzhu, et autres
Publié: (2025)
How well can LLMs Grade Essays in Arabic?
par: Ghazawi, Rayed, et autres
Publié: (2025)
par: Ghazawi, Rayed, et autres
Publié: (2025)
How do medical students' expectations shape their experiences of well‐being programmes?
par: Emmanuel Tan, et autres
Publié: (2024)
par: Emmanuel Tan, et autres
Publié: (2024)
Platform for processing medical ultrasound obstetric images enabled in the cloud
par: Kevin Jessid Figueroa Maza
Publié: (2015)
par: Kevin Jessid Figueroa Maza
Publié: (2015)
A statistical framework for detecting therapy-induced resistance from drug screens
par: Wu, Chenyu, et autres
Publié: (2024)
par: Wu, Chenyu, et autres
Publié: (2024)
Writing science : How to write papers that get cited and proposals that get funded / Joshua Schimel
par: Schimel, J
Publié: (2012)
par: Schimel, J
Publié: (2012)
Preliminary suggestions for rigorous GPAI model evaluations
par: Paskov, Patricia, et autres
Publié: (2025)
par: Paskov, Patricia, et autres
Publié: (2025)
Documents similaires
-
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs?
par: Wu, Eric, et autres
Publié: (2024) -
DataInf: Efficiently Estimating Data Influence in LoRA-tuned LLMs and Diffusion Models
par: Kwon, Yongchan, et autres
Publié: (2023) -
ClashEval: Quantifying the tug-of-war between an LLM's internal prior and external evidence
par: Wu, Kevin, et autres
Publié: (2024) -
Regulating AI Adaptation: An Analysis of AI Medical Device Updates
par: Wu, Kevin, et autres
Publié: (2024) -
Dynamic disruption index across citation and cited references windows: Recommendations for thresholds in research evaluation
par: Chen, Hongkan, et autres
Publié: (2025)