Benchmarking large language models for biomedical natural language processing applications and recommendations
Fuente:
arXiv
Saved in:
| Main Authors: | Chen, Qingyu, Hu, Yan, Peng, Xueqing, Xie, Qianqian, Jin, Qiao, Gilson, Aidan, Singer, Maxwell B., Ai, Xuguang, Lai, Po-Ting, Wang, Zhizheng, Keloth, Vipina Kuttichi, Raja, Kalpana, Huang, Jiming, He, Huan, Lin, Fongci, Du, Jingcheng, Zhang, Rui, Zheng, W. Jim, Adelman, Ron A., Lu, Zhiyong, Xu, Hua |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering
by: Hu, Yan, et al.
Published: (2023)
by: Hu, Yan, et al.
Published: (2023)
Me LLaMA: Foundation Large Language Models for Medical Applications
by: Xie, Qianqian, et al.
Published: (2024)
by: Xie, Qianqian, et al.
Published: (2024)
Watson-Crick conjugates of words and languages
by: Mahalingam, Kalpana, et al.
Published: (2022)
by: Mahalingam, Kalpana, et al.
Published: (2022)
Entry-level guide to the use of large language models for medical research
by: Jin, Qiao, et al.
Published: (2024)
by: Jin, Qiao, et al.
Published: (2024)
EHRNavigator: A Multi-Agent System for Patient-Level Clinical Question Answering over Heterogeneous Electronic Health Records
by: Qian, Lingfei, et al.
Published: (2026)
by: Qian, Lingfei, et al.
Published: (2026)
Benchmarking GPT-5 for biomedical natural language processing
by: Hou, Yu, et al.
Published: (2025)
by: Hou, Yu, et al.
Published: (2025)
BioPulse-QA: A Dynamic Biomedical Question-Answering Benchmark for Evaluating Factuality, Robustness, and Bias in Large Language Models
by: Bhattarai, Kriti, et al.
Published: (2026)
by: Bhattarai, Kriti, et al.
Published: (2026)
VOLMO: Versatile and Open Large Models for Ophthalmology
by: Qin, Zhenyue, et al.
Published: (2026)
by: Qin, Zhenyue, et al.
Published: (2026)
Enhancing Large Language Models with Domain-specific Retrieval Augment Generation: A Case Study on Long-form Consumer Health Question Answering in Ophthalmology
by: Gilson, Aidan, et al.
Published: (2024)
by: Gilson, Aidan, et al.
Published: (2024)
Information Extraction from Clinical Notes: Are We Ready to Switch to Large Language Models?
by: Hu, Yan, et al.
Published: (2024)
by: Hu, Yan, et al.
Published: (2024)
Toward Automated Cognitive Assessment in Parkinson's Disease Using Pretrained Language Models
by: Khanna, Varada, et al.
Published: (2025)
by: Khanna, Varada, et al.
Published: (2025)
Plain language adaptations of biomedical text using LLMs: Comparision of evaluation metrics
by: Kocbek, Primoz, et al.
Published: (2025)
by: Kocbek, Primoz, et al.
Published: (2025)
Streamlining evidence based clinical recommendations with large language models
by: Li, Dubai, et al.
Published: (2025)
by: Li, Dubai, et al.
Published: (2025)
Assessing biomedical knowledge robustness in large language models by query-efficient sampling attacks
by: Xian, R. Patrick, et al.
Published: (2024)
by: Xian, R. Patrick, et al.
Published: (2024)
Conversations on reasoning: Large language models in diagnosis
by: Daniel Restrepo, et al.
Published: (2024)
by: Daniel Restrepo, et al.
Published: (2024)
Factual consistency evaluation of summarization in the Era of large language models
by: Luo, Zheheng, et al.
Published: (2024)
by: Luo, Zheheng, et al.
Published: (2024)
MedicalBERT: enhancing biomedical natural language processing using pretrained BERT-based model
by: Reddy, K. Sahit, et al.
Published: (2025)
by: Reddy, K. Sahit, et al.
Published: (2025)
Strong and weak alignment of large language models with human values
by: Khamassi, Mehdi, et al.
Published: (2024)
by: Khamassi, Mehdi, et al.
Published: (2024)
Identifying signs and symptoms of urinary tract infection from emergency department clinical notes using large language models
by: Mark Iscoe, et al.
Published: (2024)
by: Mark Iscoe, et al.
Published: (2024)
EvaluateBM: a multi-agent framework for evaluating reasoning-capable language models in biomedical tasks
by: LIN, XINYI
Published: (2026)
by: LIN, XINYI
Published: (2026)
Ethical AI prompt recommendations in large language models using collaborative filtering
by: Nelson, Jordan, et al.
Published: (2025)
by: Nelson, Jordan, et al.
Published: (2025)
LEME: Open Large Language Models for Ophthalmology with Advanced Reasoning and Clinical Validation
by: Kim, Hyunjae, et al.
Published: (2024)
by: Kim, Hyunjae, et al.
Published: (2024)
Can OpenAI o1 Reason Well in Ophthalmology? A 6,990-Question Head-to-Head Evaluation Study
by: Srinivasan, Sahana, et al.
Published: (2025)
by: Srinivasan, Sahana, et al.
Published: (2025)
Improving accuracy of GPT-3/4 results on biomedical data using a retrieval-augmented language model
by: Soong, David, et al.
Published: (2023)
by: Soong, David, et al.
Published: (2023)
Methods and approaches for developing plain language versions of guidelines recommendations: Scoping review protocol
by: Tereza Friessová, et al.
Published: (2024)
by: Tereza Friessová, et al.
Published: (2024)
Peeking through the language barrier: the development of a free/open-source gisting system for Basque to English based on apertium.org
by: Jim O’Regan
Published: (2013)
by: Jim O’Regan
Published: (2013)
Potential causes of mortality for horseshoe crabs (Limulus polyphemus) during the biomedical bleeding process
by: Hurton, Lenka, et al.
Published: (2006)
by: Hurton, Lenka, et al.
Published: (2006)
Use of gender‐inclusive language in genetic counseling to optimize patient care
by: Heather Motiff, et al.
Published: (2024)
by: Heather Motiff, et al.
Published: (2024)
A multi-language toolkit for the semi-automated checking of research outputs
by: Preen, Richard J., et al.
Published: (2022)
by: Preen, Richard J., et al.
Published: (2022)
Waterproof Editor: an educational environment for proof assistants and programming languages
by: Otte, Pim, et al.
Published: (2026)
by: Otte, Pim, et al.
Published: (2026)
Dr Wenowdis: Specializing dynamic language C extensions using type information
by: Bernstein, Maxwell, et al.
Published: (2024)
by: Bernstein, Maxwell, et al.
Published: (2024)
The performance of ChatGPT and other large language models on multiple‐choice questions in biomedical disciplines: A meta‐analysis
by: Colleen M. Cheverko, et al.
Published: (2026)
by: Colleen M. Cheverko, et al.
Published: (2026)
Leveraging language models to categorize individual‐level priorities and deliver personalized brain health recommendations
by: Claudio Toro‐Serey, et al.
Published: (2025)
by: Claudio Toro‐Serey, et al.
Published: (2025)
PubTator 3.0: an AI-powered Literature Resource for Unlocking Biomedical Knowledge
by: Wei, Chih-Hsuan, et al.
Published: (2024)
by: Wei, Chih-Hsuan, et al.
Published: (2024)
Teorías del desarrollo económico / Irma Ademan; traductor, Roberto Ramón Reyes
by: Adelman, Irma
Published: (1965)
by: Adelman, Irma
Published: (1965)
Teorías del desarrollo económico / Irma Adelman ; traducción de Roberto Ramón Reyes
by: Adelman, Irma
by: Adelman, Irma
Los orígenes del terrorismo / Irma Adelman
by: Adelman, Irma
Published: (2002)
by: Adelman, Irma
Published: (2002)
Reseña de "O reencantamento do político: interpretações da contracultura" de Julie Stephens
by: Míriam Adelman
Published: (2001)
by: Míriam Adelman
Published: (2001)
Uma trajetória pessoal e acadêmica: entrevista com Raewyn Connell
by: Miriam Adelman
Published: (2013)
by: Miriam Adelman
Published: (2013)
Observando a Colombia: Albert O. Hirschman y la Economía del Desarrollo
by: Jeremy Adelman
Published: (2008)
by: Jeremy Adelman
Published: (2008)
Similar Items
-
Improving Large Language Models for Clinical Named Entity Recognition via Prompt Engineering
by: Hu, Yan, et al.
Published: (2023) -
Me LLaMA: Foundation Large Language Models for Medical Applications
by: Xie, Qianqian, et al.
Published: (2024) -
Watson-Crick conjugates of words and languages
by: Mahalingam, Kalpana, et al.
Published: (2022) -
Entry-level guide to the use of large language models for medical research
by: Jin, Qiao, et al.
Published: (2024) -
EHRNavigator: A Multi-Agent System for Patient-Level Clinical Question Answering over Heterogeneous Electronic Health Records
by: Qian, Lingfei, et al.
Published: (2026)