Building a Silver-Standard Dataset from NICE Guidelines for Clinical LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Qing, Zhang, Eric Hua Qing, Jozsa, Felix, Ive, Julia |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Emotion Aware Therapeutic Dialogue Generation: A Multi-component Reinforcement Learning Approach to Language Models for Mental Health Support
by: Zhang, Eric Hua Qing, et al.
Published: (2025)
by: Zhang, Eric Hua Qing, et al.
Published: (2025)
Clean & Clear: Feasibility of Safe LLM Clinical Guidance
by: Ive, Julia, et al.
Published: (2025)
by: Ive, Julia, et al.
Published: (2025)
Hesitation is defeat? Connecting Linguistic and Predictive Uncertainty
by: Manzo, Gianluca, et al.
Published: (2025)
by: Manzo, Gianluca, et al.
Published: (2025)
Safe Training with Sensitive In-domain Data: Leveraging Data Fragmentation To Mitigate Linkage Attacks
by: Ignashina, Mariia, et al.
Published: (2024)
by: Ignashina, Mariia, et al.
Published: (2024)
Harnessing Large Language Models for Precision Querying and Retrieval-Augmented Knowledge Extraction in Clinical Data Science
by: Jan, Juan Jose Rubio, et al.
Published: (2026)
by: Jan, Juan Jose Rubio, et al.
Published: (2026)
Building Trust in Clinical LLMs: Bias Analysis and Dataset Transparency
by: Maslenkova, Svetlana, et al.
Published: (2025)
by: Maslenkova, Svetlana, et al.
Published: (2025)
Grounding Large Language Models in Clinical Evidence: A Retrieval-Augmented Generation System for Querying UK NICE Clinical Guidelines
by: Lewis, Matthew, et al.
Published: (2025)
by: Lewis, Matthew, et al.
Published: (2025)
Persian Abstract Meaning Representation: Annotation Guidelines and Gold Standard Dataset
by: Takhshid, Reza, et al.
Published: (2022)
by: Takhshid, Reza, et al.
Published: (2022)
Combining Hierachical VAEs with LLMs for clinically meaningful timeline summarisation in social media
by: Song, Jiayu, et al.
Published: (2024)
by: Song, Jiayu, et al.
Published: (2024)
Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines
by: Abonizio, Hugo, et al.
Published: (2026)
by: Abonizio, Hugo, et al.
Published: (2026)
MoralBERT: A Fine-Tuned Language Model for Capturing Moral Values in Social Discussions
by: Preniqi, Vjosa, et al.
Published: (2024)
by: Preniqi, Vjosa, et al.
Published: (2024)
NICE: To Optimize In-Context Examples or Not?
by: Srivastava, Pragya, et al.
Published: (2024)
by: Srivastava, Pragya, et al.
Published: (2024)
ENEIDE: A High Quality Silver Standard Dataset for Named Entity Recognition and Linking in Historical Italian
by: Santini, Cristian, et al.
Published: (2026)
by: Santini, Cristian, et al.
Published: (2026)
Avoiding Copyright Infringement via Large Language Model Unlearning
by: Dou, Guangyao, et al.
Published: (2024)
by: Dou, Guangyao, et al.
Published: (2024)
LLM Assistance for Pediatric Depression
by: Ignashina, Mariia, et al.
Published: (2025)
by: Ignashina, Mariia, et al.
Published: (2025)
Learning with Silver Standard Data for Zero-shot Relation Extraction
by: Wang, Tianyin, et al.
Published: (2022)
by: Wang, Tianyin, et al.
Published: (2022)
ClinicalAgents: Multi-Agent Orchestration for Clinical Decision Making with Dual-Memory
by: Ge, Zhuohan, et al.
Published: (2026)
by: Ge, Zhuohan, et al.
Published: (2026)
A Gold Standard Dataset and Evaluation Framework for Depression Detection and Explanation in Social Media using LLMs
by: Bolegave, Prajval, et al.
Published: (2025)
by: Bolegave, Prajval, et al.
Published: (2025)
Comprehensive Modeling and Question Answering of Cancer Clinical Practice Guidelines using LLMs
by: Gupta, Bhumika, et al.
Published: (2025)
by: Gupta, Bhumika, et al.
Published: (2025)
FactEHR: A Dataset for Evaluating Factuality in Clinical Notes Using LLMs
by: Munnangi, Monica, et al.
Published: (2024)
by: Munnangi, Monica, et al.
Published: (2024)
Training and Evaluation of Guideline-Based Medical Reasoning in LLMs
by: Staniek, Michael, et al.
Published: (2025)
by: Staniek, Michael, et al.
Published: (2025)
Instruction-Tuning LLMs for Event Extraction with Annotation Guidelines
by: Srivastava, Saurabh, et al.
Published: (2025)
by: Srivastava, Saurabh, et al.
Published: (2025)
Building High-Quality Datasets for Portuguese LLMs: From Common Crawl Snapshots to Industrial-Grade Corpora
by: Almeida, Thales Sales, et al.
Published: (2025)
by: Almeida, Thales Sales, et al.
Published: (2025)
AnesSuite: A Comprehensive Benchmark and Dataset Suite for Anesthesiology Reasoning in LLMs
by: Feng, Xiang, et al.
Published: (2025)
by: Feng, Xiang, et al.
Published: (2025)
Building Multilingual Datasets for Predicting Mental Health Severity through LLMs: Prospects and Challenges
by: Skianis, Konstantinos, et al.
Published: (2024)
by: Skianis, Konstantinos, et al.
Published: (2024)
Speaking the Same Language: Leveraging LLMs in Standardizing Clinical Data for AI
by: Sett, Arindam, et al.
Published: (2024)
by: Sett, Arindam, et al.
Published: (2024)
A Decade-Scale Benchmark Evaluating LLMs' Clinical Practice Guidelines Detection and Adherence in Multi-turn Conversations
by: Tan, Andong, et al.
Published: (2026)
by: Tan, Andong, et al.
Published: (2026)
Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines
by: Xie, Shiyao, et al.
Published: (2026)
by: Xie, Shiyao, et al.
Published: (2026)
SteerConf: Steering LLMs for Confidence Elicitation
by: Zhou, Ziang, et al.
Published: (2025)
by: Zhou, Ziang, et al.
Published: (2025)
Extract Information from Hybrid Long Documents Leveraging LLMs: A Framework and Dataset
by: Yue, Chongjian, et al.
Published: (2024)
by: Yue, Chongjian, et al.
Published: (2024)
EvidenceOutcomes: a Dataset of Clinical Trial Publications with Clinically Meaningful Outcomes
by: Zhou, Yiliang, et al.
Published: (2025)
by: Zhou, Yiliang, et al.
Published: (2025)
Building Accurate Translation-Tailored LLMs with Language Aware Instruction Tuning
by: Zan, Changtong, et al.
Published: (2024)
by: Zan, Changtong, et al.
Published: (2024)
ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
by: Zong, Qing, et al.
Published: (2024)
by: Zong, Qing, et al.
Published: (2024)
When Cases Get Rare: A Retrieval Benchmark for Off-Guideline Clinical Question Answering
by: Lee, Doeun, et al.
Published: (2026)
by: Lee, Doeun, et al.
Published: (2026)
AGGA: A Dataset of Academic Guidelines for Generative AI and Large Language Models
by: Jiao, Junfeng, et al.
Published: (2025)
by: Jiao, Junfeng, et al.
Published: (2025)
Crowdsourcing Piedmontese to Test LLMs on Non-Standard Orthography
by: Vico, Gianluca, et al.
Published: (2026)
by: Vico, Gianluca, et al.
Published: (2026)
GUIDE: A Guideline-Guided Dataset for Instructional Video Comprehension
by: Liang, Jiafeng, et al.
Published: (2024)
by: Liang, Jiafeng, et al.
Published: (2024)
Building a Macedonian Recipe Dataset: Collection, Parsing, and Comparative Analysis
by: Sasanski, Darko, et al.
Published: (2025)
by: Sasanski, Darko, et al.
Published: (2025)
Benchmarking LLMs' Judgments with No Gold Standard
by: Xu, Shengwei, et al.
Published: (2024)
by: Xu, Shengwei, et al.
Published: (2024)
On the use of Silver Standard Data for Zero-shot Classification Tasks in Information Extraction
by: Wang, Jianwei, et al.
Published: (2024)
by: Wang, Jianwei, et al.
Published: (2024)
Similar Items
-
Context-Emotion Aware Therapeutic Dialogue Generation: A Multi-component Reinforcement Learning Approach to Language Models for Mental Health Support
by: Zhang, Eric Hua Qing, et al.
Published: (2025) -
Clean & Clear: Feasibility of Safe LLM Clinical Guidance
by: Ive, Julia, et al.
Published: (2025) -
Hesitation is defeat? Connecting Linguistic and Predictive Uncertainty
by: Manzo, Gianluca, et al.
Published: (2025) -
Safe Training with Sensitive In-domain Data: Leveraging Data Fragmentation To Mitigate Linkage Attacks
by: Ignashina, Mariia, et al.
Published: (2024) -
Harnessing Large Language Models for Precision Querying and Retrieval-Augmented Knowledge Extraction in Clinical Data Science
by: Jan, Juan Jose Rubio, et al.
Published: (2026)