Saved in:
Bibliographic Details
Main Authors: Hsu, Enshuo, Roberts, Kirk
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2406.06723
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866916668123381760
author Hsu, Enshuo
Roberts, Kirk
author_facet Hsu, Enshuo
Roberts, Kirk
contents The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily available or affordable. Weak supervision and in-context learning offer partial solutions to this issue, particularly using large language models (LLMs), but their performance still trails traditional supervised methods with moderate amounts of gold-standard data. In particular, inferencing with LLMs is computationally heavy. We propose an approach leveraging fine-tuning LLMs and weak supervision with virtually no domain knowledge that still achieves consistently dominant performance. Using a prompt-based approach, the LLM is used to generate weakly-labeled data for training a downstream BERT model. The weakly supervised model is then further fine-tuned on small amounts of gold standard data. We evaluate this approach using Llama2 on three different n2c2 datasets. With no more than 10 gold standard notes, our final BERT models weakly supervised by fine-tuned Llama2-13B consistently outperformed out-of-the-box PubMedBERT by 4.7% to 47.9% in F1 scores. With only 50 gold standard notes, our models achieved close performance to fully fine-tuned systems.
format Preprint
id arxiv_https___arxiv_org_abs_2406_06723
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing
Hsu, Enshuo
Roberts, Kirk
Computation and Language
Information Retrieval
The performance of deep learning-based natural language processing systems is based on large amounts of labeled training data which, in the clinical domain, are not easily available or affordable. Weak supervision and in-context learning offer partial solutions to this issue, particularly using large language models (LLMs), but their performance still trails traditional supervised methods with moderate amounts of gold-standard data. In particular, inferencing with LLMs is computationally heavy. We propose an approach leveraging fine-tuning LLMs and weak supervision with virtually no domain knowledge that still achieves consistently dominant performance. Using a prompt-based approach, the LLM is used to generate weakly-labeled data for training a downstream BERT model. The weakly supervised model is then further fine-tuned on small amounts of gold standard data. We evaluate this approach using Llama2 on three different n2c2 datasets. With no more than 10 gold standard notes, our final BERT models weakly supervised by fine-tuned Llama2-13B consistently outperformed out-of-the-box PubMedBERT by 4.7% to 47.9% in F1 scores. With only 50 gold standard notes, our models achieved close performance to fully fine-tuned systems.
title Leveraging Large Language Models for Knowledge-free Weak Supervision in Clinical Natural Language Processing
topic Computation and Language
Information Retrieval
url https://arxiv.org/abs/2406.06723