Legal Documents Drafting with Fine-Tuned Pre-Trained Large Language Model

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Lin, Chun-Hsien, Cheng, Pu-Jen
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866910475171659776
author Lin, Chun-Hsien
Cheng, Pu-Jen
author_facet Lin, Chun-Hsien
Cheng, Pu-Jen
contents With the development of large-scale Language Models (LLM), fine-tuning pre-trained LLM has become a mainstream paradigm for solving downstream tasks of natural language processing. However, training a language model in the legal field requires a large number of legal documents so that the language model can learn legal terminology and the particularity of the format of legal documents. The typical NLP approaches usually rely on many manually annotated data sets for training. However, in the legal field application, it is difficult to obtain a large number of manually annotated data sets, which restricts the typical method applied to the task of drafting legal documents. The experimental results of this paper show that not only can we leverage a large number of annotation-free legal documents without Chinese word segmentation to fine-tune a large-scale language model, but more importantly, it can fine-tune a pre-trained LLM on the local computer to achieve the generating legal document drafts task, and at the same time achieve the protection of information privacy and to improve information security issues.
format Preprint
id arxiv_https___arxiv_org_abs_2406_04202
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Legal Documents Drafting with Fine-Tuned Pre-Trained Large Language Model
Lin, Chun-Hsien
Cheng, Pu-Jen
Computation and Language
Artificial Intelligence
With the development of large-scale Language Models (LLM), fine-tuning pre-trained LLM has become a mainstream paradigm for solving downstream tasks of natural language processing. However, training a language model in the legal field requires a large number of legal documents so that the language model can learn legal terminology and the particularity of the format of legal documents. The typical NLP approaches usually rely on many manually annotated data sets for training. However, in the legal field application, it is difficult to obtain a large number of manually annotated data sets, which restricts the typical method applied to the task of drafting legal documents. The experimental results of this paper show that not only can we leverage a large number of annotation-free legal documents without Chinese word segmentation to fine-tune a large-scale language model, but more importantly, it can fine-tune a pre-trained LLM on the local computer to achieve the generating legal document drafts task, and at the same time achieve the protection of information privacy and to improve information security issues.
title Legal Documents Drafting with Fine-Tuned Pre-Trained Large Language Model
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2406.04202