Towards Efficient Large Language Models for Scientific Text: A Review

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: To, Huy Quoc, Liu, Ming, Huang, Guangyan
Natura: Preprint
Pubblicazione: 2024
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866916362931142656
author To, Huy Quoc
Liu, Ming
Huang, Guangyan
author_facet To, Huy Quoc
Liu, Ming
Huang, Guangyan
contents Large language models (LLMs) have ushered in a new era for processing complex information in various fields, including science. The increasing amount of scientific literature allows these models to acquire and understand scientific knowledge effectively, thus improving their performance in a wide range of tasks. Due to the power of LLMs, they require extremely expensive computational resources, intense amounts of data, and training time. Therefore, in recent years, researchers have proposed various methodologies to make scientific LLMs more affordable. The most well-known approaches align in two directions. It can be either focusing on the size of the models or enhancing the quality of data. To date, a comprehensive review of these two families of methods has not yet been undertaken. In this paper, we (I) summarize the current advances in the emerging abilities of LLMs into more accessible AI solutions for science, and (II) investigate the challenges and opportunities of developing affordable solutions for scientific domains using LLMs.
format Preprint
id arxiv_https___arxiv_org_abs_2408_10729
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Towards Efficient Large Language Models for Scientific Text: A Review
To, Huy Quoc
Liu, Ming
Huang, Guangyan
Computation and Language
Artificial Intelligence
Large language models (LLMs) have ushered in a new era for processing complex information in various fields, including science. The increasing amount of scientific literature allows these models to acquire and understand scientific knowledge effectively, thus improving their performance in a wide range of tasks. Due to the power of LLMs, they require extremely expensive computational resources, intense amounts of data, and training time. Therefore, in recent years, researchers have proposed various methodologies to make scientific LLMs more affordable. The most well-known approaches align in two directions. It can be either focusing on the size of the models or enhancing the quality of data. To date, a comprehensive review of these two families of methods has not yet been undertaken. In this paper, we (I) summarize the current advances in the emerging abilities of LLMs into more accessible AI solutions for science, and (II) investigate the challenges and opportunities of developing affordable solutions for scientific domains using LLMs.
title Towards Efficient Large Language Models for Scientific Text: A Review
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2408.10729