Arabic Automatic Story Generation with Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: El-Shangiti, Ahmed Oumar, Alwajih, Fakhraddin, Abdul-Mageed, Muhammad
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866914865096949760
author El-Shangiti, Ahmed Oumar
Alwajih, Fakhraddin
Abdul-Mageed, Muhammad
author_facet El-Shangiti, Ahmed Oumar
Alwajih, Fakhraddin
Abdul-Mageed, Muhammad
contents Large language models (LLMs) have recently emerged as a powerful tool for a wide range of language generation tasks. Nevertheless, this progress has been slower in Arabic. In this work, we focus on the task of generating stories from LLMs. For our training, we use stories acquired through machine translation (MT) as well as GPT-4. For the MT data, we develop a careful pipeline that ensures we acquire high-quality stories. For our GPT-41 data, we introduce crafted prompts that allow us to generate data well-suited to the Arabic context in both Modern Standard Arabic (MSA) and two Arabic dialects (Egyptian and Moroccan). For example, we generate stories tailored to various Arab countries on a wide host of topics. Our manual evaluation shows that our model fine-tuned on these training datasets can generate coherent stories that adhere to our instructions. We also conduct an extensive automatic and human evaluation comparing our models against state-of-the-art proprietary and open-source models. Our datasets and models will be made publicly available at https: //github.com/UBC-NLP/arastories.
format Preprint
id arxiv_https___arxiv_org_abs_2407_07551
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Arabic Automatic Story Generation with Large Language Models
El-Shangiti, Ahmed Oumar
Alwajih, Fakhraddin
Abdul-Mageed, Muhammad
Computation and Language
Artificial Intelligence
Large language models (LLMs) have recently emerged as a powerful tool for a wide range of language generation tasks. Nevertheless, this progress has been slower in Arabic. In this work, we focus on the task of generating stories from LLMs. For our training, we use stories acquired through machine translation (MT) as well as GPT-4. For the MT data, we develop a careful pipeline that ensures we acquire high-quality stories. For our GPT-41 data, we introduce crafted prompts that allow us to generate data well-suited to the Arabic context in both Modern Standard Arabic (MSA) and two Arabic dialects (Egyptian and Moroccan). For example, we generate stories tailored to various Arab countries on a wide host of topics. Our manual evaluation shows that our model fine-tuned on these training datasets can generate coherent stories that adhere to our instructions. We also conduct an extensive automatic and human evaluation comparing our models against state-of-the-art proprietary and open-source models. Our datasets and models will be made publicly available at https: //github.com/UBC-NLP/arastories.
title Arabic Automatic Story Generation with Large Language Models
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2407.07551