PersLLM: A Personified Training Approach for Large Language Models

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Zeng, Zheni, Chen, Jiayi, Chen, Huimin, Yan, Yukun, Chen, Yuxuan, Liu, Zhenghao, Liu, Zhiyuan, Sun, Maosong
Format: Preprint
Published: 2024
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866918020799004672
author Zeng, Zheni
Chen, Jiayi
Chen, Huimin
Yan, Yukun
Chen, Yuxuan
Liu, Zhenghao
Liu, Zhiyuan
Sun, Maosong
author_facet Zeng, Zheni
Chen, Jiayi
Chen, Huimin
Yan, Yukun
Chen, Yuxuan
Liu, Zhenghao
Liu, Zhiyuan
Sun, Maosong
contents Large language models (LLMs) exhibit human-like intelligence, enabling them to simulate human behavior and support various applications that require both humanized communication and extensive knowledge reserves. Efforts are made to personify LLMs with special training data or hand-crafted prompts, while correspondingly faced with challenges such as insufficient data usage or rigid behavior patterns. Consequently, personified LLMs fail to capture personified knowledge or express persistent opinion. To fully unlock the potential of LLM personification, we propose PersLLM, a framework for better data construction and model tuning. For insufficient data usage, we incorporate strategies such as Chain-of-Thought prompting and anti-induction, improving the quality of data construction and capturing the personality experiences, knowledge, and thoughts more comprehensively. For rigid behavior patterns, we design the tuning process and introduce automated DPO to enhance the specificity and dynamism of the models' personalities, which leads to a more natural opinion communication. Both automated metrics and expert human evaluations demonstrate the effectiveness of our approach. Case studies in human-machine interactions and multi-agent systems further suggest potential application scenarios and future directions for LLM personification.
format Preprint
id arxiv_https___arxiv_org_abs_2407_12393
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle PersLLM: A Personified Training Approach for Large Language Models
Zeng, Zheni
Chen, Jiayi
Chen, Huimin
Yan, Yukun
Chen, Yuxuan
Liu, Zhenghao
Liu, Zhiyuan
Sun, Maosong
Computation and Language
Artificial Intelligence
Computers and Society
Large language models (LLMs) exhibit human-like intelligence, enabling them to simulate human behavior and support various applications that require both humanized communication and extensive knowledge reserves. Efforts are made to personify LLMs with special training data or hand-crafted prompts, while correspondingly faced with challenges such as insufficient data usage or rigid behavior patterns. Consequently, personified LLMs fail to capture personified knowledge or express persistent opinion. To fully unlock the potential of LLM personification, we propose PersLLM, a framework for better data construction and model tuning. For insufficient data usage, we incorporate strategies such as Chain-of-Thought prompting and anti-induction, improving the quality of data construction and capturing the personality experiences, knowledge, and thoughts more comprehensively. For rigid behavior patterns, we design the tuning process and introduce automated DPO to enhance the specificity and dynamism of the models' personalities, which leads to a more natural opinion communication. Both automated metrics and expert human evaluations demonstrate the effectiveness of our approach. Case studies in human-machine interactions and multi-agent systems further suggest potential application scenarios and future directions for LLM personification.
title PersLLM: A Personified Training Approach for Large Language Models
topic Computation and Language
Artificial Intelligence
Computers and Society
url https://arxiv.org/abs/2407.12393