Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning

Fuente: arXiv
Salvato in:
Dettagli Bibliografici
Autori principali: Long, Do Xuan, Kawaguchi, Kenji, Kan, Min-Yen, Chen, Nancy F.
Natura: Preprint
Pubblicazione: 2023
Soggetti:
Accesso online:
Tags: Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
_version_ 1866909427605438464
author Long, Do Xuan
Kawaguchi, Kenji
Kan, Min-Yen
Chen, Nancy F.
author_facet Long, Do Xuan
Kawaguchi, Kenji
Kan, Min-Yen
Chen, Nancy F.
contents Reasoning and predicting human opinions with large language models (LLMs) is essential yet challenging. Current methods employ role-playing with personae but face two major issues: LLMs are sensitive to even a single irrelevant persona, skewing predictions by up to 30%, and LLMs fail to reason strategically over personae. We propose Chain-of-Opinion (COO), a simple four-step solution modeling which and how to reason with personae, inspired by the Value--Belief--Norm (VBN) theory. COO differentiates between explicit personae (demographics and ideology) and implicit personae (historical opinions), involves: (1) filtering irrelevant attributes from explicit personae, (2) ranking implicit personae into a preferential list for selecting top-k, (3) applying novel VBN reasoning to extract user environmental and personal value, belief, and norm variables for accurate and reliable predictions, and (4) iterating VBN reasoning with progressively larger lists of implicit personae to handle potential persona insufficiency. COO efficiently achieves new state-of-the-art opinion prediction via prompting with only 5 inference calls, improving prior techniques by up to 4%. Notably, fine-tuning LMs with COO data results in significantly better opinion-aligned models, by up to 23%.
format Preprint
id arxiv_https___arxiv_org_abs_2311_08385
institution arXiv
publishDate 2023
record_format arxiv
spellingShingle Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning
Long, Do Xuan
Kawaguchi, Kenji
Kan, Min-Yen
Chen, Nancy F.
Computation and Language
Reasoning and predicting human opinions with large language models (LLMs) is essential yet challenging. Current methods employ role-playing with personae but face two major issues: LLMs are sensitive to even a single irrelevant persona, skewing predictions by up to 30%, and LLMs fail to reason strategically over personae. We propose Chain-of-Opinion (COO), a simple four-step solution modeling which and how to reason with personae, inspired by the Value--Belief--Norm (VBN) theory. COO differentiates between explicit personae (demographics and ideology) and implicit personae (historical opinions), involves: (1) filtering irrelevant attributes from explicit personae, (2) ranking implicit personae into a preferential list for selecting top-k, (3) applying novel VBN reasoning to extract user environmental and personal value, belief, and norm variables for accurate and reliable predictions, and (4) iterating VBN reasoning with progressively larger lists of implicit personae to handle potential persona insufficiency. COO efficiently achieves new state-of-the-art opinion prediction via prompting with only 5 inference calls, improving prior techniques by up to 4%. Notably, fine-tuning LMs with COO data results in significantly better opinion-aligned models, by up to 23%.
title Aligning Large Language Models with Human Opinions through Persona Selection and Value--Belief--Norm Reasoning
topic Computation and Language
url https://arxiv.org/abs/2311.08385