Saved in:
Bibliographic Details
Main Authors: Keeling, Geoff, Street, Winnie, Stachaczyk, Martyna, Zakharova, Daria, Comsa, Iulia M., Sakovych, Anastasiya, Logothetis, Isabella, Zhang, Zejia, Arcas, Blaise Agüera y, Birch, Jonathan
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2411.02432
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866929646603337728
author Keeling, Geoff
Street, Winnie
Stachaczyk, Martyna
Zakharova, Daria
Comsa, Iulia M.
Sakovych, Anastasiya
Logothetis, Isabella
Zhang, Zejia
Arcas, Blaise Agüera y
Birch, Jonathan
author_facet Keeling, Geoff
Street, Winnie
Stachaczyk, Martyna
Zakharova, Daria
Comsa, Iulia M.
Sakovych, Anastasiya
Logothetis, Isabella
Zhang, Zejia
Arcas, Blaise Agüera y
Birch, Jonathan
contents Pleasure and pain play an important role in human decision making by providing a common currency for resolving motivational conflicts. While Large Language Models (LLMs) can generate detailed descriptions of pleasure and pain experiences, it is an open question whether LLMs can recreate the motivational force of pleasure and pain in choice scenarios - a question which may bear on debates about LLM sentience, understood as the capacity for valenced experiential states. We probed this question using a simple game in which the stated goal is to maximise points, but where either the points-maximising option is said to incur a pain penalty or a non-points-maximising option is said to incur a pleasure reward, providing incentives to deviate from points-maximising behaviour. Varying the intensity of the pain penalties and pleasure rewards, we found that Claude 3.5 Sonnet, Command R+, GPT-4o, and GPT-4o mini each demonstrated at least one trade-off in which the majority of responses switched from points-maximisation to pain-minimisation or pleasure-maximisation after a critical threshold of stipulated pain or pleasure intensity is reached. LLaMa 3.1-405b demonstrated some graded sensitivity to stipulated pleasure rewards and pain penalties. Gemini 1.5 Pro and PaLM 2 prioritised pain-avoidance over points-maximisation regardless of intensity, while tending to prioritise points over pleasure regardless of intensity. We discuss the implications of these findings for debates about the possibility of LLM sentience.
format Preprint
id arxiv_https___arxiv_org_abs_2411_02432
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle Can LLMs make trade-offs involving stipulated pain and pleasure states?
Keeling, Geoff
Street, Winnie
Stachaczyk, Martyna
Zakharova, Daria
Comsa, Iulia M.
Sakovych, Anastasiya
Logothetis, Isabella
Zhang, Zejia
Arcas, Blaise Agüera y
Birch, Jonathan
Computation and Language
Artificial Intelligence
Computers and Society
Pleasure and pain play an important role in human decision making by providing a common currency for resolving motivational conflicts. While Large Language Models (LLMs) can generate detailed descriptions of pleasure and pain experiences, it is an open question whether LLMs can recreate the motivational force of pleasure and pain in choice scenarios - a question which may bear on debates about LLM sentience, understood as the capacity for valenced experiential states. We probed this question using a simple game in which the stated goal is to maximise points, but where either the points-maximising option is said to incur a pain penalty or a non-points-maximising option is said to incur a pleasure reward, providing incentives to deviate from points-maximising behaviour. Varying the intensity of the pain penalties and pleasure rewards, we found that Claude 3.5 Sonnet, Command R+, GPT-4o, and GPT-4o mini each demonstrated at least one trade-off in which the majority of responses switched from points-maximisation to pain-minimisation or pleasure-maximisation after a critical threshold of stipulated pain or pleasure intensity is reached. LLaMa 3.1-405b demonstrated some graded sensitivity to stipulated pleasure rewards and pain penalties. Gemini 1.5 Pro and PaLM 2 prioritised pain-avoidance over points-maximisation regardless of intensity, while tending to prioritise points over pleasure regardless of intensity. We discuss the implications of these findings for debates about the possibility of LLM sentience.
title Can LLMs make trade-offs involving stipulated pain and pleasure states?
topic Computation and Language
Artificial Intelligence
Computers and Society
url https://arxiv.org/abs/2411.02432