Deanthropomorphising NLP: Can a Language Model Be Conscious?

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Shardlow, Matthew, Przybyła, Piotr
Format: Preprint
Published: 2022
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866917866522017792
author Shardlow, Matthew
Przybyła, Piotr
author_facet Shardlow, Matthew
Przybyła, Piotr
contents This work is intended as a voice in the discussion over previous claims that a pretrained large language model (LLM) based on the Transformer model architecture can be sentient. Such claims have been made concerning the LaMDA model and also concerning the current wave of LLM-powered chatbots, such as ChatGPT. This claim, if confirmed, would have serious ramifications in the Natural Language Processing (NLP) community due to wide-spread use of similar models. However, here we take the position that such a large language model cannot be sentient, or conscious, and that LaMDA in particular exhibits no advances over other similar models that would qualify it. We justify this by analysing the Transformer architecture through Integrated Information Theory of consciousness. We see the claims of sentience as part of a wider tendency to use anthropomorphic language in NLP reporting. Regardless of the veracity of the claims, we consider this an opportune moment to take stock of progress in language modelling and consider the ethical implications of the task. In order to make this work helpful for readers outside the NLP community, we also present the necessary background in language modelling.
format Preprint
id arxiv_https___arxiv_org_abs_2211_11483
institution arXiv
publishDate 2022
record_format arxiv
spellingShingle Deanthropomorphising NLP: Can a Language Model Be Conscious?
Shardlow, Matthew
Przybyła, Piotr
Computation and Language
Artificial Intelligence
This work is intended as a voice in the discussion over previous claims that a pretrained large language model (LLM) based on the Transformer model architecture can be sentient. Such claims have been made concerning the LaMDA model and also concerning the current wave of LLM-powered chatbots, such as ChatGPT. This claim, if confirmed, would have serious ramifications in the Natural Language Processing (NLP) community due to wide-spread use of similar models. However, here we take the position that such a large language model cannot be sentient, or conscious, and that LaMDA in particular exhibits no advances over other similar models that would qualify it. We justify this by analysing the Transformer architecture through Integrated Information Theory of consciousness. We see the claims of sentience as part of a wider tendency to use anthropomorphic language in NLP reporting. Regardless of the veracity of the claims, we consider this an opportune moment to take stock of progress in language modelling and consider the ethical implications of the task. In order to make this work helpful for readers outside the NLP community, we also present the necessary background in language modelling.
title Deanthropomorphising NLP: Can a Language Model Be Conscious?
topic Computation and Language
Artificial Intelligence
url https://arxiv.org/abs/2211.11483