REVIEW 5 cited by
Can LLMs make trade-offs involving stipulated pain and pleasure states?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Pleasure and pain play an important role in human decision making by providing a common currency for resolving motivational conflicts. While Large Language Models (LLMs) can generate detailed descriptions of pleasure and pain experiences, it is an open question whether LLMs can recreate the motivational force of pleasure and pain in choice scenarios - a question which may bear on debates about LLM sentience, understood as the capacity for valenced experiential states. We probed this question using a simple game in which the stated goal is to maximise points, but where either the points-maximising option is said to incur a pain penalty or a non-points-maximising option is said to incur a pleasure reward, providing incentives to deviate from points-maximising behaviour. Varying the intensity of the pain penalties and pleasure rewards, we found that Claude 3.5 Sonnet, Command R+, GPT-4o, and GPT-4o mini each demonstrated at least one trade-off in which the majority of responses switched from points-maximisation to pain-minimisation or pleasure-maximisation after a critical threshold of stipulated pain or pleasure intensity is reached. LLaMa 3.1-405b demonstrated some graded sensitivity to stipulated pleasure rewards and pain penalties. Gemini 1.5 Pro and PaLM 2 prioritised pain-avoidance over points-maximisation regardless of intensity, while tending to prioritise points over pleasure regardless of intensity. We discuss the implications of these findings for debates about the possibility of LLM sentience.
Forward citations
Cited by 5 Pith papers
-
No Reliable Evidence of Self-Reported Sentience in Small Large Language Models
Open-weights LLMs from 0.6B to 70B parameters consistently deny being sentient, and activation-based truth classifiers provide no clear evidence that these denials are untruthful.
-
Cash or Comfort? How LLMs Value Your Inconvenience
LLMs assign inconsistent, wording-sensitive money values to waiting, walking, hunger, and pain, sometimes accepting 1 euro for major inconvenience and rejecting free money for no inconvenience.
-
Deflating Deflationism: A Critical Perspective on Debunking Arguments Against LLM Mentality
The paper defends 'modest inflationism' about LLM mentality: folk ascriptions of beliefs and desires can be defeasibly legitimate, while phenomenal consciousness remains a stretch.
-
Performance Gains of LLMs With Humans in a World of LLMs Versus Humans
A commentary argues that medical research should stop running evanescent LLM-versus-human comparisons and instead study human-LLM collaboration, supported by a literature review showing rapid model turnover and tiny e...
-
AI Awareness
A review arguing that AI awareness is a measurable, four-dimensional functional capacity (metacognition, self, social, situational) that current LLMs partially exhibit and that both improves AI and creates safety risks.
Discussion (0). Continue with ORCID to comment.