Pith. sign in

REVIEW 1 cited by

A Turing Test: Are AI Chatbots Behaviorally Similar to Humans?

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2312.00798 v2 pith:NU6JKYEC submitted 2023-11-19 cs.AI

classification cs.AI
keywords theychatbotshumanaveragebehavebehaviorbehavioralbehaviors
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We administer a Turing Test to AI Chatbots. We examine how Chatbots behave in a suite of classic behavioral games that are designed to elicit characteristics such as trust, fairness, risk-aversion, cooperation, \textit{etc.}, as well as how they respond to a traditional Big-5 psychological survey that measures personality traits. ChatGPT-4 exhibits behavioral and personality traits that are statistically indistinguishable from a random human from tens of thousands of human subjects from more than 50 countries. Chatbots also modify their behavior based on previous experience and contexts ``as if'' they were learning from the interactions, and change their behavior in response to different framings of the same strategic situation. Their behaviors are often distinct from average and modal human behaviors, in which case they tend to behave on the more altruistic and cooperative end of the distribution. We estimate that they act as if they are maximizing an average of their own and partner's payoffs.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LLMs Model Non-WEIRD Populations: Experiments with Synthetic Cultural Agents

    cs.AI 2025-01 conditional novelty 5.0 of 10

    Synthetic cultural agents built by injecting RAG-generated profiles into ChatGPT show cross-cultural variation in dictator and ultimatum game choices, with behavior the authors claim qualitatively resembles human smal...

Pith tools