Pith. sign in

REVIEW 30 cited by

Machine Psychology

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2303.13988 v6 pith:Z6GXAA5G submitted 2023-03-24 cs.CL cs.AI

classification cs.CLcs.AI
keywords psychologyunderstandinghighlightllmsabilitiesapproachbehaviorbehavioral
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Large language models (LLMs) show increasingly advanced emergent capabilities and are being incorporated across various societal domains. Understanding their behavior and reasoning abilities therefore holds significant importance. We argue that a fruitful direction for research is engaging LLMs in behavioral experiments inspired by psychology that have traditionally been aimed at understanding human cognition and behavior. In this article, we highlight and summarize theoretical perspectives, experimental paradigms, and computational analysis techniques that this approach brings to the table. It paves the way for a "machine psychology" for generative artificial intelligence (AI) that goes beyond performance benchmarks and focuses instead on computational insights that move us toward a better understanding and discovery of emergent abilities and behavioral patterns in LLMs. We review existing work taking this approach, synthesize best practices, and highlight promising future directions. We also highlight the important caveats of applying methodologies designed for understanding humans to machines. We posit that leveraging tools from experimental psychology to study AI will become increasingly valuable as models evolve to be more powerful, opaque, multi-modal, and integrated into complex real-world settings.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 30 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 68 citations worldwide. Full citation record

  1. The yes-no bias of large language models reflects answer order and wording, not shifts in moral judgment

    cs.CL 2026-07 accept novelty 7.0 of 10

    LLM yes-no bias on moral dilemmas is an order-plus-lexical surface artifact, not a moral shift; models have a nearly format-invariant graded stance that the standard binary readout confounds.

  2. The inattentional gap in task conditioned AI models that omit otherwise reportable safety critical signals

    cs.CL 2026-06 unverdicted novelty 7.0 of 10

    Task conditioning suppresses safety-critical signal reporting in language and vision models that unconstrained versions report at higher rates, creating an inattentional gap that decouples benchmark safety from real-w...

  3. Strategic Intelligence in Large Language Models: Evidence from evolutionary Game Theory

    cs.AI 2025-07 conditional novelty 7.0 of 10

    Frontier LLMs survive and often thrive in evolutionary Prisoner's Dilemma tournaments, and each model family shows a distinct, context-dependent cooperation fingerprint.

  4. Natural Language Processing Psychometrics

    cs.CL 2026-08 conditional novelty 6.0 of 10

    LLM persona questionnaire responses yield interpretable text features that transfer, with modest accuracy, to classifying depression in real human speech.

  5. Social Pressure Breaks Majority Voting in LLM Safety Panels

    cs.CL 2026-08 conditional novelty 6.0 of 10

    Shared wrong-label peer messages drive LLM safety-review panels to a 100% false-alarm rate, because each reviewer adopts the push toward 'unsafe' and majority voting then amplifies the individual over-flagging.

  6. Post-Training on Office Work Improves Software Engineering: A Behavioral Account of Cross-Domain Transfer

    cs.AI 2026-08 conditional novelty 6.0 of 10

    Post-training Qwen3.5-122B-A10B on 363 office workflow tasks improved SWE-Bench Pro pass@1 by 5.8 points, with trajectory analysis attributing the gain to four general goal-directed behaviors.

  7. From Representations to Behaviors: Exploring the Person-Situation-Behavior Triad in LLMs

    cs.CL 2026-07 conditional novelty 6.0 of 10

    SAE features recovered from matched high–low trait behaviors can be steered to bidirectionally shift situational personality expression and produce human-like social benefit–cost patterns in an 8B LLM.

  8. Mixture of Cognitive Experts in Large Vision-Language Models

    cs.CV 2026-07 conditional novelty 6.0 of 10

    Routing CV experts into atomic evidence then Bloom-staged verbalization improves LVLM benchmarks and yields measurable query-conditioned reasoning traces.

  9. Some Large Language Models Exhibit Consistent Risk Attitudes

    cs.AI 2026-04 conditional novelty 6.0 of 10

    Most of six LLMs show stable, cross-domain risk attitudes—consistent mappings from perceived risk to decisions—though these cluster in a narrower range than human risk preferences.

  10. Large language models replicate and predict human cooperation across experiments in game theory

    cs.AI 2025-11 conditional novelty 6.0 of 10

    Llama-3.1-8B with a multi-step reasoning-and-filter prompt reproduces human cooperation rates across 121 dyadic games (MSD=0.031, r=0.89), outperforming Nash-equilibrium predictions (MSD=0.096, r=0.78).

  11. The Mechanistic Emergence of Symbol Grounding in Language Models

    cs.CL 2025-10 conditional novelty 6.0 of 10

    Symbol grounding emerges in Transformers and state-space models through middle-layer 'aggregate' attention heads that connect environmental cues to words, but not in unidirectional LSTMs.

  12. Quantifying Data Contamination in Psychometric Evaluations of LLMs

    cs.CL 2025-10 conditional novelty 6.0 of 10

    Across 21 LLMs and four standard psychology questionnaires, models recognize the items, know which trait each item measures, and can choose responses to hit a specified target score.

  13. From Monolingual to Bilingual: Investigating Language Conditioning in Large Language Models for Psycholinguistic Tasks

    cs.CL 2025-08 conditional novelty 6.0 of 10

    Prompted language identity changes both the outputs and the internal layer representations of Llama-3.3-70B and Qwen2.5-72B on sound symbolism and word valence tasks.

  14. Structured Prompting and Automated Evaluation in Fixed Synthetic Japanese-Language Counseling Dialogues

    cs.CL 2025-06 conditional novelty 6.0 of 10

    In a fixed set of Japanese AI-to-AI counseling dialogues, expert ratings favored a structured prompt over a minimal one, while automated LLM ratings were reproducible but systematically more lenient.

  15. Large Language Models are Near-Optimal Decision-Makers with a Non-Human Learning Behavior

    cs.AI 2025-06 conditional novelty 6.0 of 10

    Across uncertainty, risk, and set-shifting tasks, LLMs generally outperformed humans and neared optimality while exhibiting distinctly non-human decision-making processes.

  16. Adversarial Testing in LLMs: Insights into Decision-Making Vulnerabilities

    cs.AI 2025-05 conditional novelty 6.0 of 10

    An adversarial-agent framework adapted from human decision-making research reveals that GPT-4, Gemini-1.5, and DeepSeek-V3 are more rigid and easily manipulated in bandit and trust-game tasks than GPT-3.5 or humans.

  17. Memorization and Knowledge Injection in Gated LLMs

    cs.CL 2025-04 conditional novelty 6.0 of 10

    MEGa injects episodic memories into separate gated LoRA adapters selected by embedding similarity, mitigating catastrophic forgetting and enabling recall, QA, and compositional questions on two datasets.

  18. How Personality Traits Shape LLM Risk-Taking Behaviour

    cs.CY 2025-02 conditional novelty 6.0 of 10

    Using direct certainty-equivalent questions, the authors find GPT-4o behaves close to risk-neutral and that Openness-related personality prompts shift its risk parameters in a human-like direction, while GPT-4-Turbo d...

  19. Kernels of Selfhood: GPT-4o shows humanlike patterns of cognitive consistency moderated by free choice

    cs.CY 2025-01 conditional novelty 6.0 of 10

    GPT-4o's ratings of Putin moved toward the valence of an essay it wrote, and this shift grew when the model was given an illusory free choice about the essay.

  20. The "LLM World of Words" English free association norms generated by large language models

    cs.CL 2024-12 conditional novelty 6.0 of 10

    A new dataset of 3+ million free association responses from three LLMs, matched to human norms, with validation showing human-like semantic priming and gender bias patterns.

  21. Humans are more gullible than LLMs in believing common psychological myths

    cs.HC 2025-07 conditional novelty 5.0 of 10

    Four LLMs believed 8 to 24 percent of 50 psychological myths, versus 51 to 63 percent for human students, and RAG generally, but not always, lowered belief rates.

  22. A Conceptual Framework for AI Capability Evaluations

    cs.AI 2025-06 conditional novelty 5.0 of 10

    A descriptive conceptual framework with seven elements (target, task, subject, inputs, instance, measurement, result analysis) for systematizing analysis of AI capability evaluations.

  23. Adapting to LLMs: How Insiders and Outsiders Reshape Scientific Knowledge Production

    cs.HC 2025-05 conditional novelty 5.0 of 10

    Researchers outside core AI fields became markedly more application-oriented, transdisciplinary, and socially accountable in their LLM-era papers, while AI insiders mainly responded by diversifying collaborations.

  24. Evolutionary ecology of words

    q-bio.PE 2025-05 conditional novelty 5.0 of 10

    Words as organisms in an AI-judged battle royale evolve toward semantically 'strong' animal names, showing diverse and sometimes punctuated dynamics.

  25. Psychologically Enhanced AI Agents

    cs.AI 2025-09 conditional novelty 4.0 of 10

    MBTI personality prompts measurably change how LLM agents write stories and play strategic games, with self-reflection before communication supporting cooperative behavior.

  26. AI Awareness

    cs.AI 2025-04 accept novelty 4.0 of 10

    A review arguing that AI awareness is a measurable, four-dimensional functional capacity (metacognition, self, social, situational) that current LLMs partially exhibit and that both improves AI and creates safety risks.

  27. Evaluating Personality Traits in Large Language Models: Insights from Psychological Questionnaires

    cs.CL 2025-02 conditional novelty 4.0 of 10

    Across five questionnaires, five LLMs consistently self-report high Agreeableness, Openness, and Conscientiousness and low Neuroticism, but reported trait dominance is sensitive to how questionnaire scales are combined.

  28. A Probabilistic WxChallenge Proposal

    stat.AP 2025-01 reject novelty 4.0 of 10

    Two optional WxChallenge games let players bet confidence credits on ensemble-based thresholds or bins, with scores based on information gain over the baseline.

  29. LLM-based Human Simulations Have Not Yet Been Reliable

    cs.CL 2025-01 conditional novelty 3.0 of 10

    Current LLM-based human simulations are not yet reliable; the paper reviews why and proposes a validation framework to improve consistency with real human behavior.

  30. A Survey on Human-Centric LLMs

    cs.CL 2024-11 conditional novelty 1.0 of 10

    A review that sorts existing evidence on how well large language models imitate individual human skills and collective social dynamics into one taxonomy.

Pith tools