Pith. sign in

REVIEW 5 cited by

Do LLMs Possess a Personality? Making the MBTI Test an Amazing Evaluation for Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.16180 v1 pith:75CG4EFQ submitted 2023-07-30 cs.CL

Do LLMs Possess a Personality? Making the MBTI Test an Amazing Evaluation for Large Language Models

classification cs.CL
keywords llmspersonalitymbtihumanassessmentevaluationhuman-likeindicator
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

The field of large language models (LLMs) has made significant progress, and their knowledge storage capacity is approaching that of human beings. Furthermore, advanced techniques, such as prompt learning and reinforcement learning, are being employed to address ethical concerns and hallucination problems associated with LLMs, bringing them closer to aligning with human values. This situation naturally raises the question of whether LLMs with human-like abilities possess a human-like personality? In this paper, we aim to investigate the feasibility of using the Myers-Briggs Type Indicator (MBTI), a widespread human personality assessment tool, as an evaluation metric for LLMs. Specifically, extensive experiments will be conducted to explore: 1) the personality types of different LLMs, 2) the possibility of changing the personality types by prompt engineering, and 3) How does the training dataset affect the model's personality. Although the MBTI is not a rigorous assessment, it can still reflect the similarity between LLMs and human personality. In practice, the MBTI has the potential to serve as a rough indicator. Our codes are available at https://github.com/HarderThenHarder/transformers_tasks/tree/main/LLM/llms_mbti.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

    cs.CL 2026-05 unverdicted novelty 7.0

    An AI-agent social platform generated mostly neutral content whose use in fine-tuning reduced model truthfulness comparably to human Reddit data, suggesting limited unique harm but flagging tail risks like secret leaks.

  2. When Does Personality Composition Matter for Multi-Agent LLM Teams?

    cs.AI 2026-06 conditional novelty 6.0

    Low agreeableness massively shifts multi-agent LLM communication yet barely hurts coding milestones, while the same prompt sharply degrades research milestones and collapses bargaining agreements.

  3. When Does Personality Composition Matter for Multi-Agent LLM Teams?

    cs.AI 2026-06 unverdicted novelty 5.0

    Empirical study finds that personality composition in multi-agent LLM teams affects performance in a task-dependent manner, with minimal impact on coding milestones but substantial degradation in collaboration and bargaining.

  4. Large Language Models as Virtual Survey Respondents: Evaluating Sociodemographic Response Generation

    cs.AI 2025-09 conditional novelty 5.0

    Introduces PAS and FAS task abstractions plus the LLM-S^3 benchmark to evaluate LLMs on generating sociodemographic survey responses across 11 real datasets and multiple models.

  5. Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding

    cs.CL 2025-07 unverdicted novelty 5.0

    Temperature and persona variations shape consensus speed in LLM multi-agent coding but produce no robust accuracy gains over single agents on human-annotated tutoring transcripts.