Pith. sign in

REVIEW 1 cited by

BIG5-CHAT: Shaping LLM Personalities Through Training on Human-Grounded Data

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.16491 v3 pith:YYTSWQG2 submitted 2024-10-21 cs.CL

classification cs.CL
keywords humanpersonalitymethodstraitsbig5-chatdatadatasethigher
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this work, we tackle the challenge of embedding realistic human personality traits into LLMs. Previous approaches have primarily focused on prompt-based methods that describe the behavior associated with the desired personality traits, suffering from realism and validity issues. To address these limitations, we introduce BIG5-CHAT, a large-scale dataset containing 100,000 dialogues designed to ground models in how humans express their personality in language. Leveraging this dataset, we explore Supervised Fine-Tuning and Direct Preference Optimization as training-based methods to align LLMs more naturally with human personality patterns. Our methods outperform prompting on personality assessments such as BFI and IPIP-NEO, with trait correlations more closely matching human data. Furthermore, our experiments reveal that models trained to exhibit higher conscientiousness, higher agreeableness, lower extraversion, and lower neuroticism display better performance on reasoning tasks, aligning with psychological findings on how these traits impact human cognitive performance. To our knowledge, this work is the first comprehensive study to demonstrate how training-based methods can shape LLM personalities through learning from real human behaviors.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Pok\'eAI: A Goal-Generating, Battle-Optimizing Multi-agent System for Pokemon Red

    cs.AI 2025-06 conditional novelty 3.0 of 10

    A text-based LLM battle agent reaches an 80.8% win rate on Pokémon Red wild battles, close to a single human run of 86%, with different models showing distinct playstyles.

Pith tools