Pith. sign in

REVIEW 1 cited by

Deterministic or probabilistic? The psychology of LLMs as random number generators

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.19965 v1 pith:ZBTHNPRE submitted 2025-02-27 cs.CL cs.AI

classification cs.CLcs.AI
keywords llmsbiaseslanguagemodelsrandomwhendespitedeterministic
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large Language Models (LLMs) have transformed text generation through inherently probabilistic context-aware mechanisms, mimicking human natural language. In this paper, we systematically investigate the performance of various LLMs when generating random numbers, considering diverse configurations such as different model architectures, numerical ranges, temperature, and prompt languages. Our results reveal that, despite their stochastic transformers-based architecture, these models often exhibit deterministic responses when prompted for random numerical outputs. In particular, we find significant differences when changing the model, as well as the prompt language, attributing this phenomenon to biases deeply embedded within the training data. Models such as DeepSeek-R1 can shed some light on the internal reasoning process of LLMs, despite arriving to similar results. These biases induce predictable patterns that undermine genuine randomness, as LLMs are nothing but reproducing our own human cognitive biases.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. One Token Is Enough: Fingerprinting and Verifying Large Language Models from Single-Token Output Distributions

    cs.CR 2026-07 conditional novelty 7.0 of 10

    Single-token answer distributions to everyday prompts fingerprint 165 served LLMs, recover family lineage, and verify claimed identity at 7.3% equal-error rate.

Pith tools