Pith. sign in

Paper Citation Record · LEDGER

LLM-based Human Simulations Have Not Yet Been Reliable

As of 22 August 2026, this Paper Citation Record lists 7 of 7 outbound references and 11 inbound Pith citation observations for arXiv:2501.08579.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.08579 v3

Coverage vector

measured 7 of 7 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:24:39.451190Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:57:11.383188Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

7 of 7 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation b945e715-2bc8-4460-bcaa-b799810c3a82 · outbound

This paper cites Evaluating the Performance of Large Language Models via Debates.

LLM-based Human Simulations Have Not Yet Been Reliable Evaluating the Performance of Large Language Models via Debates

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.437473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.437473Z digest=sha256:defabaeab07573500e5812965e20d7647ae25f49db191f563aff54b7ddd7fca0

Observation 0969c565-398a-446a-8852-52533896c683 · outbound

This paper cites LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals.

LLM-based Human Simulations Have Not Yet Been Reliable LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.441710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.441710Z digest=sha256:0589d5a7cc823a9b37da4683b7883fb13166f0ba465a985b1c0c545a49fba1d7

Observation df914cff-dab9-4f11-87a4-f0531eff5b46 · outbound

This paper cites From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings.

LLM-based Human Simulations Have Not Yet Been Reliable From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.446373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.446373Z digest=sha256:62a342f4d3e7bd097a5629362d8523247d1b1f742bd300a5130d64ad663b2041

Observation 44933d77-545b-4ecd-8d3f-bbf166a94eda · outbound

This paper cites Cognitive Overload Attack:Prompt Injection for Long Context.

LLM-based Human Simulations Have Not Yet Been Reliable Cognitive Overload Attack:Prompt Injection for Long Context

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.451190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.451190Z digest=sha256:dbf857e5a36fc41303fca8f77f3861f13b33a3291b7cefe291763388593ac8e7

Observation ec20af9e-bd96-42ce-a393-c16b2004ad4c · outbound

This paper cites Machine Psychology.

LLM-based Human Simulations Have Not Yet Been Reliable Machine Psychology

Reference 1971

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.422739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.422739Z digest=sha256:9316319cb479dd059e292ffc438e0b5f265ae257eb492aac5b764bc4d0210073

Observation a0b85657-327d-468b-9c6c-1d8173e80ab8 · outbound

This paper cites Navigating LLM Ethics: Advancements, Challenges, and Future Directions.

LLM-based Human Simulations Have Not Yet Been Reliable Navigating LLM Ethics: Advancements, Challenges, and Future Directions

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.428495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.428495Z digest=sha256:58f18de5b5c6f385cca6beb793d3fe965c0d36e1037eda264d66e74df92defab

Observation 20e7242e-13fc-4bc2-bbdf-0e0160359ff1 · outbound

This paper cites Exploring the Frontiers of LLMs in Psychological Applications: A Comprehensive Review.

LLM-based Human Simulations Have Not Yet Been Reliable Exploring the Frontiers of LLMs in Psychological Applications: A Comprehensive Review

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-10T20:24:39.433008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:24:39.433008Z digest=sha256:6e4bd89c4ed15e6b8dee67811f206815dc969f1d7aa721997eca23b1ae066cca

Pith citing papers

Observation 6c508819-aae0-48d9-a58b-a3125ae58f7b · inbound

PulseReddit: A Novel Reddit Dataset for Benchmarking MAS in High-Frequency Cryptocurrency Trading cites this paper.

PulseReddit: A Novel Reddit Dataset for Benchmarking MAS in High-Frequency Cryptocurrency Trading LLM-based Human Simulations Have Not Yet Been Reliable

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:57:11.383188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:57:11.383188Z digest=sha256:e1eefbec45d29234ccc0801c0e63d918ebec80beee5006ab44a15ac13c58e87c

Observation e6204bc6-531b-4604-9a15-1170d9ef14d4 · inbound

PUB: An LLM-Enhanced Personality-Driven User Behaviour Simulator for Recommender System Evaluation cites this paper.

PUB: An LLM-Enhanced Personality-Driven User Behaviour Simulator for Recommender System Evaluation LLM-based Human Simulations Have Not Yet Been Reliable

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:46:13.383319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:46:13.383319Z digest=sha256:02166b941d70e545017693b56c779f11226756292f07c199523db97c2c9af1cf

Observation ad387865-d06d-424c-b09c-c3c94edd915c · inbound

SimuPanel: A Novel Immersive Multi-Agent System to Simulate Interactive Expert Panel Discussion cites this paper.

SimuPanel: A Novel Immersive Multi-Agent System to Simulate Interactive Expert Panel Discussion LLM-based Human Simulations Have Not Yet Been Reliable

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T23:48:49.834154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:48:49.834154Z digest=sha256:ac39bc75681b3da7436b78a891fa0a9e10fbd64fbaaa34967402585397742da2

Observation 44ddd93c-3dba-4217-9a0c-92558c32dbef · inbound

Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support cites this paper.

Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support LLM-based Human Simulations Have Not Yet Been Reliable

Reference 133

Resolution
unresolved
no resolver link, observed 2026-08-04T23:43:04.470676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:43:04.470676Z digest=sha256:d4183bdb00308bbec88aab9de9f019c5240088d2ae68cf176343622549890b2d

Observation a0a6d939-9416-408d-b8db-9c6ac09dfc53 · inbound

The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies cites this paper.

The PIMMUR Principles: Ensuring Validity in Collective Behavior of LLM Societies LLM-based Human Simulations Have Not Yet Been Reliable

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T14:31:11.974395Z digest=sha256:1dee42e03c94bf8fe54e1ec9f161e512f93ebed6e1ea03d6d37456c879ba9d90

Observation 5e162e1c-84c1-4ba3-960c-f7d74c6a7632 · inbound

We Need Strong Preconditions For Using Simulations In Policy cites this paper.

We Need Strong Preconditions For Using Simulations In Policy LLM-based Human Simulations Have Not Yet Been Reliable

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:17:55.624493Z digest=sha256:84d01c229f602ccc2135baf17cfc6ae3de20b40d04d16cba31f2d16e2347e2b6

Observation d5dfb85f-5c81-4f23-b30b-03332afbea4c · inbound

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models cites this paper.

The Granularity Axis: A Micro-to-Macro Latent Direction for Social Roles in Language Models LLM-based Human Simulations Have Not Yet Been Reliable

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T10:11:08.494758Z digest=sha256:a9270a6db215451d0d440c061773fb52f14b7d8cfdd3eb661f3d035d5189ed2f

Observation 39729a45-696b-44f3-b706-f41eef5f010d · inbound

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench cites this paper.

Can LLMs Think Like Consumers? Benchmarking Crowd-Level Reaction Reconstruction with ConsumerSimBench LLM-based Human Simulations Have Not Yet Been Reliable

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T15:31:25.079191Z digest=sha256:827b2a91c8dfbd0d302adb8cd1ef646b24cec3d98f3521323826fd2ec48f7f24

Observation 8863393e-3b6c-4609-8013-66e0fa8c715a · inbound

Simulating Human Memory with Language Models cites this paper.

Simulating Human Memory with Language Models LLM-based Human Simulations Have Not Yet Been Reliable

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T22:00:26.703406Z digest=sha256:188cf5da3544d56a3171f41843ca493eb7a5b2e1063c52c764390acc3be6cd40

Observation e533bfd0-f4fc-4547-acef-2987531798a9 · inbound

Beyond Averages: Evaluating LLMs on Human Survey Replication at the Distributional Level cites this paper.

Beyond Averages: Evaluating LLMs on Human Survey Replication at the Distributional Level LLM-based Human Simulations Have Not Yet Been Reliable

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T17:03:42.094786Z digest=sha256:7d7b74375e66eae9ae9ddfd4a609a40c39ed380af6b8cea71f105d6eb240fb5d

Observation 1fab3eba-bc2e-488e-8867-a6560b25ca5e · inbound

Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework cites this paper.

Prompt Optimization for User Simulation in Conversational Recommender Systems: A Multi-Objective Framework LLM-based Human Simulations Have Not Yet Been Reliable

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-07T03:18:07.787194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-02T23:47:05.750545Z digest=sha256:43f27f0c4a5dc8d213142a66cca24b157c5d527148ea5aea9552901049ef8930