Pith. sign in

Paper Citation Record · LEDGER

Evaluating Psychological Safety of Large Language Models

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2212.10529.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.10529 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:09:43.149479Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:23.469787Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8fbd43d-e62a-4ba1-8cff-6e0ee3dce1c4 · inbound

How Personality Traits Shape LLM Risk-Taking Behaviour cites this paper.

How Personality Traits Shape LLM Risk-Taking Behaviour Evaluating Psychological Safety of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T15:28:48.251213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:28:48.251213Z digest=sha256:c49a4ae826ffae67e74499fc6da1c1876822a680a09e5d0a4372bec08e18c38f

Observation 101f22f1-87c7-44a6-92a3-1a531dd27a80 · inbound

Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications cites this paper.

Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications Evaluating Psychological Safety of Large Language Models

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-16T05:09:43.149479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:09:43.149479Z digest=sha256:1c3d9ba2636c099f94298ca36e8b464c3dd36673e817af6be0f913b5295ac972

Observation 9aab29ff-4ee9-4cea-9871-0ec4349a9d2f · inbound

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead cites this paper.

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead Evaluating Psychological Safety of Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:12:55.706406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T02:12:48.586913Z digest=sha256:0f22786fab3942d9c7ef645a10205abc0ef1d6e30e5321d8452d1fad01d80ed7

Observation 371552ed-933a-4c32-b875-302ad814d0d1 · inbound

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks cites this paper.

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks Evaluating Psychological Safety of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T09:40:45.318166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:40:45.318166Z digest=sha256:1aa65ca983c11c48a0ba577ca3213db2384898bc929d7e249cef36950d6e775a

Observation 2f72381b-4ad0-4d50-bc19-bae8f35222a4 · inbound

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes cites this paper.

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes Evaluating Psychological Safety of Large Language Models

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T01:09:20.198368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T01:08:50.547385Z digest=sha256:a85a5f183da3dde4a8307ef4434171fa8891a099c7d5693b42b5ad087c901ce7

Observation 6deb3154-bb80-48fa-b065-50dcc8357971 · inbound

Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback cites this paper.

Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback Evaluating Psychological Safety of Large Language Models

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:23.471300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-28T14:09:37.523725Z digest=sha256:752c4932aece09ee490cee42e8f6e54c21762822a0d3d493f85ee7544bc98fe3