Pith. sign in

Paper Citation Record · LEDGER

Evaluating Psychological Safety of Large Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2212.10529.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.10529 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T15:28:48.251213Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:23.469787Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b8fbd43d-e62a-4ba1-8cff-6e0ee3dce1c4 · inbound

How Personality Traits Shape LLM Risk-Taking Behaviour cites this paper.

How Personality Traits Shape LLM Risk-Taking Behaviour Evaluating Psychological Safety of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T15:28:48.251213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T15:28:48.251213Z digest=sha256:70b7c5c11fc2df84432437f5dff4b52453c165dc1efffce73cd7c0adfd102e1c

Observation 9aab29ff-4ee9-4cea-9871-0ec4349a9d2f · inbound

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead cites this paper.

Position: Stop Evaluating AI with Human Tests, Develop Principled, AI-specific Tests instead Evaluating Psychological Safety of Large Language Models

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-19T02:12:55.706406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-19T02:12:48.586913Z digest=sha256:ddb13f5fb9967bb4e03baa10bbed436b7ba6671d940366b5088b41e2b0441abc

Observation 371552ed-933a-4c32-b875-302ad814d0d1 · inbound

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks cites this paper.

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks Evaluating Psychological Safety of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T09:40:45.318166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:40:45.318166Z digest=sha256:07a8cfbc463cb2574c2b784cc81e48e69972b331d1077d60e843f69df7130e28

Observation 2f72381b-4ad0-4d50-bc19-bae8f35222a4 · inbound

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes cites this paper.

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes Evaluating Psychological Safety of Large Language Models

Reference 73

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T01:09:20.198368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T01:08:50.547385Z digest=sha256:6dc160f97ca59c6407ac574f541e64bf4f9354b8141973f9e3e7dc5c8c9e88c3

Observation 6deb3154-bb80-48fa-b065-50dcc8357971 · inbound

Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback cites this paper.

Food Noise & False Safety: A Systematic Evaluation of How LLMs Fail to Adapt to Eating Disorder Queries with Clinician Feedback Evaluating Psychological Safety of Large Language Models

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:23.471300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T14:09:37.523725Z digest=sha256:f41dbd06bc88d11e18c18e4ca64650048753faca959393f82154306e707496cf