Pith. sign in

Paper Citation Record · LEDGER

FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2310.15421.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.15421 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:45:18.777761Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T23:17:29.774947Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 376a9404-d10c-4302-84ff-918afe3a131b · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:18.777761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:18.777761Z digest=sha256:89fe9830a4ab6cb07baa73460a184db3f3d3e6ecc46becaf26390373b2b1cf3b

Observation f6ea748b-9302-42d0-b5ba-23bdd750f2f9 · inbound

Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models cites this paper.

Beyond Context to Cognitive Appraisal: Emotion Reasoning as a Theory of Mind Benchmark for Large Language Models FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:45.594867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:10:45.594867Z digest=sha256:ba8670fb8cbb572dd9f3f1a6705e7e0d535c52cbb55baeb484afed254ac3fe30

Observation e5f44678-fdfe-4b2f-8e5a-a448a1081d57 · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.634383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.634383Z digest=sha256:b66a0248d04bc4ea499b8b1d4aaa69bf88542a73b08c8118385833573adf6351

Observation 4e089954-2179-40a6-b911-2530b8d02f70 · inbound

Intentionally Unintentional: GenAI Exceptionalism and the First Amendment cites this paper.

Intentionally Unintentional: GenAI Exceptionalism and the First Amendment FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:15.706574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:15.706574Z digest=sha256:37c38d307545ffd2bde3c060913f299727a8719f07d5d989330a420a6e7ac07a

Observation 82b6199b-6fc3-4418-85b2-4319bb337cba · inbound

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind cites this paper.

The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-06T22:49:38.805574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:49:38.805574Z digest=sha256:5bd87ec9fcc738eefeeee4e4469a15a6c15604e1b9d463ab9d77e6bcab35a4f1

Observation 5e06b11d-0b66-4ad5-951d-9eef5f8694cc · inbound

Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language cites this paper.

Mind the Motions: Benchmarking Theory-of-Mind in Everyday Body Language FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:04:17.452832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T18:04:07.256885Z digest=sha256:f7c973b68a3c9bb26c4807e3f9357a2c0fb2bdc28c1b35f8870b19f39a930acb

Observation ee95f333-29cd-410b-827a-547663070bde · inbound

Shadow-Loom: Causal Reasoning over Graphical World Models of Narratives cites this paper.

Shadow-Loom: Causal Reasoning over Graphical World Models of Narratives FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T06:15:38.644850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T18:41:38.814086Z digest=sha256:0a40d316d14ec86c310f14954b92d858c76465397b9c8bf15fbb722bf81e0ea5

Observation b8612cc8-e1c1-4027-866c-0dcae289a8ec · inbound

Reinforcing Human Behavior Simulation via Verbal Feedback cites this paper.

Reinforcing Human Behavior Simulation via Verbal Feedback FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:24:02.439889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-21T07:21:48.649289Z digest=sha256:d3ec3df2435249c6aaa82fe087ba990245b0490fe222deb91a75b9ab651db811

Observation 3eec17da-3b16-4d4e-8181-5fb696eb3057 · inbound

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting cites this paper.

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:17:29.776514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T18:19:06.340769Z digest=sha256:68126ff229a3f11e3a1683d1325cadb3450c9275754f25f6e568405646a3725b

Observation ea89d470-3d85-432e-9ded-37d101a518d0 · inbound

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting cites this paper.

PerspectiveGap: A Benchmark for Multi-Agent Orchestration Prompting FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T18:11:45.966902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T18:11:45.966902Z digest=sha256:c0726cf245b3ba8738b41abc8bef2996641ef515ee499d633d9a5b054832dda5

Observation 088b21c5-b056-4d5f-bb8b-1219dc7f646d · inbound

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs cites this paper.

Triadic Werewolf: A Jester Role for Multi-Hop Theory of Mind in LLMs FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:13:53.397462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:48:58.181882Z digest=sha256:812342b2a367c462101ecec769802e83630ca52689b462e1dce97a00f1a154bd

Observation c8fcd3cf-db62-4f3a-a3ed-eeb48e03be77 · inbound

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex cites this paper.

Training with (Swap) Regret Loss in a Single-Layer Self-Attention Model: A Case Study on the Probability Simplex FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 71

Resolution
unresolved
no resolver link, observed 2026-07-31T23:51:55.026505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T23:51:55.026505Z digest=sha256:8cbfc18569f469845f502e6432801ab680faa80d51be61e77ea0f15a4bd82a50

Observation 8db2ee0f-190a-4eb2-85eb-4c69f7fed3cc · inbound

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning cites this paper.

Evaluating Theory of Mind in Reasoning Models: Robustness over Reasoning FANToM: A Benchmark for Stress-testing Machine Theory of Mind in Interactions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:56.664963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:57:56.664963Z digest=sha256:16a0eef773a6ac4c2583f758b0472db5c865ba126ed676d6d0d29a130b083783