Pith. sign in

Paper Citation Record · LEDGER

Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2502.06060.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06060 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:00:53.877395Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T07:32:09.425973Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6f4b43c0-30b2-4d6a-aa41-a68e9ba6cd59 · inbound

AI Agent Behavioral Science cites this paper.

AI Agent Behavioral Science Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-07T11:00:53.877395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:00:53.877395Z digest=sha256:b6fdb1395bffc577db9877e53a3778c6eb13e2019d1fc1dc47426ef7c02509ea

Observation 6ed57130-317f-43e1-add1-ca4c61bd1a67 · inbound

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models cites this paper.

Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:58.121890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:41:58.121890Z digest=sha256:f78a3c8dabf8d87cfa3daa39e922dec6ff65fcc91aa377db00739ec34eba03e5

Observation deb9a075-edd8-439d-a3ea-2abfb0e03a88 · inbound

Bayesian Social Deduction with Graph-Informed Language Models cites this paper.

Bayesian Social Deduction with Graph-Informed Language Models Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-19T07:32:09.427999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T07:27:20.027179Z digest=sha256:9781f8635dfed7f58347619a5fd9f5ed224741e4446c6e75a3857ccdead34fbe

Observation fb4baca6-46fd-4e13-ba7d-a435926ffbd3 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T09:50:49.345480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T09:47:47.051969Z digest=sha256:1b7ce11adb446e4215ddd2f5c255fe077fe199dc578476d33526d7913f0ec970

Observation 5d7dd24d-003d-4756-b280-7027d65ebcb3 · inbound

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic cites this paper.

Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic Training Language Models for Social Deduction with Multi-Agent Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T06:46:40.124769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:46:40.124769Z digest=sha256:892ca52de5b365bfd5b467d1ea0711198e259305f4c85b1246074092a6153480