Pith. sign in

Paper Citation Record · LEDGER

Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2407.13943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.13943 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:58:39.391646Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:49:18.497365Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a0595463-d9c1-42c6-8e6d-42b7648a6a6f · inbound

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization cites this paper.

Learning Strategic Language Agents in the Werewolf Game with Iterative Latent Space Policy Optimization Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T21:58:39.391646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:58:39.391646Z digest=sha256:55ff62a64571c9829b810c30320b9ec57f7a18dcaf671bbeb82e05c13b9a996c

Observation b354bf91-0a7a-4a52-beff-d7b843a97177 · inbound

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers cites this paper.

Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T22:50:32.949249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T22:50:32.949249Z digest=sha256:c3ba519edec5125bca4ae67dbf3c4680d167e67a696b6b4b216c3e42e42c1916

Observation cc9f7743-f87c-4dd4-92b1-1cfc14760251 · inbound

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets cites this paper.

Agents Require Metacognitive and Strategic Reasoning to Succeed in the Coming Labor Markets Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:02:22.412122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:02:22.412122Z digest=sha256:7b3becdf1c49c49baa3146868b025feb854f852dc763bfe46b8821d4d5fb063c

Observation 6df2bdba-089d-4aee-bafa-a3f3e4da6171 · inbound

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models cites this paper.

SocialMaze: A Benchmark for Evaluating Social Reasoning in Large Language Models Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:45:15.845155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:45:15.845155Z digest=sha256:1d4a9ba8ed6f28c7b8aa940e46eb9062247b88d54a69a375b61da86400226366

Observation 724b08bb-f7b8-44c3-a214-03db4ff78204 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:57.923142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:57.923142Z digest=sha256:6613413720c786586251fec3f7d7b081e9e89146ac533b4ff37673f5afb7225b

Observation 3be38aa0-8869-492d-812c-2af4af6b75b7 · inbound

Strategy Adaptation in Large Language Model Werewolf Agents cites this paper.

Strategy Adaptation in Large Language Model Werewolf Agents Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:43.264423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:42:43.264423Z digest=sha256:b5bc6fd20b13f5baec7d88234a92ffb4a53411dc03c17a020422e7dd9cf8ee93

Observation 7eeaeb2a-9457-43d6-85e2-6642c708532a · inbound

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia cites this paper.

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:26:24.820227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T13:25:17.313704Z digest=sha256:8d74f69135725227cc0038a5b3301741f93ac8da7deff994a5f9d1558e458b29

Observation 14bb5245-6fcd-4294-b7f8-0e5807abde3d · inbound

AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum cites this paper.

AIT Academy: Cultivating the Complete Agent with a Confucian Three-Domain Curriculum Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:43:49.911579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T05:08:54.560648Z digest=sha256:a0306cd68ebccb22b2c053276eeb0562a836e9b402d2e4f3a93ab69ff6d6a49b

Observation 41bb56e2-2d5c-40f1-b85f-5f907eb5614c · inbound

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse cites this paper.

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:26:19.123141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:25:24.844859Z digest=sha256:66be016fffb2fd8b64bfda91b7d074ab2a0aeab379f7d4e4f4f5a9dc6eedc8eb

Observation d4372387-bfcd-4f7f-a16a-7ad8176b37e7 · inbound

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse cites this paper.

Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:47:26.904795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T06:44:28.552513Z digest=sha256:395c9da8e62ce6944badc346bdd678ea7de50265915468769207eb9f39a9e78d

Observation 9d9b41be-bcdb-4364-a6c3-024f314de6cf · inbound

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs cites this paper.

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:23:13.222505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:15:27.939886Z digest=sha256:8f46a9fb5f22b40c506352c4a837f2ea2703a75251d94cc80f03ad6543c2045b

Observation 45e7a251-a33f-4ef9-82d8-9dec8bf25823 · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.728065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:49fd70942b96694a37c4b1d18bed43fbaffaac22ef8a9f34a6b7eb13467968b5

Observation 62b9d302-6e6f-4bc0-a0b9-8a6ae95a92a3 · inbound

Enhancing Decision-Making with Large Language Models through Multi-Agent Fictitious Play cites this paper.

Enhancing Decision-Making with Large Language Models through Multi-Agent Fictitious Play Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:49:18.499911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T20:57:49.840546Z digest=sha256:8271030cb801354ac1c9901b1e773e389d82b5623b8d663989aad3e30e43c80a

Observation 4c68c23d-7cae-4249-944a-47e68d77ebdc · inbound

Thinking Out Loud: Real-Time Deception Monitoring in Asymmetric LLM Negotiations cites this paper.

Thinking Out Loud: Real-Time Deception Monitoring in Asymmetric LLM Negotiations Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:45:35.459536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-01T07:11:25.568064Z digest=sha256:5707a19807e402b09c62e50c3ad508ce67d625cee021ec9aaf125f35ddacfe68

Observation 5b6be818-7b78-48d2-aa8f-de906eedda25 · inbound

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action cites this paper.

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:15:44.661711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T05:40:54.002702Z digest=sha256:ef118c65d4422501f0141b37976beea16def560beeeb46fd6169ac237897e913

Observation fd43ebee-2b79-4cf7-a707-d86854e5e49e · inbound

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games cites this paper.

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T10:15:59.479435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T10:15:59.479435Z digest=sha256:3e775e321c6e062d760e32620a3bbf0613a0d4faa8b997aa30411760bff79466

Observation 3b95b283-e15e-4d31-aedd-e5b1279b29c4 · inbound

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games cites this paper.

MafiaScope: Non-Invasive, Time-Resolved Belief Probing for LLM Agents in Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T07:14:59.839689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:14:59.839689Z digest=sha256:88ff1f9c62bcb61ca13e35050409402648bca2bd3a56e58e763cdf768fd0ad85

Observation d3fc3623-1d5b-4819-9f6d-f4875f209f35 · inbound

Auditing Belief-Conditioned LLM Agents in Hidden-Information Social Deduction Games cites this paper.

Auditing Belief-Conditioned LLM Agents in Hidden-Information Social Deduction Games Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T09:04:16.361753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T09:04:16.361753Z digest=sha256:433ffdda0f858f9913ec9f423626aaabb4341f8f8392e7855ff20fdff27664c4

Observation bcfcad8a-1b0e-47b7-9701-120e26b7325d · inbound

Cumulative suspicion and absorption dynamics in an agent-based Mafia game cites this paper.

Cumulative suspicion and absorption dynamics in an agent-based Mafia game Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-31T08:47:10.407663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:47:10.407663Z digest=sha256:f9e635902ef468714272e93a0b85434a3dbae8c7cd50dddc9ee054f4a367e216

Observation 4d19343b-bada-4ef0-933b-0773c7bfddbb · inbound

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems cites this paper.

Even More Deception: Objective Misalignment in Mixed-Motive LLM Multi-Agent Systems Werewolf Arena: A Case Study in LLM Evaluation via Social Deduction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T00:51:15.781283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:51:15.781283Z digest=sha256:988d545586e41d1889ff4dc4b5532b71c1ad8717ce1d7db69704cdc12d143661