Pith. sign in

Paper Citation Record · LEDGER

Evaluating Large Language Models in a Complex Hidden Role Game

As of 23 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2605.22826.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.22826 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

21 of 21 outbound references displayed

  • verified exact10
  • verified fuzzy2
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 72d80937-1187-4ed7-aa32-4c0974977f1d · outbound

This paper cites Sijing Chen, Lu Xiao, and Jin Mao.

Evaluating Large Language Models in a Complex Hidden Role Game Sijing Chen, Lu Xiao, and Jin Mao

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T00:40:08.224939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:270ac3375e54988db1da2b5cc132b19449a5a6500b05829f82c073750b6031c7

Observation da9a2541-7c66-45e8-9ccb-b86f6bfde381 · outbound

This paper cites DeepSeek-AI, Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, Xiaokang Zhang, Xingkai Yu, Yu Wu, Z.

Evaluating Large Language Models in a Complex Hidden Role Game DeepSeek-AI, Daya Guo, Dejian Yang, Haowei Zhang, Junxiao Song, Ruoyu Zhang, Runxin Xu, Qihao Zhu, Shirong Ma, Peiyi Wang, Xiao Bi, Xiaokang Zhang, Xingkai Yu, Yu Wu, Z

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:40:08.301062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:94ba2d5d6144227f935126725c76ae31ca6acfd52febc3767f2aac5c284edbe9

Observation 41136e50-8ef7-4705-b5e2-7866f1f66f97 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Evaluating Large Language Models in a Complex Hidden Role Game DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:40:08.314443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:f9daba3fe1662432a0570dad22e42f8be4f12b433346e3f8edd1d352d52045f9

Observation bae41363-7fbe-4c21-8ee5-e782791ccef3 · outbound

This paper cites doi: 10.18653/v1/2024.naacl-long.123.

Evaluating Large Language Models in a Complex Hidden Role Game doi: 10.18653/v1/2024.naacl-long.123

Reference 4

Resolution
verified exact
doi, observed 2026-05-25T00:40:08.209659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:1b0133761f5ad6b5a054ecc3ccda3a9405ce429687c718d3c9f397f02537ece0

Observation e0691d35-e69f-4153-aba2-9e76a5f49b39 · outbound

This paper cites Gemma 3 Technical Report.

Evaluating Large Language Models in a Complex Hidden Role Game Gemma 3 Technical Report

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T00:40:08.304422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:91b8cb78a8380c3a0f30e95d4e14dfe16800e990a38ca356875af381ff4525f2

Observation 560a8979-be98-4f9c-8683-0d35ed48e42d · outbound

This paper cites Suspicion-Agent: Playing Imperfect Information Games with Theory of Mind Aware GPT-4.

Evaluating Large Language Models in a Complex Hidden Role Game Suspicion-Agent: Playing Imperfect Information Games with Theory of Mind Aware GPT-4

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:40:08.307638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:8164a41cc48890f7a57085d751df174651048432514fd73d440cced70d2af91f

Observation 32c360fa-242e-4d84-bbfa-75de82dcd2c2 · outbound

This paper cites doi: 10.63562/2577-8439.1111.

Evaluating Large Language Models in a Complex Hidden Role Game doi: 10.63562/2577-8439.1111

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:40:08.228896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:79138738c737dc555a50fbd116ad6fd19b38dbbe30516f21fbdbee56c4d3b921

Observation 8367cda9-08a6-4c1d-a296-cae08531dd8b · outbound

This paper cites 37 Wenyue Hua, Lizhou Fan, Lingyao Li, Kai Mei, Jianchao Ji, Yingqiang Ge, Libby Hemphill, and Yongfeng Zhang.

Evaluating Large Language Models in a Complex Hidden Role Game 37 Wenyue Hua, Lizhou Fan, Lingyao Li, Kai Mei, Jianchao Ji, Yingqiang Ge, Libby Hemphill, and Yongfeng Zhang

Reference 8

Resolution
verified exact
doi, observed 2026-05-25T00:40:08.232126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:b76a59ab2bd83bb83737ab4f9de2aa4b97093a0db58fefdf20124be2f6ca0616

Observation e791bbf4-78ef-4352-a83d-62345b5cff33 · outbound

This paper cites Evaluating Large Language Models in Theory of Mind Tasks.

Evaluating Large Language Models in a Complex Hidden Role Game Evaluating Large Language Models in Theory of Mind Tasks

Reference 9

Resolution
metadata mismatch
doi, observed 2026-05-25T00:40:08.218573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:a3d722f2b7620cb22fd30485f126bb09091791ccf0f72cb442fc9c54559473b3

Observation fb4405f5-6503-4642-b681-5bca6de955b4 · outbound

This paper cites doi: 10.18653/v1/2024.emnlp-main.383.

Evaluating Large Language Models in a Complex Hidden Role Game doi: 10.18653/v1/2024.emnlp-main.383

Reference 10

Resolution
verified exact
doi, observed 2026-05-25T00:40:08.215174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:7ec8f6f5b1df7954cf507a4950602df34f86bd36fc4d44203c449ca33a51d792

Observation 8ff21bf5-9fe8-4e84-a088-3939da3caa96 · outbound

This paper cites Bayesian Social Deduction with Graph-Informed Language Models.

Evaluating Large Language Models in a Complex Hidden Role Game Bayesian Social Deduction with Graph-Informed Language Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-25T00:40:08.205958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:bfb3f5eec38b3ab75d11a96676354a1b58528dbcb99cfb31bd7d2a5b5d267d2c

Observation 29977b90-9ae3-41b7-9277-1c3d81b460b4 · outbound

This paper cites doi: 10.18653/v1/2024.aiwolfdial-1.6.

Evaluating Large Language Models in a Complex Hidden Role Game doi: 10.18653/v1/2024.aiwolfdial-1.6

Reference 12

Resolution
verified exact
doi, observed 2026-05-25T00:40:08.234538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:77ab0cc4755539af9af77afaf4366f3bfcf82b6b3281b2025b6d19cde48bbd3c

Observation e5617ba6-5a93-41eb-a708-be714022585c · outbound

This paper cites Shenzhi Wang, Chang Liu, Zilong Zheng, Siyuan Qi, Shuo Chen, Qisen Yang, Andrew Zhao, Chaofei Wang, Shiji Song, and Gao Huang.

Evaluating Large Language Models in a Complex Hidden Role Game Shenzhi Wang, Chang Liu, Zilong Zheng, Siyuan Qi, Shuo Chen, Qisen Yang, Andrew Zhao, Chaofei Wang, Shiji Song, and Gao Huang

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-25T00:40:08.238813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:8d3a527b1deaab3de31a6db2e74f8e633a956ad8a861a80c34dbdc625547a1ae

Observation 1bc64e0d-2fe5-4135-94cb-9d30cfedda90 · outbound

This paper cites Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf.

Evaluating Large Language Models in a Complex Hidden Role Game Exploring Large Language Models for Communication Games: An Empirical Study on Werewolf

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-25T00:40:08.311063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:be2e68472d025aefec7bdc46a6f5faa4165c0fbfbf68fa182c09966225eca6ca

Observation c6670264-8310-4436-ba75-f37c1d935c8f · outbound

This paper cites URLhttps://openreview.net/pdf?id=WE_vluYUL-X.

Evaluating Large Language Models in a Complex Hidden Role Game URLhttps://openreview.net/pdf?id=WE_vluYUL-X

Reference 15

Resolution
verified exact
doi, observed 2026-05-25T00:40:08.221008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:58ba5654fc621c8be2354a7ff1e723bd640f79cdac41f61d140ae10b87ae3721

Observation f88d43bd-6521-4825-bbed-b3961bd3ce87 · outbound

This paper cites an unresolved cited work.

Evaluating Large Language Models in a Complex Hidden Role Game Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:40:09.440282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:21a83ecd06667b3bc8db0fe9f3890d335169e648036099b3e54082ae674b552e

Observation a218551c-1a47-402e-a388-e1a5f5b17768 · outbound

This paper cites an unresolved cited work.

Evaluating Large Language Models in a Complex Hidden Role Game Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:40:09.432730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:0add89119af2d9520eeb11b3459deafbc36bd50682ba13a459277d8f7a3423e1

Observation e3569e28-9c2e-42af-aa04-30e45e4de6f9 · outbound

This paper cites an unresolved cited work.

Evaluating Large Language Models in a Complex Hidden Role Game Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:40:09.428926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:14acf4b7e1e4775452628453209ab79a1d3bb90524bf185718e786f078eb1eb3

Observation 95948deb-6288-4833-b711-7fdd40c579b4 · outbound

This paper cites This misidentification creates substantial election risk, overwhelming the liberal president’s investigative advantage and producing a moderately fascist-favored score.

Evaluating Large Language Models in a Complex Hidden Role Game This misidentification creates substantial election risk, overwhelming the liberal president’s investigative advantage and producing a moderately fascist-favored score

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:40:09.435508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:71adad5b0971e69d5a0ec8c776dc1f52843012c6504bb7c8c181774b1a1fa1d8

Observation 1f7e6e03-e846-4c3a-b7f0-fa1e8abcc403 · outbound

This paper cites an unresolved cited work.

Evaluating Large Language Models in a Complex Hidden Role Game Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-25T00:40:09.437791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:29e990398500d6c367035c45216e4c64f608b0b6b24a84a13bbbf36ab16b05e1

Observation d05096e4-4a00-48f8-be95-c2f1232b39f1 · outbound

This paper cites Secret Hitler.

Evaluating Large Language Models in a Complex Hidden Role Game Secret Hitler

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-25T00:40:09.442632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-25T00:36:49.353675Z digest=sha256:8c032b8b49ff05e062e940f5e31dce99c4ea0ccfb195a858a94d9a1fe61cb961

Pith citing papers

No inbound Pith citation observations are available.