Pith. sign in

Paper Citation Record · LEDGER

TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2410.10479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.10479 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:48:57.298899Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T05:51:25.752382Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 466212e9-ea9d-434d-95d4-9dec953ffccc · inbound

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs cites this paper.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.298899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.298899Z digest=sha256:f7cc325ecb998677385407c3125332d16c5fa0bd1809ffdb7220a1333c33b2c4

Observation f4e8a649-6131-4a51-ba14-d55eaa5e4245 · inbound

CivBench: Progress-Based Evaluation for LLMs' Strategic Decision-Making in Civilization V cites this paper.

CivBench: Progress-Based Evaluation for LLMs' Strategic Decision-Making in Civilization V TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 37

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T05:41:01.839083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T17:59:21.887765Z digest=sha256:0f2cf9357c1b9d0327c5f17a674226b71ec99ceb9db794170d5f0853c45f6cc9

Observation 107868a7-4f30-4913-bd07-08382d99b5f5 · inbound

Explicit Trait Inference for Multi-Agent Coordination cites this paper.

Explicit Trait Inference for Multi-Agent Coordination TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:48:27.519095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-10T02:45:21.304116Z digest=sha256:ef056a5d5225f5cfe16592dd7f73fd957901726067c5da2d5cf92c0e5bb334d2

Observation d9c9e404-f9b6-4d1a-8684-78de21bace1e · inbound

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models cites this paper.

Equilibrium Residuals Expose Three Regimes of Matrix-Game Strategic Reasoning in Language Models TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:25.755923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:52:47.745976Z digest=sha256:f4a5165da10f7f32dc8df5e001ca19b8d3b58946bb8d43aa0c436bc537ae813b

Observation 9105d0ac-a0a5-4c97-b8ad-3682c64b8c55 · inbound

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game cites this paper.

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-07-31T16:28:48.696176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:28:48.696176Z digest=sha256:27603e185fa94774bfd86d453b8a428b0bbede8c1bed64e8d3e8e7b2fe24f630