Pith. sign in

Paper Citation Record · LEDGER

GameArena: Evaluating LLM Reasoning through Live Computer Games

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2412.06394.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.06394 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:25.847549Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:58:02.971335Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 30692c77-ddd6-4436-9a51-59b1f45ce31e · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.847549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.847549Z digest=sha256:44f0257f23de464341c84c6ab98222fb76bb7ce5a4b2fc9149ce53f3adac48cc

Observation ff01265b-d150-4cf4-bf94-37cd01ea52a9 · inbound

lmgame-Bench: How Good are LLMs at Playing Games? cites this paper.

lmgame-Bench: How Good are LLMs at Playing Games? GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:01.953231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:01.953231Z digest=sha256:09fe10089436c44d859ebcb5ccbe5462fe963a61c96c67ff70fee02ef07c70c8

Observation 27cf108d-20a3-4b6a-a04c-3ded62521a33 · inbound

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games cites this paper.

Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:02:16.702374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T12:01:42.681135Z digest=sha256:d8ba4606c79952c022be24c38b56713c980c462507102713e29c466132e79c70

Observation 41b2cf9d-0447-4e83-aebc-1af6ac4e8fc9 · inbound

TextAtari: 100K Frames Game Playing with Language Agents cites this paper.

TextAtari: 100K Frames Game Playing with Language Agents GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:57.969662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:57.969662Z digest=sha256:54c6ad32979f504b095e4a5642a597ea9933b9a97500a1eebdc94d27326c6e98

Observation 1a9a032c-c17a-4a23-b0ab-a62459570381 · inbound

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey cites this paper.

Evolutionary Perspectives on the Evaluation of LLM-Based AI Agents: A Comprehensive Survey GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T06:00:19.542550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T06:00:19.542550Z digest=sha256:7d69407c2f09f769378664c4e14e1934d4fad900c5eff94ae7279b3ebba33b2b

Observation 1e557b64-bf58-4d69-8b69-b76e02d04dfa · inbound

Tracing LLM Reasoning Processes with Strategic Games: A Framework for Planning, Revision, and Resource-Constrained Decision Making cites this paper.

Tracing LLM Reasoning Processes with Strategic Games: A Framework for Planning, Revision, and Resource-Constrained Decision Making GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:46.240151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:06:46.240151Z digest=sha256:19288257400856e8fddb1b7f414c392d6b5c28961239256ffcc7c742a3e1aa74

Observation 066f1e33-b6fb-4c4c-b443-d864f667aa73 · inbound

Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play cites this paper.

Game Reasoning Arena: A Framework and Benchmark for Assessing Reasoning Capabilities of Large Language Models via Game Play GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T04:33:42.672656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:33:42.672656Z digest=sha256:c75ddb5a175d6f667ac5fe87d26d7bf9d0bbcaa37dddd80d2d05aeb5af19a604

Observation 25b4631d-4520-452d-81a7-219ff835e065 · inbound

DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations cites this paper.

DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:14:40.616201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T13:12:46.927103Z digest=sha256:28f48327ae1cf09cbb95f060a288795e5b2872f976dfc03fd9d00bd699b5c532

Observation 079c2a50-c0e9-432a-a59c-719b68e5b4d3 · inbound

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics cites this paper.

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:07:29.908843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T16:50:36.194650Z digest=sha256:f43e4bc7da0c2b2de9b35f5272f3946f34a0d850b54d3ea352f38836f87faa23

Observation 089bab7c-206c-45f2-9a58-7b2b58e77d5d · inbound

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application cites this paper.

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application GameArena: Evaluating LLM Reasoning through Live Computer Games

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:58:02.972737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T09:46:30.702256Z digest=sha256:d69e6c56e190959d1b5fd6133a0545b870ed03c8a3dc21093829c3761c0eb1fa