Pith. sign in

Paper Citation Record · LEDGER

Measuring General Intelligence with Generated Games

As of 23 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 8 inbound Pith citation observations for arXiv:2505.07215.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07215 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:26:35.889537Z

measured 59 of 59 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:27.682101Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:57.651515Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy19
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 133e8599-d7df-4aa6-9141-146e2ffcc246 · outbound

This paper cites ZeroSumEval: An Extensible Framework For Scaling LLM Evaluation with Inter-Model Competition.

Measuring General Intelligence with Generated Games ZeroSumEval: An Extensible Framework For Scaling LLM Evaluation with Inter-Model Competition

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.489506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.489506Z digest=sha256:36f0c7e55307d92e2ed4652bdcb0d02a4070cc9c5527cef0278c43615c554413

Observation 82079a04-8e5d-43e5-9dfa-7af5d1442578 · outbound

This paper cites Claude 3.7 Sonnet, 2025.

Measuring General Intelligence with Generated Games Claude 3.7 Sonnet, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.506665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.506665Z digest=sha256:6f9e385ab96617a7f7af099b074b1428e106d308b9d2a563cb249dee7020f176

Observation 845ea991-ad7e-4f58-8af1-96187795ec15 · outbound

This paper cites Claude’s extended thinking.

Measuring General Intelligence with Generated Games Claude’s extended thinking

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.519112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.523508Z digest=sha256:d6023b894594c35cd97e92005ecad6c0007ce0cf967b305ee420755e381ae8f4

Observation 3b4bc024-e831-4352-a1de-19ccf2d0282c · outbound

This paper cites OpenAI Gym.

Measuring General Intelligence with Generated Games OpenAI Gym

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.537991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.537991Z digest=sha256:0a8b0048800ba2f5a54e9a8aac27f2a2de8386f3ce38a8d967b200f9b56cdd68

Observation f7b1ff31-f3b0-4d0b-8bbf-5d8e4540ec7f · outbound

This paper cites Superhuman AI for heads-up no-limit poker: Libratus beats top professionals.

Measuring General Intelligence with Generated Games Superhuman AI for heads-up no-limit poker: Libratus beats top professionals

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.548471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.548471Z digest=sha256:18d7f939184c2e9870fd157b96fe23f37e8d562a1a77cd21b897616b511bd262

Observation edcf14da-ba3d-49af-b08f-f8e24c9045ab · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with GPT-4, 2023.

Measuring General Intelligence with Generated Games Sparks of artificial general intelligence: Early experiments with GPT-4, 2023

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.504453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.566647Z digest=sha256:8ead966c4c46f4f6e4cbde517d6629d4528e3b32e88ec61fcc7e4e656ce297a1

Observation e8282534-51cd-4299-92c3-af854f325a23 · outbound

This paper cites Heuristic DENDRAL: A program for generating explanatory hypotheses.

Measuring General Intelligence with Generated Games Heuristic DENDRAL: A program for generating explanatory hypotheses

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.489938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.581270Z digest=sha256:9d1af06760150cca32e316e71ce584c139e90998e83206d9c09a40a6de80baba

Observation 3420c9a6-765f-485c-ad5e-1d5ab49d9a5e · outbound

This paper cites Deep Blue.

Measuring General Intelligence with Generated Games Deep Blue

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.598304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.598304Z digest=sha256:b9530565e39c25eae6c2831c29467d5c14d8b9dbef0351c5ddb4606f3a9a41e7

Observation 28334878-f7da-44d7-89f7-4c4bfda6c1dc · outbound

This paper cites Gonzalez, and Ion Stoica.

Measuring General Intelligence with Generated Games Gonzalez, and Ion Stoica

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.476525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.612932Z digest=sha256:4ea8689bd346467bd0b4837615d80110f070e26aafb25fc80f8bc04bbb9c0506

Observation 79e745c4-d5c0-440f-b691-7ed24f7d51df · outbound

This paper cites On the Measure of Intelligence.

Measuring General Intelligence with Generated Games On the Measure of Intelligence

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.628071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.628071Z digest=sha256:377e56939e3e32eeb9cb90c83cca73200fd35f6d33cfc3ea9b4703806093e454

Observation 2f4c514a-cebf-47b7-beab-ad45ea9b2ad2 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

Measuring General Intelligence with Generated Games GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.639697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.639697Z digest=sha256:1e07b65e4b4e7b17b7fae019aa20b1e14e1b5babf1cb942e7256abaceb94f1c5

Observation 34ba052c-0c8a-46fc-8fc9-db3779152ca3 · outbound

This paper cites Gemini 2.5 Pro.

Measuring General Intelligence with Generated Games Gemini 2.5 Pro

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.465379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.649380Z digest=sha256:2f7f513bd69c90253ca8cbb82648f59a443ecd4ca01a8a498cdd0a8f3ad1f053

Observation 69e182af-982f-45bd-9b81-7e57f494bf53 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Measuring General Intelligence with Generated Games DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.660993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.660993Z digest=sha256:81e2ff45915b67b77169f32a2dc0333bb53b6609ea0b5c9f34749edfd10d17a8

Observation 3db7325f-08a5-4276-9e4a-ec84cdb8e82a · outbound

This paper cites PAL: Program-aided language models.

Measuring General Intelligence with Generated Games PAL: Program-aided language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.452602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.684060Z digest=sha256:fcf8670504caf472e3080d077549da835b190b81dd53c60158d446e45a3258e8

Observation 47c64b66-d5ec-4df8-b129-a7afe683f40c · outbound

This paper cites Frames of mind: The theory of multiple intelligences.

Measuring General Intelligence with Generated Games Frames of mind: The theory of multiple intelligences

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.697079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.697079Z digest=sha256:e71897d7e7216fa698788c06c8c36c5d66c844dc8ee8607a501b217a4def0d76

Observation 8ebb7676-7722-4d7c-afc8-701bc237ef4c · outbound

This paper cites Artificial general intelligence, volume 2.

Measuring General Intelligence with Generated Games Artificial general intelligence, volume 2

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.432174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.708897Z digest=sha256:8baa399e4125d784455dc728f5e79d103fbc2cad013004ea0fa1c66476a8c672

Observation ddcf07fd-d745-4b67-b751-6659d7c321c4 · outbound

This paper cites Interactive Fiction Games: A Colossal Adventure.

Measuring General Intelligence with Generated Games Interactive Fiction Games: A Colossal Adventure

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.717693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.717693Z digest=sha256:9bf6460037f3b38b0183cb12c462de0b1f5614e90f743342150cb0739e5b6c5c

Observation 947e82d6-ebcf-45b3-ba18-9d5d936480f1 · outbound

This paper cites Deep Reinforcement Learning from Self-Play in Imperfect-Information Games.

Measuring General Intelligence with Generated Games Deep Reinforcement Learning from Self-Play in Imperfect-Information Games

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.728726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.728726Z digest=sha256:7896ce48f994fa9a7a300c67aecfdb6c4a04ee0ccc89caa2445347ad5b496a76

Observation 61aa489d-b5f7-476b-9e6b-cb5c8af5250c · outbound

This paper cites Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan.

Measuring General Intelligence with Generated Games Jimenez, John Yang, Alexander Wettig, Shunyu Yao, Kexin Pei, Ofir Press, and Karthik Narasimhan

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.740114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.740114Z digest=sha256:5754b6c21354100bc038b3dd689dfcdbbf9b23fdaee159ad9b8e4a87dcbaaf81

Observation c84e510d-e256-4345-944e-79a2541e376c · outbound

This paper cites Dynabench: Rethinking benchmarking in NLP.

Measuring General Intelligence with Generated Games Dynabench: Rethinking benchmarking in NLP

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.757320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.757320Z digest=sha256:a6c82e8c90179669675bf9ce120db96b661b4dc6a417dc6ccd6dac92c64c7cd2

Observation 3885d6f0-5714-4022-81bd-711b16c4d1b1 · outbound

This paper cites A collection of definitions of intelligence.

Measuring General Intelligence with Generated Games A collection of definitions of intelligence

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.406377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.768293Z digest=sha256:898071c824a23eccf6a7fcd868c5b48eb1e0fa709cd2dd01dfe4ef0a3bbe1166

Observation 51f296a6-74eb-4af6-bb61-4541346caba1 · outbound

This paper cites Guha, Karen Pittman, Dexter Pratt, and Mary Shepherd.

Measuring General Intelligence with Generated Games Guha, Karen Pittman, Dexter Pratt, and Mary Shepherd

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.394631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.776253Z digest=sha256:981b8f064805d5d71940bf359b6a26657ff837d0391cb4de4a376842f335d04f

Observation 96887fc7-ed0b-4a72-a771-7ced05b51b5f · outbound

This paper cites Discovering and exploring cases of educational source code plagiarism with Dolos.

Measuring General Intelligence with Generated Games Discovering and exploring cases of educational source code plagiarism with Dolos

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.786483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.786483Z digest=sha256:c5f32db05739f7da53b3467912d2a87da4d5193052e5bc9143e14dd5721bf131

Observation 98cfb84d-6afe-41fd-8412-34a3a79e79ac · outbound

This paper cites A proposal for the Dartmouth summer research project on artificial intelligence.

Measuring General Intelligence with Generated Games A proposal for the Dartmouth summer research project on artificial intelligence

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.382760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.794744Z digest=sha256:feb2964be8ade472a01e4ed08a4e3c4012f02af0073a30acc79123cd399cea5b

Observation b88f4be3-1701-4b9d-b8d5-b9bc9f6e060c · outbound

This paper cites Min, Yangruibo Ding, Luca Buratti, Saurabh Pujar, Gail Kaiser, Suman Jana, and Baishakhi Ray.

Measuring General Intelligence with Generated Games Min, Yangruibo Ding, Luca Buratti, Saurabh Pujar, Gail Kaiser, Suman Jana, and Baishakhi Ray

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.370710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.800822Z digest=sha256:fc3eecb99bc27e4fc56b79922dfc925af2497c7eae980121132e28bb0b14381a

Observation ed9248ad-cf24-4055-a9a7-488dbfdc86cf · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Measuring General Intelligence with Generated Games Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.812305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.812305Z digest=sha256:ca2af80e640ed628317dcb094466ae37e715aa1407390478822575ff9d691b9a

Observation 631c8233-6385-4b61-bfe8-f8f97f6fa1ed · outbound

This paper cites an unresolved cited work.

Measuring General Intelligence with Generated Games Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T22:26:36.358833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.805962Z digest=sha256:90e61712b749f5bb3a83116a53bd773a6d0c12536bf6c1d6ce14cd1af2b30620

Observation 98f28dba-20d3-4bf2-a6a0-a3344ee36131 · outbound

This paper cites OpenAI o3-mini, 2025.

Measuring General Intelligence with Generated Games OpenAI o3-mini, 2025

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.341564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.819273Z digest=sha256:4e57690ce19184f925d2ee2254d71cd6137cb5846011cb3eb0070d53f112482c

Observation da876027-ae3e-4578-ac19-4c36f74bad14 · outbound

This paper cites Introducing OpenAI o1, 2024.

Measuring General Intelligence with Generated Games Introducing OpenAI o1, 2024

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.815909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.815909Z digest=sha256:4bd1dd12c43862a3523c700849abec1d03a976f6f64906e9be0f11a20a7719ae

Observation 5cb7fdb3-c202-4ae9-8543-357bf9347835 · outbound

This paper cites Stable-Baselines3: Reliable reinforcement learning implementations.

Measuring General Intelligence with Generated Games Stable-Baselines3: Reliable reinforcement learning implementations

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.320828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.824886Z digest=sha256:377ee968c608b959cd81a3ca1be4135a9dfb989829ca03f628338cdbbc015fc2

Observation d20ee3b0-c84d-471e-bd14-975a4d86827c · outbound

This paper cites Introducing o3 and o4-mini.

Measuring General Intelligence with Generated Games Introducing o3 and o4-mini

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.331968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.821973Z digest=sha256:94c494f498cabf823c3688cf340bc1616ee2afeb8ccefb206d1ab9bfa8bfd860

Observation 2652d443-5376-4930-ad95-24aa8d52da50 · outbound

This paper cites Neural theory-of-mind? on the limits of social intelligence in large LMs.

Measuring General Intelligence with Generated Games Neural theory-of-mind? on the limits of social intelligence in large LMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.830303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.830303Z digest=sha256:eb306f076367d5bb94697126b38a3361b0b546a9f69cd36c9567f9a059e570c6

Observation e348cb2e-7f2a-4584-8af7-b488e2075061 · outbound

This paper cites Artificial intelligence: a modern approach.

Measuring General Intelligence with Generated Games Artificial intelligence: a modern approach

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.827657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.827657Z digest=sha256:75dc3c8f1808f129f117f1e3dab32089c7ea37c30afdf9161963d635de6356fa

Observation 8cc3e7c0-e5bd-4c65-9dd9-b3c2bbddea8d · outbound

This paper cites Winnowing: local algorithms for document fingerprinting.

Measuring General Intelligence with Generated Games Winnowing: local algorithms for document fingerprinting

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.292482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.841442Z digest=sha256:98a5d8aa0fbde053d9b881a70c02cbfec3f201aa87c48f8428e0c9a5727741a8

Observation a25b3ad0-77af-4059-80e4-9f345d87d6b2 · outbound

This paper cites Measuring intelligence through games,.

Measuring General Intelligence with Generated Games Measuring intelligence through games,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.302077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.834094Z digest=sha256:ec24713cf5e6024c6395eb06bfe857f8c2edf7a76b66247b1e82a23f69f1f344

Observation a849fd1f-40e6-4731-9d9d-db623daf4265 · outbound

This paper cites Reflexion: Language Agents with Verbal Reinforcement Learning.

Measuring General Intelligence with Generated Games Reflexion: Language Agents with Verbal Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.847681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.847681Z digest=sha256:af67d8a9e0b63a59bef2b485361ee2292945343558bff17ddc6a84db4759037d

Observation bea15f2c-9585-4482-9127-3e0e4bf17d35 · outbound

This paper cites Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm.

Measuring General Intelligence with Generated Games Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.851037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.851037Z digest=sha256:75c3cb04f7212f00ea124b04124b3f1a4d982b27539f321af17cd0f3901926d8

Observation 25176ad3-5ebc-4a9d-a1a7-1fb309d6bd72 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Measuring General Intelligence with Generated Games Proximal Policy Optimization Algorithms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.844596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.844596Z digest=sha256:32c03245ea5fa24ccd382ba3b53c82fb96a19c92798d961f976baf01d821bbcf

Observation 48ce2e6f-28ec-4a44-8a84-35d8b492ca1c · outbound

This paper cites What is intelligence?: Contemporary viewpoints on its nature and definition.

Measuring General Intelligence with Generated Games What is intelligence?: Contemporary viewpoints on its nature and definition

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.282733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.858755Z digest=sha256:a7b44a16c5f6e86f15aa338735f3547c40f9f910504bb32cbfc2d43e4ead03e0

Observation 79ecaa26-cecd-4a4e-9f23-e79ff1b2db3b · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Measuring General Intelligence with Generated Games Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.861284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.861284Z digest=sha256:d8c81cf24f63bd40ef1ba90816484c00aee2229e12ea90804e55783bbda1fb89

Observation 72bc769b-fbcc-4611-a983-30a97404d409 · outbound

This paper cites Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models.

Measuring General Intelligence with Generated Games Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.854776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.854776Z digest=sha256:81ccb38ff0eeb27714b920135deab20d9b5277f9b33ed4262d165521c5da29c4

Observation b4467a53-3f86-40ca-8de0-e64dcd7e4956 · outbound

This paper cites Goal-driven explainable clustering via language descriptions.

Measuring General Intelligence with Generated Games Goal-driven explainable clustering via language descriptions

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.871979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.871979Z digest=sha256:a0fac5c03ecb861e7781f10ed52d78fd0cf9fd79c11408aec65827dd551ca627

Observation 4f5e4639-fb1e-4ee1-9f4c-e3ff492a5716 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Measuring General Intelligence with Generated Games Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.875812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.875812Z digest=sha256:ce8c91a94807f5bfd7fe174848ca23c85ed080dab25029acef72ce16577ecf6a

Observation c9ebdd1b-ee3b-4df1-b5ce-c38f2778643c · outbound

This paper cites Evaluating large language models with grid-based game competitions: An extensible LLM benchmark and leaderboard,.

Measuring General Intelligence with Generated Games Evaluating large language models with grid-based game competitions: An extensible LLM benchmark and leaderboard,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.272810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.865418Z digest=sha256:3dc528ffd13d6e608c4e7122c3ecc8effba67e8a9b30df92a368282d11dcbe2d

Observation c198cd6a-ed09-4756-805f-507792166ad6 · outbound

This paper cites Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard.

Measuring General Intelligence with Generated Games Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.868797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.868797Z digest=sha256:8c221fd503da5994b917245f4f709380a8aac042c35d99f436cac3a336ffb515

Observation 27550da4-078f-4621-a11c-cd4768757aee · outbound

This paper cites VideoGameBench: Research preview.

Measuring General Intelligence with Generated Games VideoGameBench: Research preview

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:26:36.262560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T22:26:35.886314Z digest=sha256:e7a38873dcc4597aa81df81c6088205782c7591a9a61f11271a26cf51fcd11f5

Observation 45a4d9c4-12de-4af2-962a-1fbacc3f3087 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Measuring General Intelligence with Generated Games Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.889537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.889537Z digest=sha256:ce57298bf87929688e92f533b796e53ae91e850685e20fd1d0da6cf434ae6312

Observation fc0fea58-d58e-4ca3-8e0c-4935f5fc9427 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Measuring General Intelligence with Generated Games ReAct: Synergizing Reasoning and Acting in Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.879068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.879068Z digest=sha256:3b32348c9f2a885336e65492fb2c64094517abcfd0fbef5997b372c43ea6dff5

Observation ac653952-9d44-4123-b12c-066b25484ba9 · outbound

This paper cites $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains.

Measuring General Intelligence with Generated Games $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.882198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.882198Z digest=sha256:c5ab63dfa54721fef7fd8dbc0b39fa4a38cca898fe3e62ee4f691c63be4720ce

Observation 34f20f2d-1ecf-4ba9-8fa5-5f66ca68afd1 · outbound

This paper cites Measuring Intelligence through Games.

Measuring General Intelligence with Generated Games Measuring Intelligence through Games

Reference 2011

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.837771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.837771Z digest=sha256:0448c51f92b74be7757aa75fc45b3ceb0020435fd50fb2db0a290618c0f3765f

Observation bfa34b4b-eb16-4c96-a133-3bd34e367920 · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Measuring General Intelligence with Generated Games SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-15T22:26:35.747667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:26:35.747667Z digest=sha256:f11d6a8f15395f57c72e0bd574cde4416e299e6e0e90ce235939425ade226462

Pith citing papers

Observation b380370e-9acc-4f3e-a302-b99f85dadf03 · inbound

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation cites this paper.

KORGym: A Dynamic Game Platform for LLM Reasoning Evaluation Measuring General Intelligence with Generated Games

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.682101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.682101Z digest=sha256:4194f7c5a837f3c6f3808182a322546e1fc3582bf5d28d8a4fc3b9987d630091

Observation ebacd626-1aae-45ad-a2f4-29465a40b284 · inbound

Assessing Adaptive World Models in Machines with Novel Games cites this paper.

Assessing Adaptive World Models in Machines with Novel Games Measuring General Intelligence with Generated Games

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T16:41:19.917526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:41:19.917526Z digest=sha256:b361a3922febf571b2bfff147987dd34f95d9b012957eaf769137b7c6c51816e

Observation bae068f2-831c-4d1c-82fb-0e578cd9484d · inbound

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment cites this paper.

HEALing Entropy Collapse: Enhancing Exploration in Few-Shot RLVR via Hybrid-Domain Entropy Dynamics Alignment Measuring General Intelligence with Generated Games

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T05:25:55.254384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T05:23:08.478393Z digest=sha256:185cf6986ab703f98bb99db1c24e9a130ea756900ca9b317cd1ffc121c109a2d

Observation 00cc2af3-2f7d-4d80-b951-208dc667a35b · inbound

Scalable Environments Drive Generalizable Agents cites this paper.

Scalable Environments Drive Generalizable Agents Measuring General Intelligence with Generated Games

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:23:12.204325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T10:19:14.829125Z digest=sha256:a1f7a72802e152c2d20a23bb0a291a1cfe2fd016165e2c9eddc81df4c43cad84

Observation 40d95776-71ee-4c53-9e12-e1115671faa6 · inbound

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models cites this paper.

GENSTRAT: Toward a Science of Strategic Reasoning in Large Language Models Measuring General Intelligence with Generated Games

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:45:20.673786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-25T04:41:35.532363Z digest=sha256:5e059c4fbc907abf8c7e773dd1ffee3380aa7d665a49e5486565a1faac9919c5

Observation 0b83c40d-3bbc-4f5b-ade5-7efab9cdaee3 · inbound

Distilling Game Code World Model Generation into Lightweight Large Language Models cites this paper.

Distilling Game Code World Model Generation into Lightweight Large Language Models Measuring General Intelligence with Generated Games

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.687987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T13:58:37.956333Z digest=sha256:17fb26d1859417580d9b394f718cbe0ed1901149cbe404661e11b0475bafdd21

Observation 544c19b7-a014-44f3-b289-06bd07b418d8 · inbound

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games cites this paper.

Using Cognitive Models to Improve Language Model Simulation of Human Persuasion Games Measuring General Intelligence with Generated Games

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:57.653000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T01:03:49.101568Z digest=sha256:0709abf5c3bf166c437f9fceadfc7d024b2c9b17cda07962239590e1ebc8129d

Observation e4b33542-ef48-4d14-9e5a-ef52e875905c · inbound

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning cites this paper.

Spatial Reasoning in LLM Game Agents: Impact of Causal Context and Multi-Step Planning Measuring General Intelligence with Generated Games

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-01T10:58:16.945355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:58:16.945355Z digest=sha256:dd967d427c3d039cf62a48c5144045c63cdaacc8996145c19ef99f39ad641ea4