Pith. sign in

Paper Citation Record · LEDGER

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

As of 8 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2508.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11944 v1

Coverage vector

measured 38 of 38 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:48:57.786431Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:38:14.922388Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

38 of 38 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved31
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation e303be8b-d118-49da-83fd-debb922c0a42 · outbound

This paper cites GPT-4 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.816976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.816976Z digest=sha256:ae5441cd7c53f992b3623e62e2ad92282f2e0daed44b75ccf23959745be93ac1

Observation 59b3bbf1-b490-4d5c-bbf6-379836ff166d · outbound

This paper cites Akata, L.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Akata, L

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:02.677810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:53.901582Z digest=sha256:ccd7a274b0af53feac108bc5cef13d213666110edc50c35cd869a62e3797f379

Observation 8b04781d-4edc-44cf-ba33-23cc0ee101bf · outbound

This paper cites DeepSeek LLM: Scaling Open-Source Language Models with Longtermism.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:53.984633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:53.984633Z digest=sha256:095c9207d1923b8cf2e01011a47f33b701ea3b01d3af223e0adb0045f618ea20

Observation 8d10f084-86a3-4d5e-8038-75417201380e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.434647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.142406Z digest=sha256:91b33add4cbeb3a37e88c6811e31fb017e6408c39ab8593f96e2636bda93d255

Observation 257e9564-3cfa-4a29-ab2f-289dcd0a884d · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:02.177388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.291423Z digest=sha256:212f27e05c17f1e77d97000a5e3159576a06d169e3e0df4db2455d56ed8c0976

Observation a21a191a-6f5b-40c0-9fdb-3b3ae9a3084e · outbound

This paper cites Chong, T.-H.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Chong, T.-H

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.912672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.446528Z digest=sha256:d971c36cb7681e916ee0106fc1667e7efeb12224dcac233b99cf222bfe27f273

Observation 3843406b-2c53-4970-b085-fa8a20ecafa8 · outbound

This paper cites Costa-Gomes, V.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Costa-Gomes, V

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.645005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.541089Z digest=sha256:1ad2608921946f3350fa332942686654a91146c38499cfad9c1b9e79b3abdf59

Observation df416bba-30db-46f5-8714-5221c907d163 · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.647303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.647303Z digest=sha256:9f88da854be550accce6e0c64971e408cdb736f0965a664219cdde04f9151661

Observation 429390c3-b5d6-4da9-8863-ed5495f19bf6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.465721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:54.758623Z digest=sha256:a7da2fd47da99c0f0de0b59879666491eb4a42032dd113811c4d70da5fc2f772

Observation 909c6723-8d06-42bc-977a-fb92d03ebb61 · outbound

This paper cites A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Social Agents in Game-Theoretic Scenarios

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:54.905621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:54.905621Z digest=sha256:f14afd531a232e58ee3dc1dc262e32f9c74ee509943e56c0d2152081039b12f8

Observation ec70a49a-e9c7-4d09-b7fe-9998a73e605b · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:01.256598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.063515Z digest=sha256:0c53e2c77531a3c45888ccafabff5f4efe53d596740b324949fddbddeba405b5

Observation 1a714281-9d29-48eb-b7b4-27f5823c0775 · outbound

This paper cites Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.222817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.222817Z digest=sha256:13752054b93e131e75fd45d1aa86f53bf8aa10438a9521db5247071bb3e1ffa7

Observation c630cc8a-08cb-46c0-8133-c0ca8d957c88 · outbound

This paper cites Fudenberg and J.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Fudenberg and J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:01.066870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.350398Z digest=sha256:646c363abbc474195d81fc8ef85505eb100822dd5631018685224f633bdf74d1

Observation d3e8c5dc-09fd-49b6-a091-577ead17bf61 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.912683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.499041Z digest=sha256:d84da03a5545c468e2b7b0500e4ed9e0d0a88b74e0d48a2ab5781bfd4f63e162

Observation d96cafc8-7b40-4594-89d4-bff81bf4cb91 · outbound

This paper cites The Llama 3 Herd of Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs The Llama 3 Herd of Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:55.653643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:55.653643Z digest=sha256:787b5ef1c728e91535bf0083171a8ba1bf6a743d0cc37f1df771c92968534f88

Observation ad6c2434-3691-40f2-aa21-6e23926de8b9 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.680640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.758196Z digest=sha256:f64cb6bcee3b610567d8961524e704170eaf73d9677274936c2755edcb8ad411

Observation 7162f826-bfe5-47c2-bd8d-c66074baf6d7 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.510380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:55.881168Z digest=sha256:b72894837121cbebe8d8eed228148a4cb2575e2188554753a96255dfbe75efbe

Observation b88e6566-f4a1-4f4f-8ccc-a75650918512 · outbound

This paper cites A Survey on Large Language Model-Based Game Agents.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs A Survey on Large Language Model-Based Game Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.005817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.005817Z digest=sha256:718dbe4fd6ecdc2e21e8627122360bc90176510a0479df150816818734c5b442

Observation ce17b52e-199c-40a2-949f-6ca5081c00bc · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.126881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.126881Z digest=sha256:4ab7f56867082dcba4b8351c7d0c92052445925a9fe6bfd4824f2dfed477cd1a

Observation 22f9f327-c4fb-4439-8656-ffee46d16aae · outbound

This paper cites Lorè and B.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Lorè and B

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:49:00.280558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.219061Z digest=sha256:1bd8cf03a93b4e9f384e68f600cdecfb29f2ed3ce87f972724c8b7a8bd9d2f08

Observation 0bf3f676-7b40-4777-b545-a9222a18c758 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:49:00.112476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.340141Z digest=sha256:58c4f871b1f531fc4db301ea2e0d3fad00943644644f4b7d66a4218979d340a6

Observation 58cc6ebe-2cba-432f-9598-7dea43a8b1c6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.906769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.417387Z digest=sha256:a529023d6bd9d060fe2179c7b24225678bd9f68e3ae5befaba18790a772d64f6

Observation 99134207-b27a-4412-abc1-1b21383bf8f4 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.595239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.536855Z digest=sha256:7ad4cda463048e205e1a405ef6e8049afd6aec9e4f81f73b9bc85facbadbf46c

Observation 35fef22c-dd53-4fe4-8c62-2601bf32a03e · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.420270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.612362Z digest=sha256:661eee3238f215e251c86338c6f6ef96ebe60f3bf8de2f2c9d903571fbfa5a52

Observation c187cfa9-8834-46ac-a58a-cb99298d8be8 · outbound

This paper cites LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM economicus? Mapping the Behavioral Biases of LLMs via Utility Theory

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.677470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.677470Z digest=sha256:60d54821f07f864c6680975796001cf4dd8a54c64e6ccdc2075b61877e7e8bb1

Observation a87b0631-deb3-4b2c-900b-a4e21463f7e5 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.188161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.793368Z digest=sha256:583af01c98a98dbb1bd941572ff8751c1f0708057222c41c69cfbc82585b0ea7

Observation 934bacbe-ba0a-4994-ba64-9a68b3f426b6 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:59.014871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:56.897295Z digest=sha256:004ea7f2fc42b4a895ebba1728c3f6886eb7ed81156461c9dfa143eadbfbddfc

Observation 40691c42-96d4-4064-b748-14398244dee2 · outbound

This paper cites Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Game Theory Meets Large Language Models: A Systematic Survey with Taxonomy and New Frontiers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:56.972734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:56.972734Z digest=sha256:2489078b26841a6ecaa3c058cdee4b65e7deabe00e5e0f226975a321e38e7343

Observation d8542316-4165-4814-b143-c6a247efa951 · outbound

This paper cites Do Large Language Models Exhibit Spontaneous Rational Deception?.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Do Large Language Models Exhibit Spontaneous Rational Deception?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.044552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.044552Z digest=sha256:21a203fcbcb3807b810f134311c7d63b0a0077ba51c8a9e9ed8dfe81ec2191b2

Observation c4fe8893-7ef0-4a65-ab68-f1fff782b201 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.113858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.113858Z digest=sha256:9262d8bfa3f773f134be841d3930dccad2541ecc7864a744dbaa5887e01cea4d

Observation 8e62b07a-a745-4297-877c-b2f2fdde2ae1 · outbound

This paper cites V on Neumann and O.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs V on Neumann and O

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.793228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.188015Z digest=sha256:043b8ef91cd9c98b53ccf6a1352e0ee5384102d0a4f91e274e4a7b517863c827

Observation 466212e9-ea9d-434d-95d4-9dec953ffccc · outbound

This paper cites TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs TMGBench: A Systematic Game Benchmark for Evaluating Strategic Reasoning Abilities of LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.298899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.298899Z digest=sha256:d4ac92efe4628150c7f5fe117071be2ea70f02942774611c62f455837d8c603a

Observation e9dbb274-4c68-433e-9893-3fbd28f16a44 · outbound

This paper cites Wright and K.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Wright and K

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T19:48:58.622211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.369503Z digest=sha256:c02061dd08f7738c7a2df43308c57758a4895c47edb2aedd30c3bc8055318d81

Observation cb3c3c2f-d1cf-40ec-a07f-edb370ef1b75 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.467609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.452686Z digest=sha256:bf027a08f22d0bc0b83483a4d7aeb9ee5ec44311b4a2194a100a16dcb38a84a1

Observation 8ac9ef5e-177e-4c42-98a2-6a93a1c701e2 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.283296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.519879Z digest=sha256:4f45aa7ff462abbb3410ed294bed87776a2ab4be8d73308656f6fed3e91e5911

Observation 4d1f26ad-5048-4b03-afe6-e57a5495cd56 · outbound

This paper cites an unresolved cited work.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-05T19:48:58.074219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T19:48:57.618873Z digest=sha256:58b1c16b9ec0f55539d3b5421a413bb25e3481935bd0c30d2aca01a1cb1c7f2d

Observation 70ba9204-49ab-4d69-95b3-db41edff3ada · outbound

This paper cites Qwen2.5 Technical Report.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs Qwen2.5 Technical Report

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.703996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.703996Z digest=sha256:f763a6abc3b52ba73abac184bf32c4cfc02b520ce281d22764889154206429dd

Observation ad174fc3-992c-4687-b16a-087d837f79bb · outbound

This paper cites LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models.

CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T19:48:57.786431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:48:57.786431Z digest=sha256:45b3841b443c67611beb63b9195a4d250afa3ea7ab6a6e01ef9e53ac82d8f0b8

Pith citing papers

Observation f03bcd7d-52dc-47aa-be67-13b89306fced · inbound

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks cites this paper.

BehaviorBench: Benchmarking Foundation Models for Behavioral Science Tasks CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-06-26T00:38:42.728182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T00:38:14.922388Z digest=sha256:87c02bf412ca9e86423b48ae7cd4907d31d53ab5b2f1513f0d2f60574d004f51