Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:53:58.448516Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2502.07087.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:53:58.448516Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T08:40:40.910461Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T08:40:41.541945Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0d3d385b-6226-4e31-a870-161a0705d6b8 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1a12ddb-7f55-45ba-ad4f-73451fe84156 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Claude 3.5 Sonnet , 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d50af5b8-f711-45f8-98f6-893b9babe3ee · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring GPT-4 Can't Reason
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65fa7dd1-f9fb-4fb8-831e-9e143f8790c4 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring DeepSeek-R1 : Incentivizing reasoning capability in LLMs via reinforcement learning, 2025 a
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9abc6e86-9f0a-46a1-a9ad-464abcfbd369 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring DeepSeek-R1 model card, 2025 b
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b2b7225e-4314-44fa-8c28-6ac86d1a5b4d · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring L., Jiang, L., Lin, B
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4326786a-5b26-4ffd-a8d6-14dc41f2c514 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Kimi k1.5 : Scaling reinforcement learning with LLMs , 2025
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6fb7d72f-9a0e-45c9-86f4-ac70a80e85ce · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring K., Dasgupta, I., Chan, S
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e6082ac6-0cec-4c51-8f52-da3b1e5774ec · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Introducing Llama 3.1 : Our most capable models to date, 2024
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 48f5b795-eb6c-4f2c-90fe-82f8459f3a0c · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a570a10-0985-418f-b7c4-58685e3469b8 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring A Comprehensive Overview of Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b468e4-69b4-432f-b894-221578b91067 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Hello GPT-4o , 2024 a
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dd7cb98b-9f09-42b7-852a-0ce6e76eac11 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Learning to reason with LLMs , 2024 b
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9eace8b9-8428-49f0-a6a7-cacd9842b9de · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring OpenAI o1-mini , 2024 c
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d099ac27-7d05-421c-83e9-4e7be89a6911 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring and Hassabis, D
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 434f714e-610f-4443-81eb-c96787744229 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Measuring and narrowing the compositionality gap in language models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 308db8f5-f1ff-4f17-9dcf-044cde671c0e · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring The Butterfly Effect of Altering Prompts: How Small Changes and Jailbreaks Affect Large Language Model Performance
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87e53046-c0a7-4f52-81dc-71d90f61234c · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring H., Sch\" a rli, N., and Zhou, D
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7981cc38-c4eb-4601-9ffc-71540aa28923 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring On the Self-Verification Limitations of Large Language Models on Reasoning and Planning Tasks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a252b1d1-2e84-4f6c-997d-bb3ee9f24002 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring LLMs Still Can't Plan; Can LRMs? A Preliminary Evaluation of OpenAI's o1 on PlanBench
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50496835-9d32-4382-ad34-5be903cfd868 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Grokked Transformers are Implicit Reasoners: A Mechanistic Journey to the Edge of Generalization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7cadd4a-96ec-44d4-bab7-8ac4d3c9a951 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Tree of thoughts: Deliberate problem solving with large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f15cfa4-2872-4e42-8026-5218c68e6246 · outbound
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring Larger and more instructable language models become less reliable
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 89aa5d30-c9ab-4169-ac72-b89355b454db · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
Reference 259
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.