Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:13:15.438371Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 1 inbound Pith citation observation for arXiv:2507.00726.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T21:13:15.438371Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T08:57:02.003590Z
A source-named dated measurement, never combined with another source.
Source: cited_works
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4f523d3a-94c7-4424-857c-5d1edba60f18 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a3217f-1ab5-4a3c-bac4-a3fa983b97bb · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 50699f88-3097-4635-9f23-0e6dc5c007cf · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Chessgpt: Bridging policy learning and language modeling
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1478d935-a06a-48d6-bd50-aada7116b6cc · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess The Llama 3 Herd of Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9e949a-8796-4961-8374-ef58ad103bde · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a27acc06-c949-4c8c-8792-6b2e0fbe7e01 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Learning to Reason for Long-Form Story Generation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 219e7b23-2f66-41dd-a309-695a3ec33382 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Improving regression performance with distributional losses
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 67fda46f-0282-43ae-ab8e-130dc8072787 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Bridging the gap between expert and language models: Concept-guided chess commentary generation and evaluation
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37d0b883-01d8-4c91-8fb0-10d8a25cfbaa · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12671650-3fc0-4ac1-a651-65143299ec16 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess LLMs Can Easily Learn to Reason from Demonstrations Structure, not content, is what matters!
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7265002d-5809-4e71-8fea-62d52dd94b3e · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Understanding R1-Zero-Like Training: A Critical Perspective
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf7ab677-32a9-444a-bad0-724cf9baecde · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6de8796-08e3-4c08-83e8-a16d9920ca50 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5723c9a-921d-4b84-8298-04b430d1952d · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Qwen2.5 Technical Report
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 720c2cf2-b5e0-4e70-ab0c-5b5526cd2638 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Amortized planning with large-scale transformers: A case study on chess
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf7edc7d-f230-4347-9a5c-06a11217689c · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Spurious Rewards: Rethinking Training Signals in RLVR
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f464150f-8539-46dd-9e52-8aaf5f9616dc · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dd87893-dd80-4af9-8fa7-5460330dd696 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess HybridFlow: A Flexible and Efficient RLHF Framework
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9e4522-fbce-4ae5-ac6d-afa829209472 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac25e364-83af-4805-bc3c-61e9b5e52045 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Attention is all you need
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32e410e0-aecd-43f8-b37e-3f295d39d486 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Explore the Reasoning Capability of LLMs in the Chess Testbed
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d94fe74a-0bdd-4957-91ea-b8875bf22864 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Reinforcement Learning for Reasoning in Large Language Models with One Training Example
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e8cd449-4306-4088-8306-33b72a2f453f · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3dbd584-e7dc-47d8-80f3-c14f00d2d54f · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffb13532-71c1-4418-b679-f952e7921f18 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30d2e018-9101-4a3e-b509-31e9071a9cfa · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess LLM as a Mastermind: A Survey of Strategic Reasoning with Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27f828f5-7772-42e0-878a-8a4b404ad8a7 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Complete chess games enable LLM become a chess master
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e6717eb-45a6-40a0-ab59-0e9be9669c13 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Distill Not Only Data but Also Rewards: Can Smaller Language Models Surpass Larger Ones?
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d424eeb-176a-41e2-9fe5-0ff6b21aa5d9 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Absolute Zero: Reinforced Self-play Reasoning with Zero Data
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbb9ee45-ebdf-46fe-9dd0-c32d2aec739b · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 325fed64-2de9-4f52-a07f-19e0f6dd0cb4 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f8ed20e-7cc6-4427-a66c-026994472dbf · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess @esa (Ref
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbae588f-ffbb-4ce6-8647-1327ecf9de9e · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c06d96-8721-4c7a-93a9-4a33430379b9 · outbound
Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7f9eae7-12e7-4379-aaea-ec90088c20bb · inbound
The Weight of Silence: A Causal Case for Weights Over the Scratchpad in Latent Chess Reasoning Can Large Language Models Develop Strategic Reasoning? Post-training Insights from Learning Chess
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.