Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:44:00.020826Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2507.07498.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:44:00.020826Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T08:40:40.910461Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T08:40:41.143168Z
56 of 56 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e97d3ca6-0613-403f-a167-1f9d2740b4e0 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Program Synthesis with Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3498fbf5-206a-4956-b0ee-4d4b3abd7f41 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Evaluating Large Language Models Trained on Code
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7db6b0e-5d22-48a2-b67e-999a3b922f9f · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3e74ca6-911d-4584-919d-901d130cc690 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a182342f-0939-4216-a5ad-c4014354a35f · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be213f86-508b-4eec-bd9d-7ce01221e1b0 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7be902d-f847-4705-b203-fc78ebf9373e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Min, Gail E
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00ba49be-2db1-48ac-bad9-5f6480f21142 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code SemCoder: Training Code Language Models with Comprehensive Semantics Reasoning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2b5160f-9463-4fb9-82bb-23e5fda2d77e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Min, Gail E
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63d3f85b-8bcb-49b3-acb4-4b9c2a70d255 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Kaiser, Wei Le, and Baishakhi Ray
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8999b04-aa9a-4d01-b602-c9f5584ecc5c · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9933a054-1d12-4da6-8fde-87d0afd5cedc · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9ee7c40b-946e-46c1-9991-3dda6ac8a2ae · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Are We Done with MMLU?
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2524e37f-3834-445b-ad3f-2766ed9f426a · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Pseudocode-Injection Magic: Enabling LLMs to Tackle Graph Computational Tasks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ea1aca0-17cf-44ce-bab3-d6912a0b2bc1 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89b15af0-e0ae-43b1-876b-86fffe9e5918 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b770dbb4-18f9-4cd7-8879-0dc42f6fe197 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5066ef0-7bd0-4cd0-84fa-93f0e1316849 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81f6f63a-f19c-4e28-9807-46b81b6a8dbd · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Measuring Mathematical Problem Solving With the MATH Dataset
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a38939c-a5c4-4976-952f-08e297d0c85a · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f54101d-8a92-4ba9-ae5e-9ccf6a4cad05 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Qwen2.5-Coder Technical Report
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c01f51ce-a523-47e5-ba9f-cdb2e14be61a · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code OpenAI o1 System Card
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc6c8bd9-bc18-447b-8b49-db539a07e775 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f775d378-1c6f-41d0-93b1-f1171f34f13f · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 493bf569-5162-4e7c-a312-81eb33bb145a · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca038d3-4070-4412-93fe-258d7d9f6090 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Tulu 3: Pushing Frontiers in Open Language Model Post-Training
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 705bca43-4958-4332-8364-62c8169b340e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code CodeI/O: Condensing Reasoning Patterns via Code Input-Output Prediction
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8d37bcf-d2fd-4899-ae03-7805c3e87a0c · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code MTR-Bench: A Comprehensive Benchmark for Multi-Turn Reasoning Evaluation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abe9e047-01dc-4272-a012-e5869a9c0d20 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code HellaSwag-Pro: A Large-Scale Bilingual Benchmark for Evaluating the Robustness of LLMs in Commonsense Reasoning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf41a461-6ba7-4a9d-97a6-6c755177b0fc · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 351480f6-b042-44dc-a133-69877045730e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 711345b0-5fcf-4d58-aae5-ae6f48db94f9 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code On the Impact of Fine-Tuning on Chain-of-Thought Reasoning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a09bcd-2a34-4758-b151-876c4c53b183 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b02a2f70-fb09-44ee-a3ad-ec2234bab7ff · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code KOR-Bench: Benchmarking Language Models on Knowledge-Orthogonal Reasoning Tasks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0270f618-f1c0-4595-84dd-2b347d0ec3fc · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Qwen2.5 Technical Report
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17feeeea-e0d2-4784-99ce-83f14916907e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67fb2b71-9929-4c31-8644-0cc18dac0ee4 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b48e226-7534-4af7-a9c1-1b5fcc0ea1a2 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62d1e1e1-8df0-4206-9df7-212aa93f435d · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4c774638-4500-4e1d-beac-2ba61ae4e4b4 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7dde5ff-939b-4562-b0e6-95af6b34ff72 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code SuperGPQA: Scaling LLM Evaluation across 285 Graduate Disciplines
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff4a753f-d84a-4aaf-b0b8-e6b537bc62dc · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code OpenMathInstruct-1: A 1.8 Million Math Instruction Tuning Dataset
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88366e24-9723-48b8-933d-b688da49350d · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38559dbf-f1ae-4979-aae0-abcdcdc53f6f · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code MathCoder: Seamless Code Integration in LLMs for Enhanced Mathematical Reasoning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db14ff2f-6586-4fba-a676-ab1b1c522760 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Le, Ed H
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2e0b8cd6-af3b-40a5-8efa-256c0de4201b · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bf2044ff-eb66-4004-bf51-d582acf36600 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Chi, Tatsunori Hashimoto, Oriol Vinyals, Percy Liang, Jeff Dean, and William Fedus
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 92874828-cc70-4d5e-b29b-aec626613d50 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b8b0e9da-da06-47ad-8c4d-09724995ea1a · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99a139d5-32ad-498a-880e-a8739508a026 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 498cf9e9-8b5a-42f6-9848-f9b4f31190ec · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Scaling Relationship on Learning Mathematical Reasoning with Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 536563bd-f0ae-4a38-ac31-fe7cc09dc943 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Does Reinforcement Learning Really Incentivize Reasoning Capacity in LLMs Beyond the Base Model?
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85f4a8c5-380c-40c7-ac76-6aedd63ed9e2 · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Learning to Execute
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7351096-4071-4a91-8b50-e1129681d1ba · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e5406b-ede2-49b5-b534-889a281da0ac · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code online" 'onlinestring :=
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cff075c-7ed1-42d4-929a-396445ae561e · outbound
Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code write newline
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd083f39-b368-43e4-831b-cbe8ff80f22e · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Teaching LLM to Reason: Reinforcement Learning from Algorithmic Problems without Code
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.