Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-26T11:10:56.755558Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2606.22317.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-26T11:10:56.755558Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:08:22.453766Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T00:08:30.124848Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e539e23a-43b7-4966-b543-bbe156f47c81 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Rl for reasoning by adaptively revealing rationales
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 787e5272-d8f9-4bc4-b96c-1ffedd6855f8 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Online difficulty filtering for reasoning oriented reinforcement learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a135eb6-9869-4771-b9e7-e43eefdb1b66 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 53c92407-fa65-4bcd-a92d-6d8c29db6b0e · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Unveiling the key factors for distilling chain-of- thought reasoning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e2a08b-f532-431c-846a-ddc673baa1ad · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Training Verifiers to Solve Math Word Problems
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 64a05973-4fc6-49f6-af85-097bb1ea1560 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3151905a-5e48-41e1-b11a-8801af193a51 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning.Nature, 645(8081):633–638, 2025
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad351469-98a8-4e59-8bd9-405ba2ef7126 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Omni-math: A universal olympiad level mathematic benchmark for large language models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8d10b29-0fe4-4ea3-b166-67a56fbff509 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation efe3cd24-14c7-4002-9329-7febe15112df · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Rewarding the unlikely: Lifting grpo beyond distribution sharpening
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d75c0abc-3b57-43a6-87f4-ba96f7b181e6 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Measuring mathematical problem solving with the math dataset
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe5532e9-d0c2-4fd2-ae39-4ada51a27e06 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Distilling step-by-step! outperforming larger language models with less training data and smaller model sizes
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb9e5769-20cf-4bb0-9e43-29e2eb5fca8d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model R-Zero: Self-Evolving Reasoning LLM from Zero Data
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d3229f20-a436-4298-9568-fa00c3fededc · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Mitigating Catastrophic Forgetting in Large Language Models with Self-Synthesized Rehearsal
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 323258c1-76b1-450c-9c1d-2d951894027c · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Mitigating Catastrophic Forgetting in Large Language Models with Forgetting-aware Pruning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 891db1dc-2849-41a6-beab-792e3241c235 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Unlocking the power of function vectors for characterizing and mitigating catastrophic forgetting in continual instruction tuning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17048d55-71be-4e19-acdb-9a311510ee34 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Tacler: Tailored curriculum reinforcement learning for efficient reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 17451c98-d139-4776-a6b5-bc6c2dd4e791 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Language models can easily learn to reason from demonstrations
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation decc794b-e1f0-437d-9233-7f06515cb487 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Prorl: Prolonged reinforcement learning expands reasoning boundaries in large language models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8ad80ca-2fa3-494f-8906-9fb1e7df4371 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7482cacf-f7f2-4d09-a49e-852c9a44988d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model An empirical study of catastrophic forgetting in large language models during continual fine-tuning.IEEE Transactions on Audio, Speech and Language Processing, 33:3776–3786, 2025
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edaefd69-d03e-481a-8969-5ab596b96f69 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model s1: Simple test-time scaling
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 723e8753-b778-4129-922a-d40edde2e06e · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Openai o1 system card, 2024
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 016d0c7a-9751-40a5-8080-0e77cc613d27 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Curriculum reinforcement learning from easy to hard tasks improves llm reasoning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8362972f-329f-4e83-bec4-80781be32119 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Seed1.5-thinking: Advancing superb reasoning models with reinforce- ment learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b8c90504-fcc3-44f6-8b25-198bfd340ed9 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 7ef419e3-2fbe-40db-8337-6893563729c0 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ab311a9c-566d-43ae-b144-be5d30a34c1d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a0133171-3f07-48a0-9401-bcdd1fa39746 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8c149a40-a299-411c-8ba8-5a3bf370ad9c · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Continual gradient low-rank projec- tion fine-tuning for llms
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06b8fae7-b780-4c1f-8dbe-e5be94315c3f · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Inscl: A data-efficient continual learning paradigm for fine-tuning large language models with instructions
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7643db9-d66f-4523-901c-3a4510bdcdbb · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Reasoning scaffolding: Distilling the flow of thought from llms
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5286da5-28a9-4042-aef8-60ae367d41e6 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Reinforcement learning with verifiable rewards im- plicitly incentivizes correct reasoning in base llms
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b972c2da-bbc1-4150-bcd6-8612b0d312db · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Enhancing long-chain reasoning distillation through error- aware self-reflection
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c5d14caa-9f01-4289-b406-a09d72b93085 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Training large language models for reasoning through reverse curriculum reinforcement learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 839f41f2-2611-43a2-8774-b6641b528f98 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Learning to Reason under Off-Policy Guidance
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fbefa8f7-7a04-451b-93e7-c40665df5795 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 718c7f29-425f-4845-8d62-542048da84d5 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Qwen3 Technical Report
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 95a29bb8-2dc5-4a77-bfbf-602c4f97dede · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a70a5b3d-ba88-4aea-8d5f-8f72a9783c5d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Does reinforcement learning really incentivize reasoning capacity in llms beyond the base model? InThe Thirty-Ninth Annual Conference on Neural Information Processing Systems, 2025
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7a37b8b-b608-4975-b3e9-7ecac6a0816d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Simplerl-zoo: Investigating and taming zero reinforcement learning for open base models in the wild
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c27be518-5a20-411d-b4b7-0546b5c51438 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Cures: From gradient analysis to efficient curriculum learning for reasoning llms
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18cc8457-84ee-4f9a-b06a-a2255195ebb4 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model On the interplay of pre-training, mid-training, and rl on reasoning language models, 2025
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36f0bdf7-7d43-43d0-b75f-8fb4811e8bfd · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Kakade, Cengiz Pehlevan, Samy Jelassi, and Eran Malach
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c9bfd3-6cf0-48e0-8cb6-b7c8f7025cd5 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Automatic curricu- lum expert iteration for reliable llm reasoning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25c971c0-9475-447f-a550-646d002e9c3e · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model ProcessBench: Identifying Process Errors in Mathematical Reasoning
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9860daca-8e29-48b4-8ae9-2150d7e84b98 · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Group sequence policy optimization, 2025
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f250876-a5da-45b9-91d1-02e9001bbc6d · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model Spurious Forgetting in Continual Learning of Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c0543458-410c-43b3-813c-6a4ba4d1f2cc · outbound
Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model TTRL: Test-Time Reinforcement Learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 98747308-700e-46db-9866-9d4b7a981365 · inbound
Question Begets Question: Self-Evolving Curriculum for Reinforcement Fine-Tuning on Competition Mathematics Curriculum Reinforcement Learning Can Incentivize Reasoning Capacity in LLMs Beyond the Base Model
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.