Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T07:34:57.251159Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2607.24833.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T07:34:57.251159Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
48 of 48 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 81b18e99-c7e4-4435-892c-76782d9ce4ee · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2026 , note =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 244961c9-8f48-4f6b-8f4c-cce26612c28c · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2025 , note =
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b29a83-3f63-4861-b4e5-29a1d2c9d395 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef391fe9-6136-4169-a24a-8f0556461d0b · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bacec8d-f120-4ce3-a2b7-db03af9e9a12 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87a9d1a3-6d2d-4968-946a-4afa7233e0eb · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cd58a9d-cc36-4cf6-98aa-2e3635c3c67f · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35971f69-dcd6-420b-b6df-2ec85b50d6cf · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56b98f94-39ec-4418-b207-d4eaaeecbfb5 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03fce6f1-890c-438b-91ae-5704d09df196 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a91bba5-0c5d-40b4-823a-56c3226fb3c6 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning , booktitle =
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 501596e3-1294-47b5-ae47-ff65e4dff05b · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning , journal =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebb1ed94-b1b7-42a9-8472-a027ac6eb1e8 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caa7f80d-8fa4-40ef-9fb1-2d1aef61493f · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning and Zhang, Hao and Stoica, Ion , booktitle =
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1df552fc-4245-4de6-ab8b-e5cb5d41a851 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2025 , note =
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecadc99e-6b6c-40d1-b1cd-007ea25f2f97 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2023 , note =
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27bc778d-a9ab-4877-a035-5bc786942a90 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b691810-3838-4793-961b-6550e5a5eb06 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a4c51a-2945-45bc-a7bd-2a6e11a2109a · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8569224e-5a1b-4503-bc37-8c8a8ad3c12c · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Machine Learning (ICML) , year =
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0fd8522-fee6-4c9a-b44e-29d9c7ae87e5 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d5b5313-5282-441e-9b9e-0622dcf199e2 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Don't Tell the Answer, Truly Guide the Reasoning During
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c9d8cfe-5ecc-4540-9d30-8741ced11c91 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45429833-a9ff-4208-938f-93cba4343624 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26c06b32-36c7-43d3-ba09-8a252be4e03a · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Staying in the Sweet Spot: Responsive Reasoning Evolution via Capability-Adaptive Hint Scaffolding
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6635fc8-6287-485f-9f4f-f0b72cdaf244 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Boosting
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dfea157-c72a-4a0e-aaac-d803c40e3597 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Train at Moving Edge: Online-Verified Prompt Selection for Efficient
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19fca2ce-0a71-4027-b71a-d83b82273fa7 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0733587a-51d5-43fe-bd4f-8fb3e44c8c28 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6156148b-691e-451a-9ee1-c8ad2f6d81d7 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c30af885-3f78-4ee6-8acf-9c45f11113fd · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Measuring Mathematical Problem Solving With the
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2172d4fc-2eaa-4f1a-bd87-a8cd7a30081e · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Training Verifiers to Solve Math Word Problems
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30538229-4d0d-43fc-8c6b-b9800ed7363b · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02d054e1-546e-4910-b105-ac5b7fbd7c13 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 036d2c91-1866-401b-bfe9-3eb6cbb782ba · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d372dc61-e76c-4364-9a82-5502f9bf78b9 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72e1f0de-0a34-46c2-a54b-8f58af08b9a5 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b6b5067-8c38-4037-b5e7-595390f58244 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c956bc34-fda9-4db0-b9a6-724bbd67aa86 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Back to Basics: Revisiting
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53302985-149c-4556-896b-0afd4b18c722 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2024 , note =
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 897bbd8e-ef9a-4d1b-a8bc-958357b59bb2 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Advances in Neural Information Processing Systems (NeurIPS) , year =
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1535bde9-33fa-4139-9187-8d966b499f80 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Reinforced Self-Training (
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a60d578-c1c2-4792-8300-0387304d60ce · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Transactions on Machine Learning Research (TMLR) , year =
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46f0c937-c1c1-42b7-b42e-7b9169cd431c · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Teaching Large Language Models to Reason with Reinforcement Learning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 389b8b2e-45e0-4354-bbc4-4edca3600dab · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning 2016 , note =
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b79782fc-1867-4cda-847b-521e8cb4e170 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning International Conference on Learning Representations (ICLR) , year =
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 294c8b0e-9dfd-4cb3-8a34-8b5fa2165ff2 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Evaluating Large Language Models Trained on Code
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f533e7ed-2f38-4c6f-8b49-9056c522fab1 · outbound
AdaKP: Online Adaptive Knowledge-Point Selection for Reasoning-Oriented Reinforcement Learning Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.