Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:31:28.767651Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 0 inbound Pith citation observations for arXiv:2507.15844.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:31:28.767651Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 282d0983-76a0-4e7d-9269-597ef0fabc77 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc1f98b9-2226-48e4-a27a-b217e678a2d8 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URL https://doi.org/10
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a4f7ba3-08d8-4a34-9586-a2519ef98720 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fc7dd44-8184-4f22-b1a9-b918be534d1f · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c0aaae6-8018-48ee-8579-0d863a965da8 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Thinkless: LLM Learns When to Think
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87229f48-e3d9-4439-b38f-18c0427e85b8 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c3f112-3400-466c-8a33-b1f78825ba81 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c534d5d1-4d77-4b2e-b298-c0cf235de80c · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URLhttps://doi.org/10.48550/arXiv.2505.11225
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0714b9d-f4a4-46c2-9aee-ba6de4f8bcca · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Think Only When You Need with Large Hybrid-Reasoning Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02bf1114-67cd-4d06-8c9d-f67a7d202e2f · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3754db1e-dab9-47d9-9389-9b7d56f7f4d3 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning ThinkSwitcher: When to Think Hard, When to Think Fast
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06e9ff18-fa4c-457d-ab68-213df15f9805 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c52a92a-bdbe-41f5-8298-c7fb10f3f786 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Reasoning Models Can Be Effective Without Thinking
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 040d2d23-8868-452a-8f41-466916fccb9b · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Reasoning Models Can Be Effective Without Thinking
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35312d21-cb31-4f3e-8bda-dc2b7aae9586 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Imitate, Explore, and Self-Improve: A Reproduction Report on Slow-thinking Reasoning Systems
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5968c0c8-97cd-47c5-a1f7-f0cf736660e6 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning s1: Simple test-time scaling
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08e6a92c-a9df-4c47-970e-d32541be5277 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Accessed: 2025-07-22
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b823124c-5b7f-4289-abb1-10c120f67e88 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URL https://doi.org/ 10.48550/arXiv.2505.04881
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a077caf7-ac67-47b6-8c0f-5df45d927f99 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Learning when to think: Shaping adaptive reasoning in r1-style models via multi-stage RL
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c5a568d-8990-4f95-a77f-857af800e004 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URL https://doi.org/ 10.48550/arXiv.2505.10832
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a5e8ee1-c570-490f-991a-e77582d4a762 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning PATS: Process-Level Adaptive Thinking Mode Switching
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44edbed9-a921-4666-a661-fa77e446d855 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URL https://doi.org/10.48550/arXiv.2505.20258
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d4571c65-7566-4e78-958d-fde3db65fa6b · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Scalable Chain of Thoughts via Elastic Reasoning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092df8c9-b248-4eca-91ce-1631970e42fb · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning URLhttps://doi.org/10.48550/arXiv.2504.15895
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a884b321-ad3e-4dd1-8214-a6eead52d5c3 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68b378fb-239c-496d-a120-d98241007ce4 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b63b925-88cd-4732-a3ce-000e546a4658 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7fbf56c-13cb-4258-ac3f-e4ae51ea39cc · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4af2d2de-57bc-4c3d-b51b-d6bed7cff138 · outbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.