Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2503.04548.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:02:23.397708Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T04:27:36.567324Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 58656352-dd46-4280-83b0-75243c5d2710 · inbound
LongReD: Mitigating Short-Text Degradation of Long-Context Large Language Models via Restoration Distillation An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f33de9cf-eec2-4453-83db-1f09bebf594e · inbound
DAPO: An Open-Source LLM Reinforcement Learning System at Scale An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 06469bab-0143-4fcb-82dd-0f9ab749777c · inbound
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 76b6808a-bca6-4a3f-96b9-3f5d8dfea00a · inbound
ReTool: Reinforcement Learning for Strategic Tool Use in LLMs An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4381c426-e715-40cb-b99e-6afc554e7039 · inbound
Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 29297c44-ea38-4cf0-8954-a3cb1074da32 · inbound
WebThinker: Empowering Large Reasoning Models with Deep Research Capability An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1026facc-1e92-45cb-be99-89ba8001c0c8 · inbound
DiagnosisArena: Benchmarking Diagnostic Reasoning for Large Language Models An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b8875ab-b37b-43d1-add0-a48501d7948a · inbound
Prior Prompt Engineering for Reinforcement Fine-Tuning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 1901
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6f741f9-565d-4385-8f11-ea6c15b61ef1 · inbound
GUI-G1: Understanding R1-Zero-Like Training for Visual Grounding in GUI Agents An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5f52244-f3d1-4478-92df-621630b3f6d6 · inbound
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe80a669-587b-42f0-83c4-ecb28f4aebd1 · inbound
DeepRec: Towards a Deep Dive Into the Item Space with Large Language Model Based Recommendation An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 198ecfa4-6606-4742-af44-4607ce100a41 · inbound
LARES: Latent Reasoning for Sequential Recommendation An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 622a08fe-2874-4516-9be6-0009b06f3410 · inbound
Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b20eca8-405a-43c7-b82f-25ee47b0b245 · inbound
How Much Backtracking is Enough? Exploring the Interplay of SFT and RL in Enhancing LLM Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1dd7f1c-5a02-496d-920e-2483e7447a1f · inbound
Towards Effective Code-Integrated Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91706d69-cc3c-4652-96bf-4e1f55e08da2 · inbound
ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a14cc001-f48a-4c0f-801b-fff9e23f21f7 · inbound
Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7a66c95-c231-4ac5-800a-9fb4f645469d · inbound
Consistent Paths Lead to Truth: Self-Rewarding Reinforcement Learning for LLM Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66a3de8d-8fc8-4511-9668-9911376d2011 · inbound
CoRT: Code-integrated Reasoning within Thinking An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c297cfb4-f2b0-4ecc-91df-2dc7fee1cf47 · inbound
Act-With-Think: Chunk Auto-Regressive Modeling for Generative Recommendation An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d6b4f97-6902-456c-ba54-d30bb5150131 · inbound
Reasoning-Driven Retrosynthesis Prediction with Large Language Models via Reinforcement Learning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a097a0b-09ff-4ebf-8733-eafab4a5893d · inbound
Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f1e3433-c858-49cf-852f-f44ed4fbd56d · inbound
Why Does Reasoning Length Converge? Unveiling the Underfitting-Overfitting Trade-off in Chain-of-Thought An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b17cbcda-a229-4ac6-a328-5428da7d0df3 · inbound
Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c756cda-6e36-440b-b588-c5ad41ddeee5 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 628b4a99-10af-4df0-843a-96075de8680e · inbound
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f6ece2d3-1975-4510-b898-b35c9b2453fa · inbound
EEPO: Exploration-Enhanced Policy Optimization via Sample-Then-Forget An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a9173ba0-c924-44b1-b186-7d8790d0b7f9 · inbound
How You Begin is How You Reason: Driving Exploration in RLVR via Prefix-Tuned Priors An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6e29e227-7d4b-4b70-bf5f-5f7909f88ea4 · inbound
PruneTIR: Inference-Time Tool Call Pruning for Effective yet Efficient Tool-Integrated Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6e0ee3e9-a6f6-4d9c-8ee7-5edacf66493d · inbound
TimelineReasoner: Advancing Timeline Summarization with Large Reasoning Models An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 28301dbc-933c-438e-98be-bb5d6bead1b3 · inbound
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9a52c77f-62d9-4e31-bdcd-bf21981882e7 · inbound
RLVR Datasets and Where to Find Them: Tracing Data Lineage for Better Training Data An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02371df3-e049-457f-a350-f324403fbc61 · inbound
Trust Region On-Policy Distillation An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 289
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f19636a-2dcb-4fd8-9d09-89c43a4236a4 · inbound
GUI-AC: Enhancing Continual Learning in GUI Agents An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9ea71d90-980c-4f64-9ae2-54906282b8b8 · inbound
GUI-AC: Enhancing Continual Learning in GUI Agents An Empirical Study on Eliciting and Improving R1-like Reasoning Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.