Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T05:19:11.738106Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2605.21468.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-21T05:19:11.738106Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
14 of 14 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 045dee46-67ac-4fc4-92e9-20585a9b324b · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5df37922-13a6-4002-bb1d-74195db52626 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7dedee24-7b67-40a0-a187-35432e041e24 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54a19620-adf0-4f3a-a51b-a3ce7e9f0f15 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories On the Emergence of Implicit Curriculum in RLVR Learning Dynamics
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 83764595-42ad-45c5-8e80-3190892d8a55 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Scaling Laws for Neural Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 102f4e88-2b6f-4b7c-be28-6fa098c5dfc3 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Olmo 3
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55a123c7-fb5f-4313-9655-089cf116aeca · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d57c11c3-5dc4-4f5c-ad18-85729c301de2 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Linear Dynamics in the RLVR Training of Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6c925f6-e75f-4990-aa3b-339165ac8fe2 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4e27f99-c2b6-4a74-8344-1a82f5bb7122 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Qwen3 Technical Report
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00206f39-4726-464a-aabe-8eef1fc32c5f · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories On the Implicit Reward Overfitting and the Low-rank Dynamics in RLVR
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4db0f03-3c1b-4abc-9cac-110cfbffa626 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories arXiv preprint arXiv:2506.07998 , year=
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70d1e977-8f55-4950-aa5b-09ed2d940e21 · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Pan, Zhangyang Wang, Yuandong Tian, and Kai Sheng Tai
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9c986d15-323e-4577-a3c1-134a955de13c · outbound
You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.