Pith. sign in

Paper Citation Record · LEDGER

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories

As of 7 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2605.21468.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.21468 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T05:19:11.738106Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact11
  • verified fuzzy0
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 045dee46-67ac-4fc4-92e9-20585a9b324b · outbound

This paper cites Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.174452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:93f119c18c0466cb8b33db3ef8e76e174161d14436e5f6d6d566c905fc62241b

Observation 5df37922-13a6-4002-bb1d-74195db52626 · outbound

This paper cites Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Beyond Benchmarks: MathArena as an Evaluation Platform for Mathematics with LLMs

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.166197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:116cbd2f9bb773189060f806c3d7a4be89197b51727985e76819e45c0beb00c4

Observation 7dedee24-7b67-40a0-a187-35432e041e24 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.189119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:e4f23fa496bb46df61775d5f96d5de759a6d56bb9c1c4b014219526f2eb62f6d

Observation 54a19620-adf0-4f3a-a51b-a3ce7e9f0f15 · outbound

This paper cites On the Emergence of Implicit Curriculum in RLVR Learning Dynamics.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories On the Emergence of Implicit Curriculum in RLVR Learning Dynamics

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.157678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:2fa55419fc720f3553e7b2db3b687efdaafd3c660a32e4bd51049a686dcb52e2

Observation 83764595-42ad-45c5-8e80-3190892d8a55 · outbound

This paper cites Scaling Laws for Neural Language Models.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Scaling Laws for Neural Language Models

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.163459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:fdc32e739b1485b2d99db6fdf02417b35a9607cbda272af77b570899a7b19e94

Observation 102f4e88-2b6f-4b7c-be28-6fa098c5dfc3 · outbound

This paper cites Olmo 3.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Olmo 3

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T05:19:39.171370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:ba7879535e282d0b45d479552098d298fa517bba3f324614dd3eabd5e99e9801

Observation 55a123c7-fb5f-4313-9655-089cf116aeca · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.168876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:af92b88235e0abf7a8262ca4fa13db6da81a6715da840cfe03d0632f3e577a58

Observation d57c11c3-5dc4-4f5c-ad18-85729c301de2 · outbound

This paper cites Linear Dynamics in the RLVR Training of Large Language Models.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Linear Dynamics in the RLVR Training of Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:03:53.599274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:67402a9279d1befac93276d2a6ed27c7c14e422f7254942612b93bf4231b0538

Observation b6c925f6-e75f-4990-aa3b-339165ac8fe2 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.191872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:d5aeb11f8bb6dcec0bd65ef691ef736701a6a280ef5c848a466886cc083001b4

Observation a4e27f99-c2b6-4a74-8344-1a82f5bb7122 · outbound

This paper cites Qwen3 Technical Report.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Qwen3 Technical Report

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.160804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:7afa1d96e6262c339b2a5ac7e3816ed5718ba82d26cd3d382f1ddde39343ab8e

Observation 00206f39-4726-464a-aabe-8eef1fc32c5f · outbound

This paper cites On the Implicit Reward Overfitting and the Low-rank Dynamics in RLVR.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories On the Implicit Reward Overfitting and the Low-rank Dynamics in RLVR

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:19:39.180936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:900388a374c7a449946b80f310e1577cc8bdeaa64855740ba0c50c8b06b83d63

Observation f4db0f03-3c1b-4abc-9cac-110cfbffa626 · outbound

This paper cites arXiv preprint arXiv:2506.07998 , year=.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories arXiv preprint arXiv:2506.07998 , year=

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:19:39.183691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:aeaac50938e5cc427b63ad58e972d27aa9db80f338f5de5aa42a77427cb70807

Observation 70d1e977-8f55-4950-aa5b-09ed2d940e21 · outbound

This paper cites Pan, Zhangyang Wang, Yuandong Tian, and Kai Sheng Tai.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Pan, Zhangyang Wang, Yuandong Tian, and Kai Sheng Tai

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T05:19:39.186539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:4f605e3f53d7fa899d46b2b1aed6c2075a7d5e12a0e48274587d5bd39b37a4c6

Observation 9c986d15-323e-4577-a3c1-134a955de13c · outbound

This paper cites an unresolved cited work.

You Only Need Minimal RLVR Training: Extrapolating LLMs via Rank-1 Trajectories Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-05-21T05:19:39.605503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:19:11.738106Z digest=sha256:67820c08542e858e79978b9d72dc19ccae71655e9a9fdf52a2a27529bffb76bd

Pith citing papers

No inbound Pith citation observations are available.