Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2605.03871.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T01:47:26.420486Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-03T20:08:56.018910Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation c2d48edd-559d-4464-89d5-8d2be67033fb · inbound
Self-CTRL: Self-Consistency Training with Reinforcement Learning EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cac28d9e-e93f-4c0b-ab92-811ff92d3df5 · inbound
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b1a90f9-52fb-4d0b-83e5-5c32b008ee46 · inbound
SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51aa791f-3ed6-4313-9953-0cc108fd7c01 · inbound
SERPO: Self-Evolving Rubric Policy Optimization for Open-Ended Test-Time Reinforcement Learning EvoLM: Self-Evolving Language Models through Co-Evolved Discriminative Rubrics
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.