Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2502.12853.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:51:30.691200Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T06:54:01.056295Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 3c58cff3-de1e-4c63-b657-f717887b1988 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 200
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 2d5e7b6f-4e1c-4faf-9f00-b351615363dd · inbound
Boosting LLM Reasoning via Spontaneous Self-Correction S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e812ee5e-535c-40df-9a12-ddc265680545 · inbound
PAG: Multi-Turn Reinforced LLM Self-Correction with Policy as Generative Verifier S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31143d57-8403-4307-9567-3d65736b6bca · inbound
CogDual: Enhancing Dual Cognition of LLMs via Reinforcement Learning with Implicit Rule-Based Rewards S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 686afce3-f261-4af1-b3ef-9204c8fdb512 · inbound
Self-Reflective Generation at Test Time S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 391e202b-2abc-490a-af25-b77342bd120f · inbound
Towards Sparse Video Understanding and Reasoning S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa851a64-26fc-49e8-a29b-ea8e5102c6e2 · inbound
SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f2f9c1d9-9e73-4d04-9ba7-3a2762563404 · inbound
SVSR: A Self-Verification and Self-Rectification Paradigm for Multimodal Reasoning S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e91f14a7-9f89-43fd-94ad-3a06bde8597e · inbound
Multi-modal Reasoning with LLMs for Visual Semantic Arithmetic S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 90a1b760-73b2-47f8-8e9b-649ec3c0717e · inbound
Internalizing Outcome Supervision into Process Supervision: A New Paradigm for Reinforcement Learning for Reasoning S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 68b6e2e8-4366-47fe-95da-e734b1f24516 · inbound
The Hidden Signal of Verifier Strictness: Controlling and Improving Step-Wise Verification via Selective Latent Steering S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b5942e9d-b718-4427-8708-efc52af0e119 · inbound
Spectral Origins of the Self-Correction Blind Spot in Autoregressive Generation S$^2$R: Teaching LLMs to Self-verify and Self-correct via Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.