Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2311.09724.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T20:42:44.448540Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T01:42:19.143537Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9956b270-e8c7-4a25-bc1f-1de8cec547a5 · inbound
Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dda4c43d-4eda-4702-894c-c170c1b47d5e · inbound
Improve Mathematical Reasoning in Language Models by Automated Process Supervision OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 36265d67-bc8c-45c6-a7fa-360777269ff2 · inbound
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation deb6154c-f8e9-43c0-b3a3-4d8d2e812720 · inbound
Reward-Guided Speculative Decoding for Efficient LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a378e7c-2876-4756-ab69-67ca96a89215 · inbound
From System 1 to System 2: A Survey of Reasoning Large Language Models OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 183
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c1e2bd89-95e1-4960-a808-d574518db6c9 · inbound
SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd0674c3-a3c1-4f74-af92-f26678f2827c · inbound
ProgRM: Build Better GUI Agents with Progress Rewards OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c048bd41-ef28-4f99-888c-ac0ddaf58a90 · inbound
UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a07ceb-aa06-48d0-9f29-7c11db37d6ad · inbound
BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3bb7a7e-4fed-436d-96e9-e780a1c9cff7 · inbound
Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 96475b4b-1752-4fe7-a134-21bb3150d19c · inbound
Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4bcc19fd-c171-4037-ac66-9ee1108858cb · inbound
Process Supervision of Confidence Margin for Calibrated LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation daf0c399-4c67-45a7-b326-e87d8f4b3b52 · inbound
Reducing Credit Assignment Variance via Counterfactual Reasoning Paths OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.