Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:40:21.583101Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2606.02113.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:40:21.583101Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-28T09:56:36.860369Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
18 of 18 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation 4ae6c1e7-fad5-4de1-90c8-13d89c3c5446 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Miranda, Hanyang Zhao, Mohammad Rifqi Farhansyah, Garry Kuwanto, Derry Wijaya, and Genta Indra Winata
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 808a10a8-30b8-44f4-979b-5d8fad005d58 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Distillation Scaling Laws
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9f9e72e2-5555-44cf-8ce9-38574fba3712 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Evaluating Large Language Models Trained on Code
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dd66c078-fe3a-462f-8955-428831c3e560 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Training Verifiers to Solve Math Word Problems
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3f39b1b4-ded4-4070-8a1b-a2fb37d013ba · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Mind2Web: Towards a Generalist Agent for the Web
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6b6c7bd1-eb96-4875-9c45-fd7ca5f451fc · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works WorkArena: How Capable Are Web Agents at Solving Common Knowledge Work Tasks?
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 82041640-7447-4d81-9384-28a5337c6b9f · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Rubrics as Rewards: Reinforcement Learning Beyond Verifiable Domains
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 56d37dab-4fe5-41a2-8e94-82c19423c538 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Omni-RRM: Advancing Omni Reward Modeling via Automatic Rubric-Grounded Preference Synthesis
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 67e9b3c8-9058-4b56-9138-3a6d0a996b31 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Let's Verify Step by Step
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7d2a08dc-3f14-428a-b890-da61d4791cf8 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1b858ce7-4e44-4d77-a08b-8ca5bbb14aa8 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Magistral
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3ef11069-d661-469e-9df3-f117a04d2957 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works AndroidWorld: A Dynamic Benchmarking Environment for Autonomous Agents
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fa177d23-21bd-4573-af6e-34494a0d7843 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works AppWorld: A Controllable World of Apps and People for Benchmarking Interactive Coding Agents
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b4e51fb7-0546-453b-bd12-eb062a79c0af · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 182ff33b-f889-44ee-819c-54eae366d611 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works Naturalreasoning: Reasoning in the wild with 2.8 m challenging questions
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 59bda563-1572-4571-b99d-1c27eae965db · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works InThe Tenth In- ternational Conference on Learning Representations, ICLR 2022, Virtual Event, April 25-29, 2022
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 117f0cc5-3636-4ff7-aef3-fa8454790d5f · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works TAT-QA: A Question Answering Benchmark on a Hybrid of Tabular and Textual Content in Finance
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 335b65e9-9f07-4abe-9224-25d17d13e7e9 · outbound
A Primer in Post-Training Reasoning Data: What We Know About How It Works JudgeLM: Fine-tuned Large Language Models are Scalable Judges
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c4c5b44-b8a2-4d3f-9901-babc99db77f9 · inbound
RealClawBench: Live OpenClaw Benchmarks from Real Developer-Agent Sessions A Primer in Post-Training Reasoning Data: What We Know About How It Works
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.