Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2508.03682.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T21:09:21.389396Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:58:58.316829Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 8c8d2f87-bfbc-4ea3-8334-e0f83cb25660 · inbound
A Survey of Reinforcement Learning for Large Reasoning Models Self-Questioning Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a084081-9947-4121-9bec-9dc3ad2cca4f · inbound
EvoLMM: Self-Evolving Large Multimodal Models with Continuous Rewards Self-Questioning Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5e37e46-4a44-4e0f-8c28-3a05763a3798 · inbound
Toward Training Superintelligent Software Agents through Self-Play SWE-RL Self-Questioning Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2880e25e-a321-4959-935d-15a0d6ae23cf · inbound
Toward Training Superintelligent Software Agents through Self-Play SWE-RL Self-Questioning Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e32f2d2d-3250-43b9-bbf7-511323954271 · inbound
CPMobius: Iterative Coach-Player Reasoning for Data-Free Reinforcement Learning Self-Questioning Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0bfe9b8-82a5-4af8-9402-35c51397c8d1 · inbound
RoboAgent: Chaining Basic Capabilities for Embodied Task Planning Self-Questioning Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 43027edd-be8f-4a3b-b244-6388bd7b93de · inbound
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision? Self-Questioning Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8f25e5f4-4669-4a97-8c08-dd1aa556969b · inbound
$\pi$-Play: Multi-Agent Self-Play via Privileged Self-Distillation without External Data Self-Questioning Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8182d35b-5aeb-4e97-86eb-d7dbbf51b48b · inbound
Scaling Self-Play with Self-Guidance Self-Questioning Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 78b4d7ec-11b1-4f04-9e1e-47962ac8ffb8 · inbound
$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Self-Questioning Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 021d189e-29a1-4fe7-a7c1-363e4ff8ef89 · inbound
$S^3$-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data Self-Questioning Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4e726b3-3b80-4d03-850a-2e1b03df4772 · inbound
OracleTSC: Oracle-Informed Reward Hurdle and Uncertainty Regularization for Traffic Signal Control Self-Questioning Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c7a8fe9-b9e9-46b2-a8a9-86342496dbc4 · inbound
SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation Self-Questioning Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9def4cfd-1b4e-4b59-b8a8-7b96f13bd77e · inbound
Trust Region On-Policy Distillation Self-Questioning Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a761f4d1-1808-416b-b76b-3a7871294ac0 · inbound
Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning Self-Questioning Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fa189e1b-2278-49a5-8765-59acbdebf8b3 · inbound
Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning Self-Questioning Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 850dfe31-0c2a-4615-9b2b-24a655c05d08 · inbound
From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning Self-Questioning Language Models
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5f606229-6c63-4d23-a099-918d50891d13 · inbound
Anchored Self-Play for Code Repair Self-Questioning Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.