Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2411.00062.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:09.274466Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 7bce9548-c0fb-4385-a5b5-0fbd53fc1461 · inbound
Lifelong Safety Alignment for Language Models Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d49a9bb5-c74c-41b5-b597-caccd9fa8c15 · inbound
Chasing Moving Targets with Online Self-Play Reinforcement Learning for Safer Language Models Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 294a7bac-dff0-4969-b010-1ade01a7fecc · inbound
Teaching Models to Teach Themselves: Reasoning at the Edge of Learnability Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea7dfb6-a5bb-4ab1-b5fc-eb9b51d96dc9 · inbound
FrontierSmith: Synthesizing Open-Ended Coding Problems at Scale Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 00b19379-6e98-4624-ad30-e618639a9bb2 · inbound
PopuLoRA: Co-Evolving LLM Populations for Reasoning Self-Play Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ff97028-1fd7-4a5f-a61d-4d5f6b572086 · inbound
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL Scalable Reinforcement Post-Training Beyond Static Human Prompts: Evolving Alignment via Asymmetric Self-Play
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.