Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:1803.03835.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T22:09:27.590185Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T21:47:36.272825Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 9bbcc9ef-b5f9-43d7-926e-fff712bd30c2 · inbound
Attentive Multi-Task Deep Reinforcement Learning Kickstarting Deep Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b06cf659-0c8d-4075-8eee-f93eac5afe06 · inbound
Red Teaming Language Models with Language Models Kickstarting Deep Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b5f2b6c2-94f6-4b43-89a0-7f17f3947cb0 · inbound
Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models Kickstarting Deep Reinforcement Learning
Reference 287
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0070de22-2eaa-4108-927a-28b1890f182b · inbound
Proximal Policy Distillation Kickstarting Deep Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b3d45bd2-b340-4d88-ae73-794929bd22c4 · inbound
Towards an Autonomous Test Driver: High-Performance Driver Modeling via Reinforcement Learning Kickstarting Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92c2c69f-d3e7-4a1b-bb6f-741fa4ca8df8 · inbound
All You Need in Knowledge Distillation Is a Tailored Coordinate System Kickstarting Deep Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b5891fb-3032-4457-a142-a333a4535392 · inbound
Embodied CoT Distillation From LLM To Off-the-shelf Agents Kickstarting Deep Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c28a6a4-bf45-4af6-8a3d-eac8f7dd2c37 · inbound
Energy-Based Transfer for Reinforcement Learning Kickstarting Deep Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b143ffd-91bf-4246-9cb8-0f381273ab27 · inbound
ThinkTuning: Instilling Cognitive Reflections without Distillation Kickstarting Deep Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f63d06b-0b17-4988-93e5-828fe488091a · inbound
EAPO: Enhancing Policy Optimization with On-Demand Expert Assistance Kickstarting Deep Reinforcement Learning
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63603a1f-eb18-448c-9a70-5af36f61f324 · inbound
TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations Kickstarting Deep Reinforcement Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 665fdad2-4582-4307-9275-b3d615d47f89 · inbound
TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations Kickstarting Deep Reinforcement Learning
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d55c6583-052f-4f63-8556-312c744e48f4 · inbound
MotionPyramid: Hierarchical Motion Representation and Residual Interfaces Kickstarting Deep Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 52ca7b67-a72c-4c0f-b474-31b7fc32263c · inbound
Efficient Long-Horizon Learning for Learned Optimization Kickstarting Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d380fda5-b604-4a91-8d03-f0e006fcaa4a · inbound
Efficient Long-Horizon Learning for Learned Optimization Kickstarting Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b3ff2b-e8ba-4721-bba3-699c5eb6393d · inbound
Efficient Long-Horizon Learning for Learned Optimization Kickstarting Deep Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d68ed56d-7968-4fa7-afb0-f96d69f4236a · inbound
Leveraging Offline Supervision for Efficient and Generalizable Reinforcement Learning in Large-Scale Vision-Language-Action Models Kickstarting Deep Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.