Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:54:34.992032Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2506.12366.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:54:34.992032Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ff819525-73b8-41fa-b0f1-209cf67f3208 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 62341f45-5ae2-4e76-97d0-b1477f3800f1 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Playing Atari with Deep Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9c691f7-9c49-429c-aa7c-8f33f9ce667d · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc84f432-7f24-4c6c-bdac-2375dc1ef1b7 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62b23a52-8c10-4b28-8f78-7cf8a2647170 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Solving Rubik's Cube with a Robot Hand
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3628204b-3c6f-4f7e-aed3-8e1bc22dc60a · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Assessing the Impact of Distribution Shift on Reinforcement Learning Performance
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8067b51-5833-42ea-bdd8-8ddf3eaa0c66 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f0afb74-432c-4faa-af13-36455bc8470c · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eae48b2e-276e-42bd-a27d-a3ad041d95a6 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning & Krueger, D
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b85e1bf7-f3de-45f9-ab9b-376693f31989 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f4514670-e34f-4f9d-8d03-cc0053e9feb2 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1b0bfa77-eac1-43fc-8dea-8a54105fb0c4 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Explaining Reinforcement Learning Agents Through Counterfactual Action Outcomes
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cae01744-3e5b-4045-b837-214e3d193530 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning & Fal- cone, F
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a8df109-521d-4be5-b39c-fc5d20ccc265 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 05b5c852-5616-4152-b1c2-c744954b8c4e · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning ARMADA: Augmented Reality for Robot Manipulation and Robot-Free Data Acquisition (2024)
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24ff9008-bd64-4e78-9ffa-eaa4f21ff008 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Simulated augmented reality and virtual immersive technologies
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2acbf66a-3a1c-4421-8c16-87365c2f07d0 · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning & Yablon, Z
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60cc0211-e0a1-4795-80a4-b7c81e549c6c · outbound
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c28be68d-25d9-4e01-bd62-8f058b35881b · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cb1e6732-4b91-4ebf-a565-2c1f2eb8f6a3 · outbound
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6735d5a7-b288-4389-a2a1-463332410c9d · outbound
Ghost Policies: A New Paradigm for Understanding and Learning from Failure in Deep Reinforcement Learning Unity: A General Platform for Intelligent Agents
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.