Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T23:17:52.997124Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2509.06714.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T23:17:52.997124Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4c6c90fd-7ab9-4913-afb6-a314dfb57db7 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Mastering Atari Games with Limited Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c0ed7d4-ee9a-45aa-a5db-40db248b73c6 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Benchmarking deep reinforcement learning for continuous control,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 18cb8405-05ea-4bf5-808b-2629b29ce7b5 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Challenges of Real-World Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d1b1fb9-2eff-4342-af17-76a135ed5001 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Deep rein- forcement learning in a handful of trials using probabilistic dynamics models,
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79dc81a1-e6d5-48e5-87c3-59d3c67a005f · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms pytorch implementation of PETS,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 91d08693-7794-4eb0-a448-98289fa54a57 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Temporal difference learning for model predictive control,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 548e413c-2f3b-4059-bf33-79a1ffd82d68 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms A tutorial on the cross-entropy method,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d912b7dc-cd61-4b6d-ac16-2ac9599f30c3 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Physics-informed model and hybrid planning for efficient dyna-style reinforcement learning,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1f429d61-5cef-4fc6-a29e-139f40dba3a5 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Swing-up control of inverted pendulum using pseudo-state feedback,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0771da8d-6426-4dda-987e-382c391ab11e · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Exploring Model-based Planning with Policy Networks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c06e1088-12c5-45ca-98b4-f21f0ac7a269 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Blending MPC & Value Function Approximation for Efficient Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4c7ba73-8e9e-4eec-9d31-3881148487f2 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Markov decision processes with delays and asynchronous cost collection,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 55e75dce-bca5-4cea-a543-0a9077aa5a68 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Delay-aware model-based reinforcement learning for continuous control,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 737389b3-d76b-49b6-bdfd-91c40f7c0d0d · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Reinforcement Learning with Random Delays
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6f0662c-5921-48e1-939e-f5f0d224bcf5 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Real-time reinforcement learning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9765e955-04fd-46fc-b5c1-b2eec546fc1c · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Thinking While Moving: Deep Reinforcement Learning with Concurrent Control
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c94fc2-3f69-4c1a-996f-3e05f072397d · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Asynchronous reinforcement learning for real-time control of physical robots,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7a2b46ba-fbf4-420c-bf80-95161b1c80c2 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Enabling Realtime Reinforcement Learning at Scale with Staggered Asynchronous Inference
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b97198-a9d4-4c58-863e-6e31743cf741 · outbound
RT-HCP: Dealing with Inference Delays and Sample Efficiency to Learn Directly on Robotic Platforms Addressing function approxi- mation error in actor-critic methods,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.