Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T10:18:25.102620Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:1908.11494.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T10:18:25.102620Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c30b289a-d30d-4f23-90a4-f5fd0b49f44a · outbound
Reinforcement learning with world model Addressing Function Approximation Error in Actor-Critic Methods
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3e4474c-660f-4f0b-85ea-146e0a1186e2 · outbound
Reinforcement learning with world model Q-Prop: Sample-Efficient Policy Gradient with An Off-Policy Critic
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e15af6-b280-4c48-be43-25f9573e870a · outbound
Reinforcement learning with world model Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 23a28fc0-1fc6-4a87-9258-82a4d2bd732f · outbound
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7f11a55-fe57-4f24-9b4c-7ce59a8e92fd · outbound
Reinforcement learning with world model Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0741495-2176-46c1-8476-ce40822600bf · outbound
Reinforcement learning with world model Soft Actor-Critic Algorithms and Applications
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f92161fa-1217-4924-bfd6-bd973296a385 · outbound
Reinforcement learning with world model Learning Latent Dynamics for Planning from Pixels
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 400f5975-43d4-4c70-a986-0fa70c7adccd · outbound
Reinforcement learning with world model and Stone, P
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5e2bdabb-1414-4b60-b7d0-3b6b0dcb1eed · outbound
Reinforcement learning with world model Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8d8b4e75-6295-4be8-a79b-7c8ccbb5dde1 · outbound
Reinforcement learning with world model Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6be48e78-51bf-4234-b2ea-83d42ce6f826 · outbound
Reinforcement learning with world model Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7f808845-d084-4410-bb73-801edff98fac · outbound
Reinforcement learning with world model Adam: A Method for Stochastic Optimization
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcd3d62a-93b8-4da9-a63e-3559a09c515c · outbound
Reinforcement learning with world model Model-Ensemble Trust-Region Policy Optimization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5af25c26-5953-430a-96cc-44a20346927a · outbound
Reinforcement learning with world model Continuous control with deep reinforcement learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30f91050-9507-46ec-8060-743c4eeca76e · outbound
Reinforcement learning with world model Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 54828f07-5e14-408f-913b-84ea4fc2d222 · outbound
Reinforcement learning with world model Algorithmic Framework for Model-based Deep Reinforcement Learning with Theoretical Guarantees
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e1a90cd-f5f9-417c-8d19-bb0964f8bb04 · outbound
Reinforcement learning with world model A., Veness, J., Bellemare, M
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 14b9fd5d-a4e2-40bc-88f3-8ba3642272af · outbound
Reinforcement learning with world model Proximal Policy Optimization Algorithms
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4396dc90-809d-464f-a702-f66f82188335 · outbound
Reinforcement learning with world model R., Yang, C., McGreavy, C., and Li, Z
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4c6054ca-8ea3-400c-a716-afae2bed813d · outbound
Reinforcement learning with world model Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc5025e0-db4b-4a08-9b1d-e20cacb75f12 · outbound
Reinforcement learning with world model Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
No inbound Pith citation observations are available.