Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T05:19:09.239168Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:1909.01500.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T05:19:09.239168Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T05:22:09.716785Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T13:25:20.593899Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee727af1-0c2d-4cf6-8fd0-270121f07507 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Playing Atari with Deep Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c79b4e0-7a25-4270-98e1-6f6823e0cc88 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Trust region policy optimization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40dfdb77-ff30-4630-9d08-5721f4fe8af0 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Automatic differentiation in PyTorch
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce29919-2576-4037-89bd-dfdadb2617f1 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Alphastar
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7341aaff-8341-4d70-a4f8-7a0501821e62 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Openai five
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 532c3c91-0ed4-417b-8fda-182bfec2589a · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Recurrent experience replay in distributed reinforcement learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 05be1ce0-186d-4ea7-8abd-cdabb46eced9 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Openai gym, 2016
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91b2453c-1b24-42e0-bb52-358795a52a8f · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Asynchronous methods for deep reinforcement learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 433f9895-e312-4791-b09f-f74779a9cc46 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Proximal Policy Optimization Algorithms
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31e71905-a7b0-4fc4-ad7d-06ce1bd82dec · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Deep reinforcement learning with double q-learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ac7a8f62-c5ea-4d1e-9fe4-c26c5c789f60 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Dueling Network Architectures for Deep Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8067a7a9-2fed-4f15-b86b-375b2c35d75b · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch A distributional perspective on reinforce- ment learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2cb9aac9-b65d-4d20-a1f5-711632c8cb6d · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Rainbow: Combining improvements in deep reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20c04ee0-549f-4d4d-93a7-531168ba3ab4 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Distributed Prioritized Experience Replay
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8e12ed6-839e-455c-a176-fc0925a7dd36 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Implicit Quantile Networks for Distributional Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db2b5157-2bbd-472a-a24d-654f0c7d5180 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Continuous deep q-learning with model-based acceleration
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9753ee44-b572-4930-9919-d3b185e347d4 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Addressing Function Approximation Error in Actor-Critic Methods
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42ad99cc-37ac-44f4-808c-feb29af19c4b · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd02575e-29ed-425d-a429-d737560c5d33 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Soft Actor-Critic Algorithms and Applications
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8413eae-e346-4bcd-8e46-3e269d1fa5cb · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Distributed Distributional Deterministic Policy Gradients
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bc417dd-f7dc-4eaa-9d4a-58301850b34d · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Prioritized Experience Replay
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8db25fa9-0c94-4297-919e-579686a8bd8e · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch The arcade learning environment: An evaluation platform for general agents
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98c0d123-7b6c-40aa-98f9-fa478a6dd09d · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Mujoco: A physics engine for model-based control
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b9a4144-fef1-4762-9cb8-74f30c148890 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Accelerated Methods for Deep Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4f22f19d-c9e5-4953-802e-27a8fea1562f · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Openai spinning up
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 89e8521c-c76f-4522-a235-5b25ad259f33 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Theano: new features and speed improvements
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608edfe3-7719-4b08-bf07-bfc6426c76be · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch An Empirical Model of Large-Batch Training
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ebe22d1-d061-47ae-9489-d88134d99935 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Benchmarking deep reinforcement learning for continuous control
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0de24811-f854-4f53-8642-34621f9224a2 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Openai baselines
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 79b021e3-4ea5-4886-b20f-e5b9d1ff7f70 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Dopamine: A Research Framework for Deep Reinforcement Learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0761670-d21a-4574-b7d2-54a1d3ef2a3c · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Tensorflow: A system for large-scale machine learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 723cded4-26b8-4ade-a27c-b71d298e3f7f · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch RLlib: Abstractions for Distributed Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19b4a751-53ef-4531-8ffc-363a2908f03a · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Ray: A distributed framework for emerging{AI} applications
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3f75761d-a2b1-46fb-8592-5ef783c64fe5 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Horizon: Facebook's Open Source Applied Reinforcement Learning Platform
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21dae2e6-90c6-4ae4-b291-86cce34f19dd · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch Hogwild: A lock-free approach to parallelizing stochastic gradient descent
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0ce2a332-97e6-4e8b-a248-e6440d32b5c1 · outbound
rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch cuDNN: Efficient Primitives for Deep Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5130b66b-b7bf-41fc-a2e8-19851e9801e3 · inbound
Towards Fault Tolerance in Multi-Agent Reinforcement Learning rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0bd5300-83ce-46a4-98d1-076dbd8826fe · inbound
Tilted Quantile Gradient Updates for Quantile-Constrained Reinforcement Learning rlpyt: A Research Code Base for Deep Reinforcement Learning in PyTorch
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.