Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:50.823234Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2506.05716.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:18:50.823234Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 33b4becb-6c3c-484f-b044-5671c5f06bac · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Averaged-dqn: Variance reduction and stabilization for deep reinforcement learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bcf78f58-43f0-4ac9-b332-8160d5e4c52d · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Ensemble Reinforcement Learning in Continuous Spaces -- A Hierarchical Multi-Step Approach for Policy Training
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 14e3c160-e51c-4969-9fc3-440e199c68fb · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Mushroomrl: Simplifying reinforcement learning research
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b222622-40b9-4bb9-8135-53c5aa2cfb3b · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Addressing function approximation error in actor-critic methods
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b62061ba-c9dd-42b8-aaed-bf34cb00963d · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Understanding Multi-Step Deep Reinforcement Learning: A Systematic Study of the DQN Target
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 174abb56-8cf2-45ff-ad77-9f6507898be5 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Wd3: Taming the estimation bias in deep reinforcement learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2c2aaa1-ae24-4f13-9962-2d9aa4d8231a · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Rainbow: Combining improvements in deep reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfd774fa-c87d-4041-852c-1d176ad83b9c · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Bring color to deep q-networks: Limitations and improvements of dqn leading to rainbow dqn
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 46c982f7-5272-484e-986c-c405eddb4eb8 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Elastic step ddpg: Multi-step reinforcement learning for improved sample efficiency
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b808023-ad6d-4c8d-86c8-875a23d6a104 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Elastic step dqn: A novel multi-step algorithm to alleviate overestimation in deep q-networks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9086a3f5-1f5e-405f-a0a9-1042649faf05 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Maxmin Q-learning: Controlling the Estimation Bias of Q-learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78c45b44-0ade-4be2-9f43-d2709e2006af · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Human-level control through deep reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77e2d3d5-0248-4d0c-98b1-276b66674e6a · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f35cc21c-bff7-485b-ae2e-f58f3538ba08 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Revisiting Rainbow: Promoting more Insightful and Inclusive Deep Reinforcement Learning Research
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5634b6a-8ec5-4164-aa8e-5ebcbc7f9d8e · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Issues in using function approximation for reinforcement learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ba0a0d1d-eb7c-4e1d-bf74-8d37a9cbc8e3 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Deep Reinforcement Learning and the Deadly Triad
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 299c696d-4d85-414d-a6f9-e422f2d815ce · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Deep reinforcement learning with double q-learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 137e1cfe-eb05-4c37-9e30-8732445e98ce · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Controlling underestimation bias in reinforcement learning via quasi-median operation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb2d4639-c01a-4524-8ed4-ddb292c4ae87 · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning MinAtar: An Atari-Inspired Testbed for Thorough and Reproducible Reinforcement Learning Experiments
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6c0836e-6e67-4db9-82e9-9a314c87e66f · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning A novel multi-step q-learning method to improve data efficiency for deep reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d80f9e1b-7935-4a6e-b508-740b9734ffdb · outbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning beta-dqn: Improving deep q-learning by evolving the behavior
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.