Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2407.04811.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:23:16.184323Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 1cbf3a39-31ac-4f64-b5ea-4588a997c971 · inbound
Plasticity Loss in Deep Reinforcement Learning: A Survey Simplifying Deep Temporal Difference Learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f23f055e-cb93-49fb-bd93-46373db78d0a · inbound
Hadamax Encoding: Elevating Performance in Model-Free Atari Simplifying Deep Temporal Difference Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9805582-53ad-478a-984d-dfa78092c0f0 · inbound
Universal Value-Function Uncertainties Simplifying Deep Temporal Difference Learning
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3349cc25-8617-4799-b1a8-445fb5b43491 · inbound
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control Simplifying Deep Temporal Difference Learning
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5e31804-8085-4696-9cd2-b998016c275e · inbound
The Impact of On-Policy Parallelized Data Collection on Deep Reinforcement Learning Networks Simplifying Deep Temporal Difference Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 625ffa4f-d64e-42a5-9af4-316d21545853 · inbound
Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models Simplifying Deep Temporal Difference Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab8b6e32-c3dd-4343-aa03-25fb4bda2d62 · inbound
Benchmarking Massively Parallelized Multi-Task Reinforcement Learning for Robotics Tasks Simplifying Deep Temporal Difference Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a5445b3-83e8-45f8-bdab-b98429e1a043 · inbound
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies Simplifying Deep Temporal Difference Learning
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c61f33e-1cc3-470c-b869-fd968e9c35dc · inbound
Priors Matter: Addressing Misspecification in Bayesian Deep Q-Learning Simplifying Deep Temporal Difference Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1877c9b2-7022-41d7-9ccd-9958b2bb4981 · inbound
TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning Simplifying Deep Temporal Difference Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 18f36599-0f5f-4134-bf72-988583957402 · inbound
TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning Simplifying Deep Temporal Difference Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 700bd6c9-d87a-4d83-b166-9073f9a2d975 · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control Simplifying Deep Temporal Difference Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 66d0fc71-81f6-4aee-b9ff-a19f156460cf · inbound
FlashSAC: Fast and Stable Off-Policy Reinforcement Learning for High-Dimensional Robot Control Simplifying Deep Temporal Difference Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 99db9eed-557d-4a49-9d21-11518d8db892 · inbound
A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations Simplifying Deep Temporal Difference Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a98f02b8-939f-483e-a53d-b0622f339fc3 · inbound
Scalable Reinforcement Learning via Adaptive Batch Scaling Simplifying Deep Temporal Difference Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7fb16431-0e58-43f7-b718-674279663c0f · inbound
Scalable Reinforcement Learning via Adaptive Batch Scaling Simplifying Deep Temporal Difference Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7d2bc9c-9f63-4cf8-a1f9-e1d79c416930 · inbound
Goal-Conditioned Agents that Learn Everything All at Once Simplifying Deep Temporal Difference Learning
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 610deba9-2a2d-4ea0-baf2-60ece410dbf5 · inbound
Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion Simplifying Deep Temporal Difference Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55d89c66-9695-4d48-b86b-6f1e692f7511 · inbound
Task diversity produces systematic transfer but inhibits continual reinforcement learning Simplifying Deep Temporal Difference Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8fbaf87d-acd4-4039-b25b-b2f5d992d2c6 · inbound
Trace-Mediated Peak Bias: Bridging Temporal Credit Assignment and Cognitive Heuristics in Deep Reinforcement Learning Simplifying Deep Temporal Difference Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03362496-b669-41a4-b8d4-ec3f5ad205a7 · inbound
Modularized Reinforcement Learning on LLMs: From MDP Creation to Exploration and Learning Simplifying Deep Temporal Difference Learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3aef960c-934a-4938-8528-e304a4dd1541 · inbound
Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning Simplifying Deep Temporal Difference Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.