Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2403.00514.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:58:51.075651Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-18T13:01:23.635839Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 19548d37-4bec-43bc-af16-66f226479d82 · inbound
Bigger, Regularized, Categorical: High-Capacity Value Functions are Efficient Multi-Task Learners Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fd96c6b-db8e-4849-a3f8-57a82ec57793 · inbound
Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77e2d3d5-0248-4d0c-98b1-276b66674e6a · inbound
Ensemble Elastic DQN: A Step Dependent Ensemble Approach for Reducing Overestimation in Deep Value-Based Reinforcement Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e48555d3-1a4f-4afc-9fb0-98007b4db202 · inbound
A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7e28391-58e3-45e9-8de3-f0ae950b80ef · inbound
On the Effect of Regularization in Policy Mirror Descent Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b7ccadc-2df1-43f7-987f-f5286c24b362 · inbound
Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fc359b0-f3cd-4b89-b0ec-b6118a569801 · inbound
Activation Function Design Sustains Plasticity in Continual Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 90130f14-8157-4070-aee3-50ebfdb0a4d1 · inbound
Forager: a lightweight testbed for continual learning with partial observability in RL Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d6fde913-fe6b-47d7-b25d-6c1a436a441f · inbound
Memory Merge DQN: Sensitivity Weighted Target Updates for Stable Value Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6af8d223-f4fb-493e-b5ab-f11e853aa1d2 · inbound
Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.