Pith. sign in

Paper Citation Record · LEDGER

Deep Reinforcement Learning that Matters

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:1709.06560.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1709.06560 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T22:43:27.794855Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

366
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b4b9ea1a-2cf5-4ec2-a603-c88bd4c641f5 · inbound

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor cites this paper.

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor Deep Reinforcement Learning that Matters

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:48:10.739212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T01:48:10.689906Z digest=sha256:086cfefa97cdae789a8a1d41db8f8a97fd61993728134973c208187bd7902794

Observation 4d1bfb09-22c5-4c4d-9d21-488274632bc0 · inbound

Soft Actor-Critic Algorithms and Applications cites this paper.

Soft Actor-Critic Algorithms and Applications Deep Reinforcement Learning that Matters

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-13T14:32:42.027745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T14:32:41.948156Z digest=sha256:5e8816548d7bc3690a5bb70368701ad251de2f2986fd81523adcea60a2568e67

Observation ea5dab91-6196-4c2c-8df0-1b8cd61dfcf0 · inbound

Reproducibility in Machine Learning for Health cites this paper.

Reproducibility in Machine Learning for Health Deep Reinforcement Learning that Matters

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-25T11:05:40.127539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T11:03:15.737223Z digest=sha256:6be659f3cd236a8a072fed2b0e5341e01d2be5afe0ccb220b46754ef202f5ca9

Observation a9eeaa5e-8e8c-4253-8604-2d5bf0bf0ae8 · inbound

Enhancing Reinforcement Learning in 3D Environments through Semantic Segmentation: A Case Study in ViZDoom cites this paper.

Enhancing Reinforcement Learning in 3D Environments through Semantic Segmentation: A Case Study in ViZDoom Deep Reinforcement Learning that Matters

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T22:43:27.794855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T22:43:27.794855Z digest=sha256:03f879842c9c2df57ac2596aa2e74f5db3c96f25dfa46aaec96cd229499afea5

Observation 75750cd0-2c0d-46aa-8205-b3d9a8f037e5 · inbound

Regimes of Scale in AI Meteorology cites this paper.

Regimes of Scale in AI Meteorology Deep Reinforcement Learning that Matters

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-10T19:20:44.962186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:17:52.935347Z digest=sha256:bec884b5f732298b4f25cff6136db1450a23bb6a7fe2b82267b7295c00fcd1be

Observation 02d3754b-0679-4dd9-8385-0dfd5b4748dd · inbound

Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture cites this paper.

Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture Deep Reinforcement Learning that Matters

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:29:06.418008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T13:59:40.638085Z digest=sha256:894983c8d72afc8d4dbb4e423b534bf141102daea5327d4ba7af03ece108f8d4

Observation 42fd072f-03fe-4584-a842-ce4a5d844f3b · inbound

Temporal Reasoning Is Not the Bottleneck: A Probabilistic Inconsistency Framework for Neuro-Symbolic QA cites this paper.

Temporal Reasoning Is Not the Bottleneck: A Probabilistic Inconsistency Framework for Neuro-Symbolic QA Deep Reinforcement Learning that Matters

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:16:07.897798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T17:45:44.270122Z digest=sha256:136219ef4bbfda1a21c3487b8f3b9c230cad809855d57a364d5b94e241a6cef4

Observation e4590b28-abec-49b3-b43f-d28a15d96b44 · inbound

Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity cites this paper.

Rescaled Asynchronous SGD: Optimal Distributed Optimization under Data and System Heterogeneity Deep Reinforcement Learning that Matters

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-14T19:32:51.156278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-14T19:31:12.149482Z digest=sha256:8704480afe52c05676f064d70cb5aa8e83c4e5ff7e7893ff122bb97a1ed246f6

Observation c6e3f277-1583-475f-9ee6-c12680adcf98 · inbound

Some Essential Constructive Foundations for Systems and Control cites this paper.

Some Essential Constructive Foundations for Systems and Control Deep Reinforcement Learning that Matters

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T23:47:28.421001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T17:48:35.402903Z digest=sha256:bc9901acbacb865f2d740069014129162182b2f7f706f1bf9e0d36295b4289f8