Pith. sign in

Paper Citation Record · LEDGER

Generalization and Regularization in DQN

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:1810.00123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1810.00123 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:03:11.438031Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:09:52.653481Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5e08a654-056c-4ee5-807b-5b34b1eac042 · inbound

In Hindsight: A Smooth Reward for Steady Exploration cites this paper.

In Hindsight: A Smooth Reward for Steady Exploration Generalization and Regularization in DQN

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-25T17:41:05.708898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T17:38:45.913293Z digest=sha256:af1fbe4cc37c96fcedf1e3e0fa6a3ad27abbde08e12f22c5041ef5438ec55250

Observation a9fc6219-3c8c-4ac1-a151-91cc96beb823 · inbound

Generalizing from a few environments in safety-critical reinforcement learning cites this paper.

Generalizing from a few environments in safety-critical reinforcement learning Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-25T11:00:40.168151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T10:57:45.056957Z digest=sha256:fd4e4fbd70ea75fe3e77be21cf30beb908e4dfd1bc530d40a625fa206677e454

Observation 30d6457c-ad63-4e72-b30f-4c1d74fdf07d · inbound

Reasoning and Generalization in RL: A Tool Use Perspective cites this paper.

Reasoning and Generalization in RL: A Tool Use Perspective Generalization and Regularization in DQN

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-25T09:20:34.443759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T09:20:04.036102Z digest=sha256:4de1bdb9fc637269741f756d30f6db50496bb4a196251f7e55f32f10d8d81dc3

Observation ca5ac14d-9989-4a86-8d63-531a5dc3b7e5 · inbound

Scaling Laws for Reward Model Overoptimization cites this paper.

Scaling Laws for Reward Model Overoptimization Generalization and Regularization in DQN

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:04:53.245501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T09:04:53.129737Z digest=sha256:70edd3bba64d037795835ac4495f46a4df247a1d74c774a1d952885e52522693

Observation 6cded5cc-c288-4cbe-b46c-e7aeea35f7a0 · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey Generalization and Regularization in DQN

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:50.692503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:02ad8e8d0695fcbfcff344e457013356f6ea810d98d1a9f9788ad12f0d70dcdb

Observation 9a183f81-0082-471d-beb2-14c07b8694a8 · inbound

A Research Agenda for Usability and Generalisation in Reinforcement Learning cites this paper.

A Research Agenda for Usability and Generalisation in Reinforcement Learning Generalization and Regularization in DQN

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T05:58:46.294480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T05:58:46.294480Z digest=sha256:0819c0e3a409445bc8fe9c413dbbf0b63ad6d98691877ff29ad816974eeded9f

Observation 63b110f6-35da-4ff4-9f02-171cea97c91b · inbound

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning cites this paper.

Fisher-Guided Selective Forgetting: Mitigating The Primacy Bias in Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-09T17:46:07.253309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T17:46:07.253309Z digest=sha256:ece5ea9fad506a9d35cf05e8d32d6ca66570eb69957ca76e6a5ba33627dd8f08

Observation ec076800-faf5-4fb0-86b8-f4e561d1f171 · inbound

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing cites this paper.

LLM Bandit: Cost-Efficient LLM Generation via Preference-Conditioned Dynamic Routing Generalization and Regularization in DQN

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-09T11:22:28.868513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T11:22:28.868513Z digest=sha256:1dc9d11f08d5c40c325ac5784f022c922b52773e7a45dc57016d39600de58592

Observation 6247186c-16c4-4a41-83a9-9cb24acebc57 · inbound

Bidirectional Distillation: A Mixed-Play Framework for Multi-Agent Generalizable Behaviors cites this paper.

Bidirectional Distillation: A Mixed-Play Framework for Multi-Agent Generalizable Behaviors Generalization and Regularization in DQN

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:03:11.438031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:03:11.438031Z digest=sha256:cc82ccaae9adbed54792fd804a5a5e0c00f0277a92f25915723997fd0f0bbc64

Observation 73fc2854-b51f-4b46-831d-6332dc046698 · inbound

Zero-Shot Reinforcement Learning Under Partial Observability cites this paper.

Zero-Shot Reinforcement Learning Under Partial Observability Generalization and Regularization in DQN

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T19:37:29.607617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:37:29.607617Z digest=sha256:a52138ad50cee3e836c96206704ddcab08a8e9bdee3f2dbd6771b489970470cb

Observation 91e7bc31-a76d-4b71-9efb-2ad424663a00 · inbound

Hierarchical Successor Representation for Robust Transfer cites this paper.

Hierarchical Successor Representation for Robust Transfer Generalization and Regularization in DQN

Reference 2003

Resolution
unresolved
no resolver link, observed 2026-08-02T23:46:56.405225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:46:56.405225Z digest=sha256:7f852347fad373ddf58c00c8c85aa3a2d39031d2b24970eb4df7f90de8a8157b

Observation 13e803c1-d8d9-416a-8132-ae5706b5c6dd · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.720409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:c64526283ea61993939ea0f10e3374409bfe2ec5cb87cbcce333061793d47874

Observation ab828019-ff6e-41fd-8278-6b2e670873be · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Generalization and Regularization in DQN

Reference 225

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.822838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:ac5bc07f45556f0b84fe85e4db9bfb68431b2b409e37f0aef7ed98d7f69c92d3

Observation 0d9ae354-4fd0-407f-88ab-d791e2642e16 · inbound

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight cites this paper.

Bridging Performance and Generalization in Reinforcement Learning for Agile Flight Generalization and Regularization in DQN

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:52.655148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T04:33:20.981811Z digest=sha256:cb4def96281f7ed5ad8d7b0af4dda98648208b9d109ce188254c27d16f4ba9db

Observation 2a199256-a862-47db-81e5-fb823b2538c1 · inbound

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control cites this paper.

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control Generalization and Regularization in DQN

Reference 236

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:46.569656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:46.569656Z digest=sha256:e10cc5cb4012679022723a858ec88bba410d9a9122ff8bdad9eccaa9a699e957