Pith. sign in

Paper Citation Record · LEDGER

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning

As of 20 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2509.04815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04815 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:32:12.820798Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact7
  • verified fuzzy4
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0c2071b-d611-486c-ba19-c2a27d812a64 · outbound

This paper cites A Definition of Continual Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning A Definition of Continual Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.684556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.684556Z digest=sha256:bf24e5932f78f2e667b366366068ec3b19345999cc4b3acd0691e1e6e8f860a4

Observation 80a059ca-65cb-4687-9ac0-1fbd292a5d79 · outbound

This paper cites An Optimistic Perspective on Offline Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning An Optimistic Perspective on Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.689193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.689193Z digest=sha256:5438c5c000e7d24d97e912acbec12b2000f49bbe32ea41491324bb09a4e6646e

Observation 90e46a99-ea99-48ff-8cef-637c74e0e4ac · outbound

This paper cites Averaged-DQN: Variance Reduction and Stabilization for Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Averaged-DQN: Variance Reduction and Stabilization for Deep Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.405978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.694140Z digest=sha256:8c2edb3cf0e4517dde5e41fe215c106db6597e81be830a3c2a0b73a7761f60b3

Observation f32fefd8-8e9a-4eef-9c90-86c48d4885bc · outbound

This paper cites A Distributional Perspective on Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning A Distributional Perspective on Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.699176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.699176Z digest=sha256:b0ac49904c12e1c723c661cfad165c10d040ca66d8c407259582a208730a5111

Observation 848c583f-d52d-4bd8-95d9-0b2147424300 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-15T16:32:13.377658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.703447Z digest=sha256:0e90eadb3833684892d5e1f182999b2884aebc66b74a0e6897ca501a56a59e46

Observation 6b801350-c177-4d70-99d9-5e538633deab · outbound

This paper cites Noisy Networks for Exploration.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Noisy Networks for Exploration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.707519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.707519Z digest=sha256:213f8e88a17ebe5dbcfb9090f59590f3147e20a5612d67e19dbbd608287ad4b2

Observation 0497429f-2440-4e08-9fe0-dde679a9c566 · outbound

This paper cites Gl \"a scher, R.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Gl \"a scher, R

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.510565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.711758Z digest=sha256:373949f2eb649e6513a181fd631f39d7a7917a15de1947c9ccc2c19020fef1ef

Observation 25b4b8a6-4d4a-4b82-b10f-adb1da5db636 · outbound

This paper cites Deep Reinforcement Learning with Double Q-learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Deep Reinforcement Learning with Double Q-learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.716344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.716344Z digest=sha256:aa3bbd0a8202711e19699642602ad76a272f5a9b30c0387bf51823c170f9882b

Observation 75a286e1-bbd8-4466-9186-c368a88e452b · outbound

This paper cites Rainbow: Combining Improvements in Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Rainbow: Combining Improvements in Deep Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.719996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.719996Z digest=sha256:7f4a7306f9bfa71ae9bb7b7154e155554d04e99b66aaf4db5032221147df2cf6

Observation 87527406-39a9-4239-b971-1dcfc16803aa · outbound

This paper cites Randomized Exploration for Reinforcement Learning with General Value Function Approximation.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Randomized Exploration for Reinforcement Learning with General Value Function Approximation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.723333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.723333Z digest=sha256:2768af0d9b2857611cfe37808137408a209f366868898f625c7785b4902d29f3

Observation 2c184e29-7800-4ad7-ad7f-687df1c37bf4 · outbound

This paper cites Continuous Control With Ensemble Deep Deterministic Policy Gradients.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Continuous Control With Ensemble Deep Deterministic Policy Gradients

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.281562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.727288Z digest=sha256:aeb47fb7e51a58e789fa4cbca9ca10626292cf0ed38f831f2ee12ee70bdfe976

Observation deb8b41d-badc-49a6-bca1-64bc014b8cf4 · outbound

This paper cites Towards Continual Reinforcement Learning: A Review and Perspectives.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Towards Continual Reinforcement Learning: A Review and Perspectives

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.731129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.731129Z digest=sha256:04443d4f620cdade6103a163a4380cc37872bf36fe5b625d538886ee7ce0f8c2

Observation 65b9a3a4-ee08-402e-928f-ded427b2e4de · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 13

Resolution
verified exact
doi, observed 2026-08-15T16:32:12.869175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.735687Z digest=sha256:b2c8639421b8e6e1a5acc94023d054d52df521c8b7a0e46e921821705e72bb24

Observation 93f97fba-314c-4ad2-8ad4-4ca83999cfaa · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.739877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.739877Z digest=sha256:d4e971985de4a1e760d408861595849f9c97470ffc7f0f340d2724acc7bc8572

Observation c53a9a82-5ac5-4b06-8001-4270af2f6229 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.498149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.743904Z digest=sha256:a2839b6269bc68c141cdab0421d3d4c3a961df41710e91829b3b6dd110a07372

Observation 91cb9cb0-4b16-48f0-8860-5945bf7a783a · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 16

Resolution
verified exact
doi, observed 2026-08-15T16:32:12.856456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.748012Z digest=sha256:87e4d9d935601f7a3ac528ad128a5259cfc8e09d82431d8b95386bd506c57f56

Observation e184c1e9-53af-407e-9a3f-e3bf662a162a · outbound

This paper cites The Curse of Diversity in Ensemble-Based Exploration.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning The Curse of Diversity in Ensemble-Based Exploration

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.751332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.751332Z digest=sha256:06925b76a2c96cdcad96493ebdbb467181945fe2d5927dfdb009d95e159bf84e

Observation 8e3524b2-8c83-4051-b794-09d1853a9c05 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.755668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.755668Z digest=sha256:0fd9c61f0a06b78adbbeec9b60fde7ee2840807883745902efeb0937e2969362

Observation e612f9b1-1747-42e1-b835-b2b421f4953e · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.486525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.760131Z digest=sha256:c46a96c03f5a821d6c0da93696b2d4f792cb1ce7616553e168e3695b9fcb54fd

Observation 27d31070-631b-4776-91fc-b457c6fc0f66 · outbound

This paper cites UCB Exploration via Q-Ensembles.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning UCB Exploration via Q-Ensembles

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.763857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.763857Z digest=sha256:33a93b7aada8288f2c402dac9b11171ca89cc5e56b5ad5523c807c49caeceab6

Observation 79af5860-86d4-4cce-9832-826cc6439f45 · outbound

This paper cites Deep Exploration via Bootstrapped DQN.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Deep Exploration via Bootstrapped DQN

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.768406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.768406Z digest=sha256:433371ea40c1134962156a4b6eb2d1475aee6e67cfa6d6f879a67b9f540e6b79

Observation cdb9adfe-f792-4810-b19f-ce2a08c2ef8b · outbound

This paper cites Osband, C.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Osband, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.476499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.772364Z digest=sha256:780dcb86cb92634361f86ec02e4a942eae4171f7a4d271bacbbdb6cd8f74bda6

Observation 3fd6ba6b-8c1c-42de-b810-e5afae33737e · outbound

This paper cites Count-Based Exploration with Neural Density Models.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Count-Based Exploration with Neural Density Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.776053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.776053Z digest=sha256:a2bc368eb9aced67d8a2db6fbb6ef945a8099ca66e2c117a5ea3578f8447d3c2

Observation 33d0c9db-9452-4177-9541-c813feeafd27 · outbound

This paper cites Ensemble Bootstrapping for Q-Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Ensemble Bootstrapping for Q-Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.146108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.779849Z digest=sha256:62f79b609cd55472e0e11d2ce9d125dd31e4ad00cd37d04c5a34e16a74d5a529

Observation 9a8cba07-b7ca-480a-8ceb-294836cf3a5d · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 25

Resolution
verified exact
raw_fallback, observed 2026-08-15T16:32:13.129586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.783685Z digest=sha256:e7b71ce77d91408390ea88c30b37c6fc129d9554d803558619ba1f0c6063eb25

Observation c033b81b-d18c-47ed-98a8-d0f729813f39 · outbound

This paper cites Saphal, B.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Saphal, B

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.787308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.787308Z digest=sha256:1bfaeb5b6175771f76311f6f019364877e74bead33717875149af959cafe7f93

Observation 78d58b98-48b6-4985-b088-636018bd5459 · outbound

This paper cites Prioritized Experience Replay.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Prioritized Experience Replay

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.791860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.791860Z digest=sha256:bb6d2c93f122b4e8334fad1563edf33beb11604bf04fbcb76785037554149bd9

Observation 783cf942-8e64-4c30-a71c-22dad2b9379a · outbound

This paper cites Off-Policy Actor-Critic with Shared Experience Replay.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Off-Policy Actor-Critic with Shared Experience Replay

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.795306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.795306Z digest=sha256:48f90069745d3e3500d106568869cc4e9434f3bb95c4b961210e2122e15b268a

Observation 612ace01-1374-48fe-bb94-edbd8ac504dd · outbound

This paper cites Sutton and A.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Sutton and A

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.798809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.798809Z digest=sha256:00c057f0e563d07c797ec693f97621886a34923689669622aa9ce46cd27cdc49

Observation 25e81015-7609-4449-b38b-fa4462087582 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.802389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.802389Z digest=sha256:0d3efe8363d5dca4786e11ab85c7bbbc81c96391baa309b713029a689f6bf7eb

Observation 57752963-fb7b-4269-9b23-70104442ebb6 · outbound

This paper cites Thangarajah, F.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Thangarajah, F

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.466235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.806168Z digest=sha256:e4728a174493ed9250785bf430d231c8fc80cc1f606b4d3efa2b2d97ed0bae54

Observation 81f2a908-a3b6-4d8f-8245-6725c44f2d91 · outbound

This paper cites Continual Learning and Catastrophic Forgetting.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Continual Learning and Catastrophic Forgetting

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.809530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.809530Z digest=sha256:668ca99596c41c446ac731f9c0f43c1339c9775a567ed78b9169c7e2f33bbba3

Observation d0346254-a34c-4da0-ad93-a5f8ee50d6cb · outbound

This paper cites Dueling Network Architectures for Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Dueling Network Architectures for Deep Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.813302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.813302Z digest=sha256:1ff6800594a6946a87a81496a3ed20b573dcdc7ea21b3da34038aa2bf2a73eb5

Observation 0b42bfee-1b8c-48de-8a7f-6addf8c816cb · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.454855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.817215Z digest=sha256:cb14411f9829bbfa93bc7e72188491f5eb27a289d101da94218df155df07820c

Observation 11e498c3-0e35-4e39-88dc-4f137e6fa734 · outbound

This paper cites Borkar and S.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Borkar and S

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.442717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.820798Z digest=sha256:26c70518ebd6a870cf3f86518a3d65871d7de80ed872edb0762b0cbd0d958dc6

Pith citing papers

No inbound Pith citation observations are available.