Pith. sign in

Paper Citation Record · LEDGER

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2509.04815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04815 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:32:12.820798Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

35 of 35 outbound references displayed

  • verified exact7
  • verified fuzzy4
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0c2071b-d611-486c-ba19-c2a27d812a64 · outbound

This paper cites A Definition of Continual Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning A Definition of Continual Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.684556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.684556Z digest=sha256:cc94d6be5c783fe571c5393f6d2ac245394c8ccd7143dfe184e9ac0a5711d117

Observation 80a059ca-65cb-4687-9ac0-1fbd292a5d79 · outbound

This paper cites An Optimistic Perspective on Offline Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning An Optimistic Perspective on Offline Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.689193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.689193Z digest=sha256:edb0b4f0058c8f5090ccd254f3da66f63c471636251a869948526d0d37c14032

Observation 90e46a99-ea99-48ff-8cef-637c74e0e4ac · outbound

This paper cites Averaged-DQN: Variance Reduction and Stabilization for Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Averaged-DQN: Variance Reduction and Stabilization for Deep Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.405978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.694140Z digest=sha256:8e91408480d2670d2993280860387547a10fb2109babfaf52fd84b36ae2b7a90

Observation f32fefd8-8e9a-4eef-9c90-86c48d4885bc · outbound

This paper cites A Distributional Perspective on Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning A Distributional Perspective on Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.699176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.699176Z digest=sha256:b0ac49904c12e1c723c661cfad165c10d040ca66d8c407259582a208730a5111

Observation 848c583f-d52d-4bd8-95d9-0b2147424300 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 5

Resolution
verified exact
raw_fallback, observed 2026-08-15T16:32:13.377658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.703447Z digest=sha256:4fea9301f32120f7e3914b10e9c29fc09acccb29fd0cca0bf9f12c56cc180dc9

Observation 6b801350-c177-4d70-99d9-5e538633deab · outbound

This paper cites Noisy Networks for Exploration.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Noisy Networks for Exploration

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.707519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.707519Z digest=sha256:213f8e88a17ebe5dbcfb9090f59590f3147e20a5612d67e19dbbd608287ad4b2

Observation 0497429f-2440-4e08-9fe0-dde679a9c566 · outbound

This paper cites Gl \"a scher, R.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Gl \"a scher, R

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.510565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.711758Z digest=sha256:90aa09fb9ddf9003e65d6b8d70e842bf1e48abcec74cf4218ac015f72c895600

Observation 25b4b8a6-4d4a-4b82-b10f-adb1da5db636 · outbound

This paper cites Deep Reinforcement Learning with Double Q-learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Deep Reinforcement Learning with Double Q-learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.716344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.716344Z digest=sha256:4d7208bed6c61af0788afab76ba78c2df7785efb5e2dadcf2288ddf9b476fe48

Observation 75a286e1-bbd8-4466-9186-c368a88e452b · outbound

This paper cites Rainbow: Combining Improvements in Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Rainbow: Combining Improvements in Deep Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.719996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.719996Z digest=sha256:7f4a7306f9bfa71ae9bb7b7154e155554d04e99b66aaf4db5032221147df2cf6

Observation 87527406-39a9-4239-b971-1dcfc16803aa · outbound

This paper cites Randomized Exploration for Reinforcement Learning with General Value Function Approximation.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Randomized Exploration for Reinforcement Learning with General Value Function Approximation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.723333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.723333Z digest=sha256:7570dd9977a2729e832e1cf6741404a95da3f1041daa60c39777deb1fe828238

Observation 2c184e29-7800-4ad7-ad7f-687df1c37bf4 · outbound

This paper cites Continuous Control With Ensemble Deep Deterministic Policy Gradients.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Continuous Control With Ensemble Deep Deterministic Policy Gradients

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.281562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.727288Z digest=sha256:2cac4bd11b90acdb4533a4489cea594e64b713eb5e6d10c3c0d69d8bca56f8c6

Observation deb8b41d-badc-49a6-bca1-64bc014b8cf4 · outbound

This paper cites Towards Continual Reinforcement Learning: A Review and Perspectives.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Towards Continual Reinforcement Learning: A Review and Perspectives

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.731129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.731129Z digest=sha256:2256291c0a78524b420dab28e28dc97313a8f6e1bcd14d5113807d7e113a7ede

Observation 65b9a3a4-ee08-402e-928f-ded427b2e4de · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 13

Resolution
verified exact
doi, observed 2026-08-15T16:32:12.869175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.735687Z digest=sha256:deff5999e067d38c84d4ac9e62f1507aac50d2aa5eab992101a727b211eb2994

Observation 93f97fba-314c-4ad2-8ad4-4ca83999cfaa · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.739877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.739877Z digest=sha256:d4e971985de4a1e760d408861595849f9c97470ffc7f0f340d2724acc7bc8572

Observation c53a9a82-5ac5-4b06-8001-4270af2f6229 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.498149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.743904Z digest=sha256:055fcb1ec33b31295c8b2e68e44237319465263afe92dae6ab716dc70e81b8dd

Observation 91cb9cb0-4b16-48f0-8860-5945bf7a783a · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 16

Resolution
verified exact
doi, observed 2026-08-15T16:32:12.856456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.748012Z digest=sha256:12ffdf7d1d5dd5677b8557339d05fc2156b803b1fc06528212e90d651f99bb0a

Observation e184c1e9-53af-407e-9a3f-e3bf662a162a · outbound

This paper cites The Curse of Diversity in Ensemble-Based Exploration.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning The Curse of Diversity in Ensemble-Based Exploration

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.751332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.751332Z digest=sha256:d38b7b02a362d56f3f56663b9d6b869d1ce50f9931f94bdc3a897916f06f569a

Observation 8e3524b2-8c83-4051-b794-09d1853a9c05 · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Playing Atari with Deep Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.755668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.755668Z digest=sha256:cb33f32c5dd3d7bd4f0aa3b7d614993e18df1926b58f19956530f019cc543645

Observation e612f9b1-1747-42e1-b835-b2b421f4953e · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.486525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.760131Z digest=sha256:a5bbe9c35a17b767b31b9b76f07edec77bae3620a70939679ecaafa6539d99d0

Observation 27d31070-631b-4776-91fc-b457c6fc0f66 · outbound

This paper cites UCB Exploration via Q-Ensembles.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning UCB Exploration via Q-Ensembles

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.763857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.763857Z digest=sha256:33a93b7aada8288f2c402dac9b11171ca89cc5e56b5ad5523c807c49caeceab6

Observation 79af5860-86d4-4cce-9832-826cc6439f45 · outbound

This paper cites Deep Exploration via Bootstrapped DQN.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Deep Exploration via Bootstrapped DQN

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.768406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.768406Z digest=sha256:86d8111e87cb23f5d84f2c4d8ce8e227a9a321a94e89b7ddbabee28c91d239e9

Observation cdb9adfe-f792-4810-b19f-ce2a08c2ef8b · outbound

This paper cites Osband, C.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Osband, C

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.476499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.772364Z digest=sha256:9cfd839c983667fc1c6ef813c21d1cb47bb4dc845af2b0788ff14d451e299329

Observation 3fd6ba6b-8c1c-42de-b810-e5afae33737e · outbound

This paper cites Count-Based Exploration with Neural Density Models.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Count-Based Exploration with Neural Density Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.776053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.776053Z digest=sha256:08999c0b4533de9c2afce62e32a34de523d206e9b947e84842a1acd32b93444b

Observation 33d0c9db-9452-4177-9541-c813feeafd27 · outbound

This paper cites Ensemble Bootstrapping for Q-Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Ensemble Bootstrapping for Q-Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-15T16:32:13.146108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.779849Z digest=sha256:112b5cff492b06e7a93a895be41d16a5fe1e9be6385479f218f03f9af0245ba5

Observation 9a8cba07-b7ca-480a-8ceb-294836cf3a5d · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 25

Resolution
verified exact
raw_fallback, observed 2026-08-15T16:32:13.129586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.783685Z digest=sha256:a1550adc888b5bf1bd797fca0af6f546bdfe8b28681c2ca426728c16786eab6c

Observation c033b81b-d18c-47ed-98a8-d0f729813f39 · outbound

This paper cites Saphal, B.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Saphal, B

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.787308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.787308Z digest=sha256:1bfaeb5b6175771f76311f6f019364877e74bead33717875149af959cafe7f93

Observation 78d58b98-48b6-4985-b088-636018bd5459 · outbound

This paper cites Prioritized Experience Replay.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Prioritized Experience Replay

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.791860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.791860Z digest=sha256:bb6d2c93f122b4e8334fad1563edf33beb11604bf04fbcb76785037554149bd9

Observation 783cf942-8e64-4c30-a71c-22dad2b9379a · outbound

This paper cites Off-Policy Actor-Critic with Shared Experience Replay.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Off-Policy Actor-Critic with Shared Experience Replay

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.795306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.795306Z digest=sha256:42b76aeb4c3fe54779e8e9b99ac05332383820185c5800e8f028791d29905fbb

Observation 612ace01-1374-48fe-bb94-edbd8ac504dd · outbound

This paper cites Sutton and A.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Sutton and A

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.798809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.798809Z digest=sha256:00c057f0e563d07c797ec693f97621886a34923689669622aa9ce46cd27cdc49

Observation 25e81015-7609-4449-b38b-fa4462087582 · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.802389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.802389Z digest=sha256:0d3efe8363d5dca4786e11ab85c7bbbc81c96391baa309b713029a689f6bf7eb

Observation 57752963-fb7b-4269-9b23-70104442ebb6 · outbound

This paper cites Thangarajah, F.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Thangarajah, F

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.466235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.806168Z digest=sha256:d06ca4f46559827044a0143807b7a3b04acbe0f69b064abe023efa05423cfe99

Observation 81f2a908-a3b6-4d8f-8245-6725c44f2d91 · outbound

This paper cites Continual Learning and Catastrophic Forgetting.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Continual Learning and Catastrophic Forgetting

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.809530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.809530Z digest=sha256:bdcea9d85953b3ca27dd15065becaec1ade626fb982a2c9d6a296fdb025ffe61

Observation d0346254-a34c-4da0-ad93-a5f8ee50d6cb · outbound

This paper cites Dueling Network Architectures for Deep Reinforcement Learning.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Dueling Network Architectures for Deep Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T16:32:12.813302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:32:12.813302Z digest=sha256:09f5f5fb0074221c041b06c3f06116c172d72b90a34fc4367e1181db84880529

Observation 0b42bfee-1b8c-48de-8a7f-6addf8c816cb · outbound

This paper cites an unresolved cited work.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:32:13.454855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.817215Z digest=sha256:f10c34cfc0fef39be874454c0b4f1b382807355221a5a051e5928c8137b0d1f5

Observation 11e498c3-0e35-4e39-88dc-4f137e6fa734 · outbound

This paper cites Borkar and S.

An Arbitration Control for an Ensemble of Diversified DQN variants in Continual Reinforcement Learning Borkar and S

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:32:13.442717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T16:32:12.820798Z digest=sha256:e81fde514b6f88c9141a309a18753ebc5220f80e0ff0d014f623b050d0dc196e

Pith citing papers

No inbound Pith citation observations are available.