Pith. sign in

Paper Citation Record · LEDGER

ADDQ: Adaptive Distributional Double Q-Learning

As of 22 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2506.19478.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.19478 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:14:03.900092Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6b58787e-adc2-4345-9442-380a38372e8a · outbound

This paper cites N(µ 2, σ2 2)) distributed rewards.

ADDQ: Adaptive Distributional Double Q-Learning N(µ 2, σ2 2)) distributed rewards

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:07.274394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:02.256253Z digest=sha256:6414aae4ce46c9761f793180c44b69e4665921c414888af011514f3f7c3cecc3

Observation 1f0873cc-b2c5-46f5-b637-ed0c559a9042 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:14:07.024193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:02.523971Z digest=sha256:0b6a71fed5634df9bfe9b252f0ae6f1bfaa9668b65fa2eb1ed42cfa462dec946

Observation 57272fa8-0012-472b-9465-b101f58a8957 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.200827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:02.945633Z digest=sha256:ea2d14195f933ddc84a028229880fc5fb397573c5405b2cc23b7b9255662790a

Observation 527420bf-ee6c-4d98-99c0-87c20539fcf1 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:05.933292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.071047Z digest=sha256:7ad423d5dde29bbb56889cb7d800a33a3b91db55965840be3c6d4c8a70412c50

Observation 57d6f0a1-8c3c-44dd-b9f7-6b49309e41b9 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.735255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:02.711844Z digest=sha256:2ff607345fdc667b83d878772a642b1181c4539b8e29f3ef43008f6ffc9ce0cb

Observation c71b2eca-c759-432f-93cc-16b2c4d4ca6f · outbound

This paper cites , dsuch that Fn is Fn+1-measurable and for alli= 1,.

ADDQ: Adaptive Distributional Double Q-Learning , dsuch that Fn is Fn+1-measurable and for alli= 1,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:05.616463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.327831Z digest=sha256:05c56be91616f129956fab5d913ea0cce176db444006b473c7287b27910946ee

Observation ab87d9cf-f490-4e43-b9fc-27c55dc9a23a · outbound

This paper cites , dis adapted with ∞X n=1 αi,n =∞and ∞X n=1 α2 i,n <∞a.s.

ADDQ: Adaptive Distributional Double Q-Learning , dis adapted with ∞X n=1 αi,n =∞and ∞X n=1 α2 i,n <∞a.s

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:05.335693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.420474Z digest=sha256:c98921ef8b7f3ccc995fde4a97eae8a0b972753b930124826c0c4a6a4fdd7f8e

Observation 1409acd7-b18b-4d17-89d2-c1fcbf0d2329 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:05.066511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.504810Z digest=sha256:d31a89d6badf6867e608cffb408e6887b2de7d9d82716fdff2cb2e770d4c2168

Observation 5127464a-b216-45f1-9214-9bdef6185004 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.430275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.663227Z digest=sha256:01554c58c551fa77264d468f1e2d6e45436eae51475d43222a9e4a5e078756c0

Observation 024e4a3c-9eb5-4061-97b8-93d697c5b838 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:04.768990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.779183Z digest=sha256:68add84eec77228b56cfc382d77afef9bbcb0f938f2438ba339309770fc2e40a

Observation b2b31dbe-d416-43cc-9c25-6f18bc953f76 · outbound

This paper cites Proof of Thoerem F .2.Step 1: Convergence of mean values toQ ∗ The proof mainly follows (Rowland et al., 2018) and (van Hasselt, 2010).

ADDQ: Adaptive Distributional Double Q-Learning Proof of Thoerem F .2.Step 1: Convergence of mean values toQ ∗ The proof mainly follows (Rowland et al., 2018) and (van Hasselt, 2010)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:04.419404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:03.900092Z digest=sha256:9617b24c35435793bd642fcbbcb08bb14a4a4988172be35553f422109bcc4dc1

Observation 28ac423a-4283-43a5-aae1-2130bf241611 · outbound

This paper cites Sudakov, V.

ADDQ: Adaptive Distributional Double Q-Learning Sudakov, V

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.783734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.783734Z digest=sha256:5d46e0273aaf285f21a745bd66bfe1e308cd0c6efefd009cb5a5d2fc9fb397a9

Observation fcf4708f-c7dd-42cc-9c50-92c62f4b8c95 · outbound

This paper cites Deep Reinforcement Learning with Double Q-learning.

ADDQ: Adaptive Distributional Double Q-Learning Deep Reinforcement Learning with Double Q-learning

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:02.096212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:02.096212Z digest=sha256:c70cbf28d82c5c917f82bb8774e747a164e2d072f0563e20f5b4371778146785

Observation 6d401963-943c-487c-829f-75d90edf13e9 · outbound

This paper cites Dopamine: A Research Framework for Deep Reinforcement Learning.

ADDQ: Adaptive Distributional Double Q-Learning Dopamine: A Research Framework for Deep Reinforcement Learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.517365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.517365Z digest=sha256:398b8118b752e472ea737272181da202932fce6eddd19731786ba919d152a83c

Observation 49fa3b95-c673-4957-a3a2-465b7f2aee64 · outbound

This paper cites Agent57: Outperforming the Atari Human Benchmark.

ADDQ: Adaptive Distributional Double Q-Learning Agent57: Outperforming the Atari Human Benchmark

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.364990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.364990Z digest=sha256:466d05fdcec9878a5ee9d10c9dd6e716f08a9ca8ef81869d824da2782de22bcc

Observation bdd6cba9-744b-4471-890a-08812cefc87d · outbound

This paper cites Thrun, S.

ADDQ: Adaptive Distributional Double Q-Learning Thrun, S

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.936257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.936257Z digest=sha256:257d08342428ab01b5759c29876ae9072242886db35b870ee540e08e50a75cdd

Observation 81d9263f-6718-4c68-9b87-850c84d9562f · outbound

This paper cites The Potential of the Return Distribution for Exploration in RL.

ADDQ: Adaptive Distributional Double Q-Learning The Potential of the Return Distribution for Exploration in RL

Reference 2019

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:14:04.164753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:01.663649Z digest=sha256:6467d190b3952f1fe4492e720325c500e3e39d5845c297e84268400b60100006

Observation e8a5cf44-19f7-402b-9f16-9bc316766de8 · outbound

This paper cites cc/paper_files/paper/2021/file/ f514cec81cb148559cf475e7426eed5e-Paper.

ADDQ: Adaptive Distributional Double Q-Learning cc/paper_files/paper/2021/file/ f514cec81cb148559cf475e7426eed5e-Paper

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:07.582196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T23:14:01.264177Z digest=sha256:595560c882342fc3506f88b0bf23b76579a59f13cb28dd5c106b226eb1131a6c

Pith citing papers

No inbound Pith citation observations are available.