Pith. sign in

Paper Citation Record · LEDGER

ADDQ: Adaptive Distributional Double Q-Learning

As of 19 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2506.19478.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.19478 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:14:03.900092Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved11
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6b58787e-adc2-4345-9442-380a38372e8a · outbound

This paper cites N(µ 2, σ2 2)) distributed rewards.

ADDQ: Adaptive Distributional Double Q-Learning N(µ 2, σ2 2)) distributed rewards

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:07.274394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:02.256253Z digest=sha256:505f6f60eaab1f0c37810d9546ebbfa8c5a8991890893726e16daa95d5ff61d3

Observation 1f0873cc-b2c5-46f5-b637-ed0c559a9042 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T23:14:07.024193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:02.523971Z digest=sha256:c7f30569fae8d578e392ba3d8746cd548e2a7d9254646a9d654ed1df2b15125b

Observation 57272fa8-0012-472b-9465-b101f58a8957 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.200827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:02.945633Z digest=sha256:cebb9ff8bfc778e6887500095448b6e39f0b240aa49139b125bb0474367ff7ef

Observation 527420bf-ee6c-4d98-99c0-87c20539fcf1 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:05.933292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.071047Z digest=sha256:1d8ccfe4aa736f993a42ae3d208fab04f9cdb962c2a0ee6d961f960c7400b6b8

Observation 57d6f0a1-8c3c-44dd-b9f7-6b49309e41b9 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.735255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:02.711844Z digest=sha256:ffc556a8bf88f0e9ffd81f217cb6edff83c92035e2cbc382c0ec2f380c1a4323

Observation c71b2eca-c759-432f-93cc-16b2c4d4ca6f · outbound

This paper cites , dsuch that Fn is Fn+1-measurable and for alli= 1,.

ADDQ: Adaptive Distributional Double Q-Learning , dsuch that Fn is Fn+1-measurable and for alli= 1,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:05.616463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.327831Z digest=sha256:d95667089945d338bd6c134067579c4b0bc987c7f015e006c70915a7f140956a

Observation ab87d9cf-f490-4e43-b9fc-27c55dc9a23a · outbound

This paper cites , dis adapted with ∞X n=1 αi,n =∞and ∞X n=1 α2 i,n <∞a.s.

ADDQ: Adaptive Distributional Double Q-Learning , dis adapted with ∞X n=1 αi,n =∞and ∞X n=1 α2 i,n <∞a.s

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:05.335693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.420474Z digest=sha256:e1866d681624b4950783419cd5a484898af7ba7d8073146f566740e36a19b763

Observation 1409acd7-b18b-4d17-89d2-c1fcbf0d2329 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:05.066511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.504810Z digest=sha256:199e8ea57b43a77f767c8901c6a85a35fac2b9c3fa1578ca24c8b134c43db386

Observation 5127464a-b216-45f1-9214-9bdef6185004 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:06.430275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.663227Z digest=sha256:64383c38e8b759643d315a3f2ca1994bfe3adc29f65b976f4aa086ae527b7485

Observation 024e4a3c-9eb5-4061-97b8-93d697c5b838 · outbound

This paper cites an unresolved cited work.

ADDQ: Adaptive Distributional Double Q-Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T23:14:04.768990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.779183Z digest=sha256:3019fb34b851d5046caa02b667561a96b6b9da7f2ac26a70d1d708b629692983

Observation b2b31dbe-d416-43cc-9c25-6f18bc953f76 · outbound

This paper cites Proof of Thoerem F .2.Step 1: Convergence of mean values toQ ∗ The proof mainly follows (Rowland et al., 2018) and (van Hasselt, 2010).

ADDQ: Adaptive Distributional Double Q-Learning Proof of Thoerem F .2.Step 1: Convergence of mean values toQ ∗ The proof mainly follows (Rowland et al., 2018) and (van Hasselt, 2010)

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:04.419404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:03.900092Z digest=sha256:249268b40bddde7bd35cd96eb48611e44586fb30a8be5507c6b212efb2113e49

Observation 28ac423a-4283-43a5-aae1-2130bf241611 · outbound

This paper cites Sudakov, V.

ADDQ: Adaptive Distributional Double Q-Learning Sudakov, V

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.783734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.783734Z digest=sha256:5d46e0273aaf285f21a745bd66bfe1e308cd0c6efefd009cb5a5d2fc9fb397a9

Observation fcf4708f-c7dd-42cc-9c50-92c62f4b8c95 · outbound

This paper cites Deep Reinforcement Learning with Double Q-learning.

ADDQ: Adaptive Distributional Double Q-Learning Deep Reinforcement Learning with Double Q-learning

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:02.096212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:02.096212Z digest=sha256:5fdb94324849ec8f87272b6820ea7631cc0983b0cba5856b71e30e1ef67237ae

Observation 6d401963-943c-487c-829f-75d90edf13e9 · outbound

This paper cites Dopamine: A Research Framework for Deep Reinforcement Learning.

ADDQ: Adaptive Distributional Double Q-Learning Dopamine: A Research Framework for Deep Reinforcement Learning

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.517365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.517365Z digest=sha256:398b8118b752e472ea737272181da202932fce6eddd19731786ba919d152a83c

Observation 49fa3b95-c673-4957-a3a2-465b7f2aee64 · outbound

This paper cites Agent57: Outperforming the Atari Human Benchmark.

ADDQ: Adaptive Distributional Double Q-Learning Agent57: Outperforming the Atari Human Benchmark

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.364990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.364990Z digest=sha256:466d05fdcec9878a5ee9d10c9dd6e716f08a9ca8ef81869d824da2782de22bcc

Observation bdd6cba9-744b-4471-890a-08812cefc87d · outbound

This paper cites Thrun, S.

ADDQ: Adaptive Distributional Double Q-Learning Thrun, S

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T23:14:01.936257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:14:01.936257Z digest=sha256:257d08342428ab01b5759c29876ae9072242886db35b870ee540e08e50a75cdd

Observation 81d9263f-6718-4c68-9b87-850c84d9562f · outbound

This paper cites The Potential of the Return Distribution for Exploration in RL.

ADDQ: Adaptive Distributional Double Q-Learning The Potential of the Return Distribution for Exploration in RL

Reference 2019

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T23:14:04.164753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:01.663649Z digest=sha256:a28cd7d6b41c2d525ade1dc0e91eb4c1a68b41e40961ee1a223a9d7be0e01f50

Observation e8a5cf44-19f7-402b-9f16-9bc316766de8 · outbound

This paper cites cc/paper_files/paper/2021/file/ f514cec81cb148559cf475e7426eed5e-Paper.

ADDQ: Adaptive Distributional Double Q-Learning cc/paper_files/paper/2021/file/ f514cec81cb148559cf475e7426eed5e-Paper

Reference 2021

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:14:07.582196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T23:14:01.264177Z digest=sha256:7143f3a3f82451bc71beb66c55a6524601fe97863cf3fba9faa671a30c803c8e

Pith citing papers

No inbound Pith citation observations are available.