Pith. sign in

Paper Citation Record · LEDGER

Uncertainty Prioritized Experience Replay

As of 9 August 2026, this Paper Citation Record lists 95 of 95 outbound references and 3 inbound Pith citation observations for arXiv:2506.09270.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09270 v1

Coverage vector

measured 95 of 95 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:59:15.652640Z

measured 98 of 98 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T11:18:08.362562Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T11:24:38.227917Z

Reference resolution

95 of 95 outbound references displayed

  • verified exact17
  • verified fuzzy13
  • unresolved63
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fb97e573-b264-4fcd-845c-bd546b4467e0 · outbound

This paper cites Query The Agent: Improving sample efficiency through epistemic uncertainty estimation.

Uncertainty Prioritized Experience Replay Query The Agent: Improving sample efficiency through epistemic uncertainty estimation

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:19.654545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:04.530922Z digest=sha256:1dcf0b039773672dfb78c3d6c0a6bdcf74f7409ed3da141796d476b0fe2971dc

Observation a7bcd771-d342-46ff-b504-5a4df1915a15 · outbound

This paper cites Hindsight Experience Replay.

Uncertainty Prioritized Experience Replay Hindsight Experience Replay

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:04.569792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:04.569792Z digest=sha256:61a57786afe3bb091023190493a2aa28781ec1122bc62a18cdc0b2c89effd448

Observation 2dcefea6-dd98-4cfa-a4eb-a9311e031e8a · outbound

This paper cites Optimism and pessimism in optimised replay.

Uncertainty Prioritized Experience Replay Optimism and pessimism in optimised replay

Reference 3

Resolution
verified exact
doi, observed 2026-08-07T04:59:17.122946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:04.632231Z digest=sha256:e7d77111bab007562edf4f7fef08537ff6eaa891198841daa860db3d0ea4386f

Observation f141325a-5c0e-44cd-ac48-590d06a4ad6b · outbound

This paper cites Learning is planning: near Bayes-optimal reinforcement learning via Monte-Carlo tree search.

Uncertainty Prioritized Experience Replay Learning is planning: near Bayes-optimal reinforcement learning via Monte-Carlo tree search

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:19.424105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:04.690612Z digest=sha256:43680630de8eb95fae8a050df39ec2e8647e748ee9b4b3dd25d67101c9852b7a

Observation d945c53b-1d4b-4bb0-a00d-fe635ccdb1fa · outbound

This paper cites Using confidence bounds for exploitation-exploration trade-offs.

Uncertainty Prioritized Experience Replay Using confidence bounds for exploitation-exploration trade-offs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:04.743104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:04.743104Z digest=sha256:4b142568b55aa6942d304ee8ad4a1150616674a56f51075cb9d4a6a6ea710d20

Observation 8ca8a7da-2a8b-4db4-bb0e-4d6a3b316323 · outbound

This paper cites Using confidence bounds for exploitation-exploration trade-offs.

Uncertainty Prioritized Experience Replay Using confidence bounds for exploitation-exploration trade-offs

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:04.827917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:04.827917Z digest=sha256:623c6e70c6f6230bd676ec81ce8625e26a001e9baf042924cf68db526ef437d9

Observation 26a78126-7459-4a64-adf7-f6558001cd03 · outbound

This paper cites Never Give Up: Learning Directed Exploration Strategies.

Uncertainty Prioritized Experience Replay Never Give Up: Learning Directed Exploration Strategies

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:04.925034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:04.925034Z digest=sha256:d8a8412ca665ae97b6a95e52b9585858c41e4274be49dab9e8d657b2ce831f62

Observation b96206d6-dec6-46c6-89ed-cbf1ce1d34f4 · outbound

This paper cites Emergent Tool Use From Multi-Agent Autocurricula.

Uncertainty Prioritized Experience Replay Emergent Tool Use From Multi-Agent Autocurricula

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.064790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.064790Z digest=sha256:f908b54efbdb9bed71fe6d819134066de72e8cd4cd86cd7fd2afb5d5dde0322c

Observation 8a116d71-8a6f-4518-9228-0a5ab70a75be · outbound

This paper cites The Effectiveness of Memory Replay in Large Scale Continual Learning.

Uncertainty Prioritized Experience Replay The Effectiveness of Memory Replay in Large Scale Continual Learning

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:19.240771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:05.171526Z digest=sha256:cd3f12b61d86ac00382c4e782c77d07e77a61f3ef1259dc4fb88c94a504abd64

Observation 7b5cf619-8faf-47fa-8372-4bc5fc5d03c3 · outbound

This paper cites Intrinsic motivation and reinforcement learning.

Uncertainty Prioritized Experience Replay Intrinsic motivation and reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.264885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.264885Z digest=sha256:8c2a481a7e3f1ace0aa4155e383f0aff7e248df00d3d27c4ebd0c177e647ba45

Observation 568aba0d-016a-4d3d-b9e6-2dfa3c88e978 · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.357368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.357368Z digest=sha256:5da5958457448e9ea46a2715762866f540aee377271e855d2d0005ec21e46378

Observation 1f4c9586-a599-4d39-9308-790edf3b5131 · outbound

This paper cites Unifying count-based exploration and intrinsic motivation.

Uncertainty Prioritized Experience Replay Unifying count-based exploration and intrinsic motivation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.474141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.474141Z digest=sha256:8035f07ab0775137ea2031f191f25a6816e837a74a04527b9577cab5586b9e23

Observation ff75eacd-d914-4504-acda-74c41be2a9b7 · outbound

This paper cites Unifying Count-Based Exploration and Intrinsic Motivation.

Uncertainty Prioritized Experience Replay Unifying Count-Based Exploration and Intrinsic Motivation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.580974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.580974Z digest=sha256:3b2127775d1a69b9fb5af8989adec13d05e6ac69a990f9c45f4ea9e27fb9026f

Observation 75456ecb-df69-45f9-8b4f-96daa7afdc11 · outbound

This paper cites A distributional perspective on reinforcement learning.

Uncertainty Prioritized Experience Replay A distributional perspective on reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.674967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.674967Z digest=sha256:7dac8bedc8266bbe7a87c2cc7711fe3f69e7aabf727399aec5d30e07b20cf428

Observation 1683a5fb-df9f-49de-b8da-77a5d6c14600 · outbound

This paper cites Distributional Reinforcement Learning.

Uncertainty Prioritized Experience Replay Distributional Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.781873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.781873Z digest=sha256:1205a8bc805d71cd21a4d8ac4c433c00fd70566fe14e4f9fe73870c4942b6197

Observation 84e1ba60-abf4-464f-9cdf-397c5d33fe06 · outbound

This paper cites Bickel and David A.

Uncertainty Prioritized Experience Replay Bickel and David A

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:05.886453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:05.886453Z digest=sha256:c107e099d70ba9e23343d00dc8eb1895a7bb7341855b0bf45be9c8e1a41ac25e

Observation 6269d055-d160-45b3-95d1-acd92639582f · outbound

This paper cites Wang, Will Dabney, Kevin J.

Uncertainty Prioritized Experience Replay Wang, Will Dabney, Kevin J

Reference 17

Resolution
verified exact
doi, observed 2026-08-07T04:59:17.009321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:06.000589Z digest=sha256:a4b137cdf2e7f9a4bad67be26ed3817c6775def54025db9047e0112d5bcb8030

Observation a4bc6f0c-19f9-4844-8a0d-428cc5d6a071 · outbound

This paper cites Bayes-optimal reinforcement learning for discrete uncertainty domains.

Uncertainty Prioritized Experience Replay Bayes-optimal reinforcement learning for discrete uncertainty domains

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.143349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.143349Z digest=sha256:b10b9890097d94fa85938ca7e36da1a823baacaa9c9e821030e39312b3d01f8e

Observation 935169f4-f315-4aa0-b1e2-c56355142224 · outbound

This paper cites Exploration by Random Network Distillation.

Uncertainty Prioritized Experience Replay Exploration by Random Network Distillation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.277250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.277250Z digest=sha256:8245739f6928020faf6b1a8757d6f194c44219acd8a5ed805670372669846695

Observation 728a8fb2-4c7f-4286-858c-987f2b8e97d2 · outbound

This paper cites Disentangling Epistemic and Aleatoric Uncertainty in Reinforcement Learning.

Uncertainty Prioritized Experience Replay Disentangling Epistemic and Aleatoric Uncertainty in Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.385820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.385820Z digest=sha256:0eabd19bfe446d6175826a13d053283f5a258600cdec18073f9fd087cb6de964

Observation c77bd387-51a7-4dc8-b222-db8448418cfb · outbound

This paper cites Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models.

Uncertainty Prioritized Experience Replay Deep Reinforcement Learning in a Handful of Trials using Probabilistic Dynamics Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.483617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.483617Z digest=sha256:5e74cc6b67a0a7d7629cd7381dd11e7bd47a5729c1bce5a25a5e31655649b410

Observation 3d61215e-54bc-4d8f-b29d-5e3dbb6dcb63 · outbound

This paper cites Estimating Risk and Uncertainty in Deep Reinforcement Learning.

Uncertainty Prioritized Experience Replay Estimating Risk and Uncertainty in Deep Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.596214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.596214Z digest=sha256:d37964a6e63b3f037acb5ccc43f9577a8a14a4d65d5da69a6c077c94d9e16344

Observation fd13b90b-268a-45a5-9dcd-82ca2ecb0ec0 · outbound

This paper cites Active learning with statistical models.

Uncertainty Prioritized Experience Replay Active learning with statistical models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.706141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.706141Z digest=sha256:28fef9b365a059a0a17a2ed02a9b7313953ae135fd426a6e308700d0bd222e21

Observation c7f31970-2c43-431d-a05f-c84e16ad64da · outbound

This paper cites Distributional Reinforcement Learning with Quantile Regression.

Uncertainty Prioritized Experience Replay Distributional Reinforcement Learning with Quantile Regression

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.831256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.831256Z digest=sha256:40ecf14b43b7d416e214b6135949b519ae1815af8dc740c4cde6ad29cf60d68a

Observation 001546de-8a2a-4267-b1d2-a113e1c142b8 · outbound

This paper cites Daw, Yael Niv, and Peter Dayan.

Uncertainty Prioritized Experience Replay Daw, Yael Niv, and Peter Dayan

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:06.984784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:06.984784Z digest=sha256:92dbb3d6b341120d4b72481618ca50935bc6e1b17a10d7815ba8587e78570de5

Observation 6c7b6f89-f543-4833-8836-ab3f002f9c31 · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning.

Uncertainty Prioritized Experience Replay Magnetic control of tokamak plasmas through deep reinforcement learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.089903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.089903Z digest=sha256:7d12c6baf6f59157318e1df63d3e4eca18e32d2580bc52cf6e33cc63669b2537

Observation 222cf5f0-0ae4-4284-ba3e-470359698ae9 · outbound

This paper cites Revisiting fundamentals of experience replay.

Uncertainty Prioritized Experience Replay Revisiting fundamentals of experience replay

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.187982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.187982Z digest=sha256:cbf8c4d0dfc420a1e51e841e735852bd6f19c11193cdadc6dfc95ac2ccf1c4a4

Observation f26112ed-15aa-4df6-8df9-932eb9a86bb2 · outbound

This paper cites Foster and Matthew A.

Uncertainty Prioritized Experience Replay Foster and Matthew A

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.324437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.324437Z digest=sha256:2f78286cb5df614abab5f523c7d36e747509179f43ff7e99c8eed86312b27c2c

Observation 1e456dfb-18d3-4cee-9ef4-68df9164b4b3 · outbound

This paper cites Dropout as a bayesian approximation: Representing model uncertainty in deep learning.

Uncertainty Prioritized Experience Replay Dropout as a bayesian approximation: Representing model uncertainty in deep learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.496686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.496686Z digest=sha256:72bf7256b33c51f95f3152e104ade52c291c2d3ce993bcdb938caef415cc8b7e

Observation cf6aae28-93a9-4ae2-82f1-1aee9c2e8659 · outbound

This paper cites Dopamine, inference, and uncertainty.

Uncertainty Prioritized Experience Replay Dopamine, inference, and uncertainty

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.606446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.606446Z digest=sha256:7749355e92f732a460764bdfae4f4f41ba9fca7458620e3901e7130bdeaeaa4b

Observation 403041e9-fbba-4d22-8e04-1c535bdbcc73 · outbound

This paper cites Econometric analysis 4th edition.

Uncertainty Prioritized Experience Replay Econometric analysis 4th edition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.735424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.735424Z digest=sha256:3c9a5b569516210cbbfa9f9b57e1b4030b82789153184c2dcd57c10e80d02b1e

Observation 5ad65fde-b29e-4ef5-bca6-e532f341b840 · outbound

This paper cites Grewe, and João Sacramento.

Uncertainty Prioritized Experience Replay Grewe, and João Sacramento

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.881303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.881303Z digest=sha256:53492004f3361ae993d1ad495101307df1a3d849478e2e5fafa158026626eba7

Observation ed222ed2-d44b-4757-8170-b0e15885d1e8 · outbound

This paper cites Rainbow: Combining Improvements in Deep Reinforcement Learning.

Uncertainty Prioritized Experience Replay Rainbow: Combining Improvements in Deep Reinforcement Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:07.996842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:07.996842Z digest=sha256:45256f940b971dbcae9518a9e8d25e0fdc4952a9658207b5fec682b158512226

Observation 4e009739-942c-44d2-9be2-432d59200e16 · outbound

This paper cites Meta reinforcement learning as task inference.

Uncertainty Prioritized Experience Replay Meta reinforcement learning as task inference

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.137927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.137927Z digest=sha256:379328579e27505687ae013f0979d3f7796f2d4cb4764cfbcc754eaa001704d5

Observation d9a75a42-f4ed-4b0d-ae43-e81a74ae7a90 · outbound

This paper cites Aleatoric and epistemic uncertainty in machine learning: an introduction to concepts and methods.

Uncertainty Prioritized Experience Replay Aleatoric and epistemic uncertainty in machine learning: an introduction to concepts and methods

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.270325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.270325Z digest=sha256:41fb33e39f2c5b64ba7e4ffbb1587500d0f44f94ce0d04f7872dbe9496d8decb

Observation 00c624ad-5a39-40f6-9c73-c4adfa1539a2 · outbound

This paper cites On the Importance of Exploration for Generalization in Reinforcement Learning.

Uncertainty Prioritized Experience Replay On the Importance of Exploration for Generalization in Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.382492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.382492Z digest=sha256:c127bea280171607da93e16d8f53a00412b5fe6cdde0bc1673996794eda24b30

Observation da757c5c-660f-4731-b6b3-fd9749b4b84d · outbound

This paper cites Uncertainty-Aware Reinforcement Learning for Collision Avoidance.

Uncertainty Prioritized Experience Replay Uncertainty-Aware Reinforcement Learning for Collision Avoidance

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.507227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.507227Z digest=sha256:f9556fa882a099be06bb0a68aad9b96157a21afa40215d1bf55fa7e538565637

Observation 756025c3-709f-4a6f-bcd1-122a99c94773 · outbound

This paper cites Continual Reinforcement Learning with Multi-Timescale Replay.

Uncertainty Prioritized Experience Replay Continual Reinforcement Learning with Multi-Timescale Replay

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:18.904049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:08.646940Z digest=sha256:ab6ecb5a2b91b73177271505d509219c981ca52a096020b858e74636062f4992

Observation 48c589a1-7d2e-4325-b00e-4878fab141bf · outbound

This paper cites Towards Continual Reinforcement Learning: A Review and Perspectives.

Uncertainty Prioritized Experience Replay Towards Continual Reinforcement Learning: A Review and Perspectives

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.782236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.782236Z digest=sha256:93e0fdd46a260b10027bb82fa923ca68df0ae37a58fb26a34cee92ccc0d51f22

Observation dccd917f-1936-439d-b234-1a7b872a8d61 · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:08.867289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:08.867289Z digest=sha256:93ba34b2f5ce42cdd8865b3c1b5269bba5092e56a6a50d1c48a3b517c9b072dc

Observation e8a84fce-8ea7-4f84-8967-4ede4ec6f722 · outbound

This paper cites DEUP: Direct Epistemic Uncertainty Prediction.

Uncertainty Prioritized Experience Replay DEUP: Direct Epistemic Uncertainty Prediction

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:09.017908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:09.017908Z digest=sha256:6822baf6193d3866a4948cba6c93e7419c6c008a594de0535974163a705412d2

Observation c86e5450-c23f-4c92-99ec-96f14f57b313 · outbound

This paper cites Bandit Algorithms.

Uncertainty Prioritized Experience Replay Bandit Algorithms

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:09.169304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:09.169304Z digest=sha256:f4309be2c302f629168f73b35954b78a0e1fc9b9b5c825a63728a488308b6553

Observation 06c5e871-7c78-4405-a94b-5fa480bbe544 · outbound

This paper cites Continual Learning Using Bayesian Neural Networks.

Uncertainty Prioritized Experience Replay Continual Learning Using Bayesian Neural Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:09.325129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:09.325129Z digest=sha256:46dc824ce24669ce803ff9e457bba075fc1a415951a77a5cfa361effee5c427d

Observation fa4c9891-e941-45ec-90a3-bc3cc978143d · outbound

This paper cites Self-improving reactive agents based on reinforcement learning, planning and teaching.

Uncertainty Prioritized Experience Replay Self-improving reactive agents based on reinforcement learning, planning and teaching

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:09.416141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:09.416141Z digest=sha256:6cbc5681f342982020b5bfce3f03bc8301878bd11c7015417013819c481d6a37

Observation d62dff62-45aa-4d3e-b3db-ce46ddb30470 · outbound

This paper cites Distributional reinforcement learning with epistemic and aleatoric uncertainty estimation.

Uncertainty Prioritized Experience Replay Distributional reinforcement learning with epistemic and aleatoric uncertainty estimation

Reference 45

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T04:59:18.626167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:09.545903Z digest=sha256:53fe8642f2b59e23d428ef3e51e54a5be2ac51b7be13861346308df054d13169

Observation 21d6d1a4-6d46-4f8c-967c-303c6d683c39 · outbound

This paper cites The Effects of Memory Replay in Reinforcement Learning.

Uncertainty Prioritized Experience Replay The Effects of Memory Replay in Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:09.694588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:09.694588Z digest=sha256:c1e002e9907a168148b77fdf67ce3c4dccc375f669c9244a3d8781908ff8fed7

Observation c985d732-d625-4f2f-9aae-c98ebf2b03ab · outbound

This paper cites Dolan, Zeb Kurth-Nelson, and Timothy E.J.

Uncertainty Prioritized Experience Replay Dolan, Zeb Kurth-Nelson, and Timothy E.J

Reference 47

Resolution
verified exact
doi, observed 2026-08-07T04:59:16.839272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:09.811133Z digest=sha256:b1bf228b37afc4b910611aa363ee3846aada4a8d87bdc3bd3a172a6cdc200283

Observation 2a17d01a-be66-4f75-bde9-57049b616361 · outbound

This paper cites Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement Learning.

Uncertainty Prioritized Experience Replay Flipping Coins to Estimate Pseudocounts for Exploration in Reinforcement Learning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:18.436575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:09.924917Z digest=sha256:e0ed257c61a7cbe02d100ebde92ac8431bc854451aef66fa88f1659a770c4a7d

Observation 4ffe179a-8b7e-433c-a73e-787898a280f3 · outbound

This paper cites Safe Reinforcement Learning with Model Uncertainty Estimates.

Uncertainty Prioritized Experience Replay Safe Reinforcement Learning with Model Uncertainty Estimates

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:10.058735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:10.058735Z digest=sha256:880ca43d97994a538ba129c03fa144938a745cfcd29572d1aeea08daffe67ad3

Observation 16123032-5733-428e-9443-60eea35ba479 · outbound

This paper cites Sample Efficient Deep Reinforcement Learning via Uncertainty Estimation.

Uncertainty Prioritized Experience Replay Sample Efficient Deep Reinforcement Learning via Uncertainty Estimation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:10.157842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:10.157842Z digest=sha256:b64577813558f9af419cfc94be659e10aa635d233e2badfc3b7d6b23aefe57f1

Observation 58f54b4c-6f58-47eb-a4bc-18032109ccb0 · outbound

This paper cites Bayesian decision problems and markov chains.

Uncertainty Prioritized Experience Replay Bayesian decision problems and markov chains

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:22.551424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:10.277282Z digest=sha256:7fddfcde339216d5a1490d9855b96d1a8096900873f4c146eaf4f7e0cfe12903

Observation 4e776f21-8b9d-44e3-9adb-e261db8a050a · outbound

This paper cites Mattar and Nathaniel D.

Uncertainty Prioritized Experience Replay Mattar and Nathaniel D

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:10.420297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:10.420297Z digest=sha256:cf6942c3a416adc6729d2af7d5305bf3798759efd49b66ce335429fdc0f7e331

Observation d2a9ac14-f92c-4426-a720-9ca18cc63cf9 · outbound

This paper cites How to Stay Curious while avoiding Noisy TVs using Aleatoric Uncertainty Estimation.

Uncertainty Prioritized Experience Replay How to Stay Curious while avoiding Noisy TVs using Aleatoric Uncertainty Estimation

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:22.338045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:10.586221Z digest=sha256:160be466e57031091000a66f6800c461e473a2fdd9be5b889b84a7acb60a9017

Observation 861ffbb3-d831-4ff5-bfc3-ffbe65f262c3 · outbound

This paper cites McNamara, Álvaro Tejero-Cantero, Stéphanie Trouche, Natalia Campo-Urriza, and David Dupret.

Uncertainty Prioritized Experience Replay McNamara, Álvaro Tejero-Cantero, Stéphanie Trouche, Natalia Campo-Urriza, and David Dupret

Reference 54

Resolution
verified exact
doi, observed 2026-08-07T04:59:16.700629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:10.720707Z digest=sha256:6e1271205b7cac54363e24f0c1bb005c71ab7c4782253e9a6be7eab882ae8457

Observation 5c154923-54b5-4269-beb7-bd3c9451a1b0 · outbound

This paper cites Rusu, Joel Veness, Marc G.

Uncertainty Prioritized Experience Replay Rusu, Joel Veness, Marc G

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:10.836052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:10.836052Z digest=sha256:01fccff3919ccc8276ce7011bcfb1726ef6c2a84cb2ff8fe03bca0a8d5947e1e

Observation ce26a3ae-5ba9-44b4-942d-51cb16b05b3f · outbound

This paper cites Moore and Christopher G.

Uncertainty Prioritized Experience Replay Moore and Christopher G

Reference 56

Resolution
verified exact
doi, observed 2026-08-07T04:59:16.496733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:10.966094Z digest=sha256:9d82900803285bbf3cffd3d34df6080a9782644ffae968207e039b9c701a418f

Observation 295d8212-0883-4a3a-ab50-e612b91b632a · outbound

This paper cites Overcoming Exploration in Reinforcement Learning with Demonstrations.

Uncertainty Prioritized Experience Replay Overcoming Exploration in Reinforcement Learning with Demonstrations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:11.082277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:11.082277Z digest=sha256:a3c88a70d582eb971b31aa654cd8b7e19d57e5bfdf6ddad36d746018c1ff68b3

Observation ee8a293e-3fed-4d41-8168-9231d4d77e6b · outbound

This paper cites Collision Probability Matching Loss for Disentangling Epistemic Uncertainty from Aleatoric Uncertainty.

Uncertainty Prioritized Experience Replay Collision Probability Matching Loss for Disentangling Epistemic Uncertainty from Aleatoric Uncertainty

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:22.085699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:11.217545Z digest=sha256:068383c577d455eee329f25f4b982a2c25e7d5664488185434d407c78866573c

Observation 6f471bab-db8f-4b27-b9f7-46446e3b9655 · outbound

This paper cites How to measure uncertainty in uncertainty sampling for active learning.

Uncertainty Prioritized Experience Replay How to measure uncertainty in uncertainty sampling for active learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:11.299942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:11.299942Z digest=sha256:30b933938a043d351fe125dfbc5042469b9876b92780adc96a0c87e7cd40ec7e

Observation d2b85c3e-3d18-4ea0-8e71-ff0b5aaf8ae3 · outbound

This paper cites A review On reinforcement learning: Introduction and applications in industrial process control.

Uncertainty Prioritized Experience Replay A review On reinforcement learning: Introduction and applications in industrial process control

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:11.439530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:11.439530Z digest=sha256:6b437c25ced7a9f0720bb1c9ad752b26708b25e4c7d13760fdb0843a93b798eb

Observation 1ecab6c4-ebe4-4684-a23c-be41dba30d73 · outbound

This paper cites Information-Directed Exploration for Deep Reinforcement Learning.

Uncertainty Prioritized Experience Replay Information-Directed Exploration for Deep Reinforcement Learning

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:18.211314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:11.622714Z digest=sha256:4787c757285c095b1bc9ae6f993f82cac2393de34d7b4e7372760111578efbe0

Observation f313ebd7-a4bd-4626-988b-13528cf7511a · outbound

This paper cites Efficient Exploration via Epistemic-Risk-Seeking Policy Optimization.

Uncertainty Prioritized Experience Replay Efficient Exploration via Epistemic-Risk-Seeking Policy Optimization

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:11.753710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:11.753710Z digest=sha256:edba1de250dc47979f590960fa3735c0608d6bdd6b377c797c2d9f1f2ca7d5e1

Observation 3adc404b-dfde-4bd9-9085-d9fa031ef1d8 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Uncertainty Prioritized Experience Replay Dota 2 with Large Scale Deep Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:11.862987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:11.862987Z digest=sha256:c8830ce5c43ec823ece200a1c724d1b95d762ddfa94108cc189d53be802de6e8

Observation d3614bab-4f10-4f0a-8415-b79e634f3987 · outbound

This paper cites Deep Exploration via Bootstrapped DQN.

Uncertainty Prioritized Experience Replay Deep Exploration via Bootstrapped DQN

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:12.010327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:12.010327Z digest=sha256:21970d8a0dd68dd5aabc48624aa678fac4987bce2a7c8b9d6c7bd7644a069322

Observation 60be9072-be07-494c-bfc3-5ecc58643fa9 · outbound

This paper cites Randomized prior functions for deep reinforcement learning.

Uncertainty Prioritized Experience Replay Randomized prior functions for deep reinforcement learning

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:21.814717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:12.139416Z digest=sha256:de61ee7998d2cc8eeeca41bae3917ddca3f0e129297a6bafa890c8be4d81ec62

Observation 2e790b1e-e4c7-4bf8-9f9e-172a29b6a5fc · outbound

This paper cites Epistemic Neural Networks.

Uncertainty Prioritized Experience Replay Epistemic Neural Networks

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:12.300673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:12.300673Z digest=sha256:b2728a1e1e11adffb7d00c818d5ab863dc3f2ff58409f78af71551dbfe38db86

Observation 8cb2fdb0-37c4-4c93-aaa6-9221b863df8e · outbound

This paper cites Count-based exploration with neural density models.

Uncertainty Prioritized Experience Replay Count-based exploration with neural density models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:21.583223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:12.405707Z digest=sha256:0e84b4bc4ad365cb68f832937b1b43454a576a9fe2be2c57587d48b7d0ae19f5

Observation d98ce937-9f32-494c-b26b-1921c230e943 · outbound

This paper cites Count-Based Exploration with Neural Density Models.

Uncertainty Prioritized Experience Replay Count-Based Exploration with Neural Density Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:12.561871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:12.561871Z digest=sha256:e9b0e0544fcf81844b7cded6c3dde694e274fc3c222b66b18317ac4fa32d874d

Observation f6cbb842-64e2-4ccf-b3a8-3b9057cc8e0e · outbound

This paper cites What is intrinsic motivation? A typology of computational approaches.

Uncertainty Prioritized Experience Replay What is intrinsic motivation? A typology of computational approaches

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:12.683297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:12.683297Z digest=sha256:35a1055248d974a7e43ff3927887546662fd969d1b9ee53ddb061e5ebd6510f9

Observation fc5d0834-ab9a-4669-8b45-0f20b40a08af · outbound

This paper cites Understanding and mitigating the limitations of prioritized experience replay.

Uncertainty Prioritized Experience Replay Understanding and mitigating the limitations of prioritized experience replay

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:21.351685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:12.797300Z digest=sha256:1f4f67f53e27328fefc6618c67a25ca6c36446bcec11aaf4e92adf45c0dfc14a

Observation 86e7e25f-b4de-4ba3-8f75-fc2c6c573803 · outbound

This paper cites Curiosity-driven Exploration by Self-supervised Prediction.

Uncertainty Prioritized Experience Replay Curiosity-driven Exploration by Self-supervised Prediction

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:12.920187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:12.920187Z digest=sha256:60eef2377f9ba46a41efae05f89c217ead070ad50d9659df6ad9f0993c90b56d

Observation ecdba1bb-5cf8-493b-9498-c5b710d1283b · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 72

Resolution
verified exact
doi, observed 2026-08-07T04:59:16.318334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:13.034838Z digest=sha256:54ae09a44a8a9f142d7e7e03089c9a2bfefc46a4febb1060b4a50062a4e026c0

Observation 6d58fd36-4cb9-40d5-a733-62ba6a02cb27 · outbound

This paper cites Episodic Curiosity through Reachability.

Uncertainty Prioritized Experience Replay Episodic Curiosity through Reachability

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:13.161515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:13.161515Z digest=sha256:722f1de7a576bfd186a8132d92d316fb7c2b1aa49d1ecd3fe4a5a2aa07e6a863

Observation dcd68f51-7d06-4eee-909d-a6712866462b · outbound

This paper cites Prioritized Experience Replay.

Uncertainty Prioritized Experience Replay Prioritized Experience Replay

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:13.358289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:13.358289Z digest=sha256:34439535ac190be46d09e38602f50e0f8d0fab04fc74187f668b1ab532718b16

Observation bbb332ca-eb1f-4b06-a791-040ceec25aa7 · outbound

This paper cites Comparing Direct and Indirect Temporal - Difference Methods for Estimating the Variance of the Return.

Uncertainty Prioritized Experience Replay Comparing Direct and Indirect Temporal - Difference Methods for Estimating the Variance of the Return

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:21.090557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:13.560894Z digest=sha256:5bd3845be5bb77c3cc71ce9f6c2eb389b7ed9b37d323afe459cdc5294ccd5910

Observation 6e7b6193-52d2-477b-8ea9-317d503de93e · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:13.708754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:13.708754Z digest=sha256:a7fc583cc4987529972125226ff9ba695b6b9c5c46e721202a1e7071ea1717d0

Observation 246971cb-f47d-4b85-b54a-917bc75adaa6 · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 77

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:59:17.925150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:13.822579Z digest=sha256:865172b73ef349d01f18269c643495aa8faa33bdccaa186e1ba860058e4d258b

Observation ac8cd1c0-435e-479c-bb47-b536de0a502c · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

Uncertainty Prioritized Experience Replay Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:13.972098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:13.972098Z digest=sha256:65a45fa0da1fa9254e1ae4071c8b0cffa5d439b9574d557953a5d3c8a30eed05

Observation dba98005-e78d-46fd-8769-a73fcb7a77be · outbound

This paper cites Reinforcement learning and its connections with neuroscience and psychology.

Uncertainty Prioritized Experience Replay Reinforcement learning and its connections with neuroscience and psychology

Reference 79

Resolution
verified exact
doi, observed 2026-08-07T04:59:16.161963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.099713Z digest=sha256:98d3050616f565a050700594219cf637b831bf6f414edf449b81202374fd777d

Observation 5c64bd0d-2409-4c8b-a94d-31593ec6351a · outbound

This paper cites Attentive Experience Replay.

Uncertainty Prioritized Experience Replay Attentive Experience Replay

Reference 80

Resolution
verified exact
doi, observed 2026-08-07T04:59:15.990701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.224018Z digest=sha256:3558c44a52b47b67cd6857db487a76774b542392e2556f0406692501e9e4ee02

Observation 50b94434-f492-4107-b690-27876e5f7ca0 · outbound

This paper cites Reinforcement learning: An Introduction.

Uncertainty Prioritized Experience Replay Reinforcement learning: An Introduction

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:20.841670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.347077Z digest=sha256:a121f6e88c7d7455b06a5205f97bd7858ede315e94e790b24a3dd11c740032eb

Observation f3ee20cf-5ab9-4594-848d-797089234c14 · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:14.467194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:14.467194Z digest=sha256:0d20891b4bc5a9fd30fba7af305cee68be4c49b32b5f37034d27bc0146addb81

Observation 88ea7a60-935e-4688-9cdf-3646ff697ed2 · outbound

This paper cites Policy gradients with variance related risk criteria.

Uncertainty Prioritized Experience Replay Policy gradients with variance related risk criteria

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:20.605473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.542594Z digest=sha256:93bb321e11483832a56f0d78defe565525f422607f9464dc92ad23828a8d414c

Observation d56cb4d0-ea81-41e0-8620-006f79b22000 · outbound

This paper cites Learning the Variance of the Reward - To - Go.

Uncertainty Prioritized Experience Replay Learning the Variance of the Reward - To - Go

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:20.440237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.656558Z digest=sha256:267fa324e6b1b2d91a64df0f15ca8677ad23107bf775e8c76ac8d84484171963

Observation 0fa47a5e-be18-4314-a020-b3ad1e243aa8 · outbound

This paper cites \# exploration: A study of count-based exploration for deep reinforcement learning.

Uncertainty Prioritized Experience Replay \# exploration: A study of count-based exploration for deep reinforcement learning

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:20.308925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:14.746820Z digest=sha256:a19c497d068b15589366cc855f508b8e3a57f5fda7cc92c494c0a86d43c11824

Observation 8240a333-4a49-4961-afce-02f0c50f7179 · outbound

This paper cites Open-Ended Learning Leads to Generally Capable Agents.

Uncertainty Prioritized Experience Replay Open-Ended Learning Leads to Generally Capable Agents

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:14.869952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:14.869952Z digest=sha256:f1bc083f22aaccb698de3f2f3ab892d932e357ce19a755a34f9b1c5a3fe66e94

Observation acdcd362-4911-4d33-b93a-77ad88230a47 · outbound

This paper cites an unresolved cited work.

Uncertainty Prioritized Experience Replay Unresolved cited work

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:14.957572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:14.957572Z digest=sha256:f9a4b0c63810d9cf546f5fef1d93fd006be80a2b452323993482db815a1fefed

Observation 460fab3b-7e3e-4299-82d5-51724b78b9ad · outbound

This paper cites Q-learning.

Uncertainty Prioritized Experience Replay Q-learning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:15.091515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:15.091515Z digest=sha256:e1f3fb7e36c431054c6fb1f8bbf2ba44a037c000ad3b5ac0ee0607d38f643302

Observation 8e6ee2e0-2495-4e26-813b-a56d35ac5f4c · outbound

This paper cites A Review of Reinforcement Learning for Controlling Building Energy Systems From a Computer Science Perspective.

Uncertainty Prioritized Experience Replay A Review of Reinforcement Learning for Controlling Building Energy Systems From a Computer Science Perspective

Reference 89

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T04:59:17.593543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:15.148496Z digest=sha256:695bd469299ccea4d95dd585f7fe89734c94b8a9975ab8ba105c005ed60cac83

Observation 263fd8e0-8390-46f2-9d12-1823c362b03d · outbound

This paper cites An introduction to the kalman filter.

Uncertainty Prioritized Experience Replay An introduction to the kalman filter

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:15.262301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:15.262301Z digest=sha256:e2941f2f2b8fad73d8167ed14c2c1fbcd399c6e65b53137316c28e62b6cf2f51

Observation 5b5bb12f-2ef2-4404-952c-d317d65e25ce · outbound

This paper cites A Greedy Approach to Adapting the Trace Parameter for Temporal Difference Learning.

Uncertainty Prioritized Experience Replay A Greedy Approach to Adapting the Trace Parameter for Temporal Difference Learning

Reference 91

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:59:17.324070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:15.339545Z digest=sha256:ac7161cac360e4c6b34d42a77cabd574d8e708c1b9a7e88106dfcac60203dacf

Observation 62920f0f-f8c9-4246-9d0b-5e1b388d45b8 · outbound

This paper cites Minimum excess risk in bayesian learning.

Uncertainty Prioritized Experience Replay Minimum excess risk in bayesian learning

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:20.154488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:15.421052Z digest=sha256:9f26ecfba0b1faae12f26038bf7437460e954667a023d7a26f6b0a0887fe17c6

Observation 697573e4-74b8-470f-a4ff-7cdd44b0f0a1 · outbound

This paper cites Experience Replay Optimization.

Uncertainty Prioritized Experience Replay Experience Replay Optimization

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:59:19.916710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:15.469418Z digest=sha256:83ab4a924d825dd95952803697ddca218ebdd6e2967e7773aee1588261591851

Observation 7c5bbda1-3059-4e1c-9ff8-0d187fedcddc · outbound

This paper cites A survey on epistemic (model) uncertainty in supervised learning: Recent advances and applications.

Uncertainty Prioritized Experience Replay A survey on epistemic (model) uncertainty in supervised learning: Recent advances and applications

Reference 94

Resolution
verified exact
doi, observed 2026-08-07T04:59:15.810714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T04:59:15.550692Z digest=sha256:8b983ccc2d6fff4b7b73108343d016d5bd2fbd481b6076a646c83ddb24070fd2

Observation d4fa4b48-1ca3-430c-84bc-3d69928723fd · outbound

This paper cites VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning.

Uncertainty Prioritized Experience Replay VariBAD: A Very Good Method for Bayes-Adaptive Deep RL via Meta-Learning

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T04:59:15.652640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:59:15.652640Z digest=sha256:05b1dcb57e41872028e5d50c2fedf3e8f287a7cfe4f21fdc115a5ecc959e0824

Pith citing papers

Observation b9ef21d5-c932-462b-9dc2-18f73979296e · inbound

Uncertainty-Weighted Experience Replay for Continual MIMO Channel Prediction cites this paper.

Uncertainty-Weighted Experience Replay for Continual MIMO Channel Prediction Uncertainty Prioritized Experience Replay

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:30:17.954480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:27:25.479441Z digest=sha256:7b0d66056b9db88d50bc2b0fa099e08b9ac056e02294b3220c26ef8716f01154

Observation 23818c03-489d-4949-860d-6c99cabee30d · inbound

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty cites this paper.

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty Uncertainty Prioritized Experience Replay

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T16:13:36.100526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T16:08:23.588699Z digest=sha256:3df91c1f804da5fffaeccb9f027af96653bff59d401d864c544899dff7f1eb70

Observation 8acf85ba-a541-46aa-8b22-6d6d390b0119 · inbound

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty cites this paper.

Breaking the Epistemic Trap: Active Perception Under Compound Uncertainty Uncertainty Prioritized Experience Replay

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T11:24:38.229256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:18:08.362562Z digest=sha256:5932e37fb4c4eb2b7c9d6ed3a98063046830090581f98f401b1678698386ce18