Pith. sign in

Paper Citation Record · LEDGER

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment

As of 16 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 2 inbound Pith citation observations for arXiv:1908.02388.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.02388 v3

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T14:49:25.531670Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T12:33:26.408771Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-14T05:24:51.266012Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14dd12e1-55e2-42ec-8513-b8fa056e87a8 · outbound

This paper cites The relative ranking of methods differs from the one observed on MONTEZUMA ’S REVENGE.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment The relative ranking of methods differs from the one observed on MONTEZUMA ’S REVENGE

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:49:25.695872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.522824Z digest=sha256:1c6002189472c90ce70ef18da25da993bd7a98fa3f2206496bddd6120d278538

Observation 19c19b42-0809-4da6-8191-bbc32c812785 · outbound

This paper cites Count-Based Exploration with the Successor Representation.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Count-Based Exploration with the Successor Representation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.486107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.486107Z digest=sha256:09a7c7be844cd0c1092691298c5425354d2bb83806d306f3c8ab810ad3eabba5

Observation df733ac9-03d9-4b9c-bf05-8111888029d2 · outbound

This paper cites an unresolved cited work.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:49:25.669553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.531670Z digest=sha256:123360cf703bd9b85168d75dd5a5abff7e82fe7095c008e93eb740dd44a3caec

Observation 5f071908-6094-4ec0-862c-d475a41b76c8 · outbound

This paper cites Prioritized Experience Replay.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Prioritized Experience Replay

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.499942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.499942Z digest=sha256:4f99135128a973960ec1689fb2774d1583d5de19f62e7d5e79bc3fb092b1851a

Observation 393f7761-e1af-4ee2-a896-3852e8c7a2c1 · outbound

This paper cites Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Incentivizing Exploration In Reinforcement Learning With Deep Predictive Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.506043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.506043Z digest=sha256:620f4860a01ffd22f51f13bbcc60c110f0ace6187ee6637d431850aefe7285e0

Observation d79e6a2b-f834-4729-b296-81b768996423 · outbound

This paper cites Randomized Value Functions via Multiplicative Normalizing Flows.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Randomized Value Functions via Multiplicative Normalizing Flows

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-14T14:49:25.582259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.510578Z digest=sha256:3082317ee01570f42e63bceba279cdc92f583e5a92c71576bcf6fc117d064727

Observation c537773c-66bc-49d5-967b-190f627b0bef · outbound

This paper cites Pixel Recurrent Neural Networks.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Pixel Recurrent Neural Networks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.514771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.514771Z digest=sha256:9e78ff70f0d08c887793d923f19ff9a439f4aee889a65feb6039455242fdfd58

Observation 7e701219-c824-417f-9428-06cceb9f7cbd · outbound

This paper cites an unresolved cited work.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-14T14:49:25.708333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.518779Z digest=sha256:391e39e6117cddeaf216a211aa23ca8408ad03b4095540723f82556cec5d5c8c

Observation 5f1868bc-5318-4b3f-aecd-c789a3e5a82e · outbound

This paper cites Rainbow with ϵ-greedy exploration performs as well as other more complex exploration method.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Rainbow with ϵ-greedy exploration performs as well as other more complex exploration method

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:49:25.682534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.526629Z digest=sha256:674599a1bdc2ba89281fa3cac3fee7397cd93b3d616afad0b110b2ab9c553a0d

Observation bdf42e76-f177-4a78-96fa-bf52aaa7df74 · outbound

This paper cites P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment P., Mirza, M., Graves, A., Lillicrap, T., Harley, T., Silver, D., and Kavukcuoglu, K

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T14:49:25.720970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T14:49:25.490947Z digest=sha256:f19d658d74a4146d529f441334d800484ad4008c94be0404f06d495aa8fadcc1

Observation 0e9ad346-ed47-4757-b5d7-56791f537f3b · outbound

This paper cites The Uncertainty Bellman Equation and Exploration.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment The Uncertainty Bellman Equation and Exploration

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.495303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.495303Z digest=sha256:618550fe5e370702cee8a54137e25aa241cdd365c37dacaa11bcea6fb0f399ea

Observation e7e9db91-e43f-4371-b293-633742fc1b81 · outbound

This paper cites Large-Scale Study of Curiosity-Driven Learning.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Large-Scale Study of Curiosity-Driven Learning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.470844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.470844Z digest=sha256:8da9752668fad7c3e6dd51f924ca5d5c2c43d00f45ee0724d0d662ce31f69dd7

Observation dd0baacc-1615-486b-9105-0834f2409783 · outbound

This paper cites DORA The Explorer: Directed Outreaching Reinforcement Action-Selection.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment DORA The Explorer: Directed Outreaching Reinforcement Action-Selection

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.480419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.480419Z digest=sha256:15ac6c22a4698d8d11f74ce2d08db204470e591e4d1a513ca67a280763ef0712

Observation 3fd709d9-7856-4d0d-b6e2-81af30b8cd85 · outbound

This paper cites Dopamine: A Research Framework for Deep Reinforcement Learning.

Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment Dopamine: A Research Framework for Deep Reinforcement Learning

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-14T14:49:25.475786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:49:25.475786Z digest=sha256:8e634d98f44f1b43dac2cc7468895910a7fc92effd46cf307f6c577fc100f869

Pith citing papers

Observation 0493b36a-b47e-48a7-bd10-19ff93967054 · inbound

A survey on intrinsic motivation in reinforcement learning cites this paper.

A survey on intrinsic motivation in reinforcement learning Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment

Reference 182

Resolution
unresolved
no resolver link, observed 2026-08-14T12:33:26.408771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T12:33:26.408771Z digest=sha256:51a6fcb5d69ccaa94d6c9e3fee37742d64f5b8296496a71c4b9fbd9b49fcd52c

Observation 0bc158d1-2a89-435a-9969-ffe37c3f5c86 · inbound

Making Efficient Use of Demonstrations to Solve Hard Exploration Problems cites this paper.

Making Efficient Use of Demonstrations to Solve Hard Exploration Problems Benchmarking Bonus-Based Exploration Methods on the Arcade Learning Environment

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-14T05:24:51.269765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-14T05:24:51.199994Z digest=sha256:2f89c6eee97026ce144c529a5fe3e077bb33342e1a50e8b9fb1888fda31e1727