Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:10:29.525354Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 0 inbound Pith citation observations for arXiv:2506.22566.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:10:29.525354Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
21 of 21 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 4f742340-d001-4ecb-ad71-3213a896ec27 · outbound
Exploration Behavior of Untrained Policies Lipbab: Computing exact lipschitz constant of relu networks
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5e3a4c18-c91a-436b-a530-2607ec7bbadc · outbound
Exploration Behavior of Untrained Policies Radial basis functions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c93c7a78-3e4a-4457-b0fe-9f056c9059e1 · outbound
Exploration Behavior of Untrained Policies Exploration by Random Network Distillation
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aea7d39d-5684-45d1-a88a-db50d36e9520 · outbound
Exploration Behavior of Untrained Policies Rainbow: Combining im- provements in deep reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 921bc39c-40bb-4719-a0bf-387044182b2c · outbound
Exploration Behavior of Untrained Policies Neural tangent kernel: Convergence and generalization in neural networks
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4a01a4cc-b16d-411d-83f0-78a7da33c1a8 · outbound
Exploration Behavior of Untrained Policies Exploration in deep reinforce- ment learning: A survey
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 648b2eca-f65a-4e42-aa9b-c8e86a180c32 · outbound
Exploration Behavior of Untrained Policies Lipschitz constant estimation of Neural Networks via sparse polynomial optimization
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1a691914-3af9-435c-8dab-fd76864640ee · outbound
Exploration Behavior of Untrained Policies Flipping coins to estimate pseudocounts for exploration in reinforcement learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 24c8ee25-e0a5-4cff-86a8-0d3e5110a3b3 · outbound
Exploration Behavior of Untrained Policies Periodic activation functions induce stationar- ity
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a0542598-25d5-440f-9006-12808fac717e · outbound
Exploration Behavior of Untrained Policies Human-level control through deep reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3830a1c-008c-4c0e-ae77-39efec3bb782 · outbound
Exploration Behavior of Untrained Policies Bayesian learning for neural networks , volume 118
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff014368-7c2d-4e38-a154-f1cf0c515db3 · outbound
Exploration Behavior of Untrained Policies The primacy bias in deep reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 744ac8ad-46de-4499-bc53-31ac34d2eb37 · outbound
Exploration Behavior of Untrained Policies Deep reinforcement learning with plasticity injection
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e483f8f7-c5b7-41e9-b252-789164ae8490 · outbound
Exploration Behavior of Untrained Policies Deep exploration via bootstrapped dqn
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3c452bac-3594-4190-bf34-11c4375a7721 · outbound
Exploration Behavior of Untrained Policies Trust region policy optimization
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1cf297be-0c97-490d-ab4d-efcd470d9e22 · outbound
Exploration Behavior of Untrained Policies Proximal Policy Optimization Algorithms
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 419d81e5-aa47-4426-b564-2a250826ae5e · outbound
Exploration Behavior of Untrained Policies On Bonus-Based Exploration Methods in the Arcade Learning Environment
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 22c9fc1d-d33a-44f0-a281-3d3b3fef7bf1 · outbound
Exploration Behavior of Untrained Policies Lipschitz regularity of deep neural networks: analysis and efficient estimation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aec46387-79f1-45fe-b2f3-ce3a690111e0 · outbound
Exploration Behavior of Untrained Policies Q-learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c5b8af5f-2bc5-42f1-8abf-0e501571d7c2 · outbound
Exploration Behavior of Untrained Policies Simple statistical gradient-following algorithms for connectionist rein- forcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fbe3f23d-dab5-49d1-b34c-e2efb76b71c6 · outbound
Exploration Behavior of Untrained Policies Neural Architecture Search with Reinforcement Learning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.