Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:1611.05397.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T04:39:32.069321Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-25T19:06:08.995943Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 14e92d04-108f-4fd2-a2a6-f4964f1f82a5 · inbound
Continual Reinforcement Learning with Diversity Exploration and Adversarial Self-Correction Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 517d1d1e-4190-4c07-a335-bb31cf6a7801 · inbound
Shaping Belief States with Generative Environment Models for RL Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e3b90c70-0a11-426c-8aac-9df65f5d3079 · inbound
Learning Belief Representations for Imitation Learning in POMDPs Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bd93b13b-0456-4e43-a530-30037837c8d5 · inbound
Supervise Thyself: Examining Self-Supervised Representations in Interactive Environments Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 775dfe4c-c8fc-4153-8d4d-82d6f7450e9d · inbound
To Learn or Not to Learn: Analyzing the Role of Learning for Navigation in Virtual Environments Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c6802c97-daf4-4b12-9d7e-9a0ff4941d5d · inbound
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 49cb5872-2991-401c-a29a-11d762934e06 · inbound
Dream to Control: Learning Behaviors by Latent Imagination Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c2d5b451-7c76-412d-b732-bd000dfd64d1 · inbound
Mastering Diverse Domains through World Models Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cbf16847-a4e2-4bc1-a597-2b1efb5174f8 · inbound
Hierarchical Successor Representation for Robust Transfer Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc2f41c3-69d3-4b2a-8f8f-63ef9d1ff257 · inbound
Reliability-Aware Geometric Fusion for Robust Audio-Visual Navigation Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7aff776e-1348-43dc-a87f-98a81b38dd9f · inbound
Reflective Context Learning: Studying the Optimization Primitives of Context Space Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation eb595bab-a188-4494-a781-5cd0db995f48 · inbound
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 80e987fe-a219-4022-a0f3-c70591316c19 · inbound
Learning to Theorize the World from Observation Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8eea4d83-7326-4c82-a23f-34cf9df350d7 · inbound
Goal-Conditioned Agents that Learn Everything All at Once Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 30a8ef44-1c82-405a-ab56-d557f2d2ea0d · inbound
When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c47536ad-d400-4540-b82a-cb11a388130e · inbound
PAMD: Structured Adaptive Distances for Bisimulation Representations in Visual Reinforcement Learning Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e4a8bd-41a0-4f7a-87c1-794798b2141c · inbound
TAPO: Transition-Aware Policy Optimization for LLM Agents Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8024fd77-6ac0-4172-b553-47579b855246 · inbound
LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Reinforcement Learning with Unsupervised Auxiliary Tasks
Reference 228
Source-reported events for the cited work
Unavailable: canonical work link unavailable.