Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:59:16.541167Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2507.01470.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:59:16.541167Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
33 of 33 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7d4da9db-1c46-43ae-a885-322a86a7f863 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Feudal Multi-Agent Hierarchies for Cooperative Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99c0bd26-131a-419b-b279-4e8d9470bc95 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Concrete Problems in AI Safety
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8836cee5-a8d6-4999-a299-a65dc61cdfee · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Hindsight experience replay
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 14ad7b67-bf52-4977-83fa-a00466b041a1 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Dynamic programming
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4fb16a7-b28d-410a-9931-4d125ab8aa59 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Exploration by Random Network Distillation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc377857-9684-4efc-b592-66816d3ead45 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bfdeb87-e46b-4126-ac23-e2f7d48435d9 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals HiSOMA : A hierarchical multi-agent model integrating self-organizing neural networks with multi-agent deep reinforcement learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dde65c51-d02b-4875-bbc3-adef23066d2a · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals MASER : Multi-agent reinforcement learning with subgoals generated from experience replay buffer
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d311e481-1434-40ba-9ab9-c61e70c41aed · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Automatic discovery of subgoals in reinforcement learning using strongly connected components
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 467cdcee-7649-43a1-8f56-49382aafe968 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Exploration in deep reinforcement learning: A survey
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c6a3b94-79d5-418b-9334-34cc91bef9c4 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Lecun, L
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9ce9363-1912-4b43-b219-5e02f2885fd2 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Automatic discovery of subgoals in reinforcement learning using diverse density
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d82be80-cdee-46ff-818b-67cb0c575fd3 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Research on Multi -agent Sparse Reward Problem
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a082f66-e373-4a5f-a84a-68e4806cc08a · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Laser learning environment: A new environment for coordination-critical multi-agent tasks
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2c60f28c-eda0-4176-a4fa-434b66a51ee5 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals An overview of environmental features that impact deep reinforcement learning in sparse-reward domains
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ddab0dbb-b9fd-4f0d-800b-2f05df26db66 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Efros, and Trevor Darrell
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c7f1459-afb7-4eca-8ca3-3382338a08f2 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Learning to Drive a Bicycle using Reinforcement Learning and Shaping
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cbeecba4-df41-4b81-83a4-5ee1fa5b4c80 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals QMIX: Monotonic Value Function Factorisation for Deep Multi-Agent Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3724e73d-6be2-4af4-8957-7a144c55dbcc · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals The StarCraft Multi-Agent Challenge
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a62b7a4-3661-41f7-885b-ab7abe8a9ad7 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Normalized cuts and image segmentation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f0da5f00-7721-4c86-8dcd-a8d99e8b560c · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Wolfe, and Andrew G
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082ccee4-2d28-4ccb-ab72-b3b636db67bd · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Leibo, Karl Tuyls, and Thore Graepel
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6775740b-eb04-4378-a2ee-978b43fb2dca · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Faster MIL -based subgoal identification for reinforcement learning by tuning fewer hyperparameters
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 25b4aadd-313e-4b70-a586-c290970713a2 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Sutton and Andrew G
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 308c9bdf-4d87-4580-a124-1fbafbe0520f · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Sutton, Doina Precup, and Satinder Singh
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c91e146-c36f-40b5-a796-6d9299277dbe · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals \#exploration: A study of count-based exploration for deep reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8a892e58-5ae4-4b7c-bf72-9a0072e396e2 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Keeping your distance: Solving sparse reward tasks using self-balancing shaped rewards
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26d0f6c0-c6e6-46bb-b059-8b1e0516bb2b · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Deep Reinforcement Learning with Double Q-learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e0992bd-c580-4fd7-870d-383f6305ef15 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals QPLEX: Duplex Dueling Multi-Agent Q-Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2df939bb-a422-41e5-a98d-c16d404ec35e · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f4be7ec-1b16-4267-b6c5-7c5eba4fd65b · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals HAVEN : Hierarchical cooperative multi-agent reinforcement learning with dual coordination mechanism
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 23accf97-5d55-4607-9472-1b9d15576a60 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals Ng, Daishi Harada, and Stuart Russell
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff4fd410-ad6b-4afd-aa4c-6643d5dac742 · outbound
Zero-Incentive Dynamics: a look at reward sparsity through the lens of unrewarded subgoals write newline
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.