Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:33:25.016687Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.13690.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:33:25.016687Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c032ebbc-9c2e-46ee-a361-fe1d88d5b00b · outbound
Meta-learning how to Share Credit among Macro-Actions Mas- tering the game of go with deep neural networks and tree search
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5293bb1b-0cb6-4997-8349-c0e0c202749c · outbound
Meta-learning how to Share Credit among Macro-Actions Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efaa988e-376d-495e-b48f-edcb56b3070a · outbound
Meta-learning how to Share Credit among Macro-Actions Dota 2 with Large Scale Deep Reinforcement Learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1550dea7-2e6d-4b8a-8c3e-723b6dcf7f3e · outbound
Meta-learning how to Share Credit among Macro-Actions Autonomous navigation of stratospheric balloons using reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e134d38b-7f0c-48d9-8251-0e86ca64c058 · outbound
Meta-learning how to Share Credit among Macro-Actions Magnetic control of tokamak plasmas through deep reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b97099db-f115-465c-869a-d62877af7863 · outbound
Meta-learning how to Share Credit among Macro-Actions Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 89f20de9-1747-46ab-b6b9-641b87c887f6 · outbound
Meta-learning how to Share Credit among Macro-Actions Hierarchical solution of markov decision processes using macro-actions
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d911cf51-915f-40ee-a8f0-d1eec8604439 · outbound
Meta-learning how to Share Credit among Macro-Actions Fikes and Nils J
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d934cc5-c425-4162-9b9f-d21b743f5ae3 · outbound
Meta-learning how to Share Credit among Macro-Actions Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 10a3d0ce-5480-461f-abac-770f3db7d5b3 · outbound
Meta-learning how to Share Credit among Macro-Actions Durugkar, Clemens Rosenbaum, Stefan Dernbach, and Sridhar Mahadevan
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bd294195-f6c9-4baa-9985-dd1a27100afa · outbound
Meta-learning how to Share Credit among Macro-Actions Rainbow: Combining improve- ments in deep reinforcement learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 43c626f0-e847-4ef6-b22b-87cd80632f26 · outbound
Meta-learning how to Share Credit among Macro-Actions Learning macro-actions in reinforcement learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f83086f3-342b-4f35-96e1-961be60390b2 · outbound
Meta-learning how to Share Credit among Macro-Actions Macro-actions in reinforcement learning: An empirical analysis
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation aff8175a-875f-48c2-adaf-321be739ae24 · outbound
Meta-learning how to Share Credit among Macro-Actions Meta learning shared hierarchies
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e1bf3579-d282-4194-b486-5fac2345c4ce · outbound
Meta-learning how to Share Credit among Macro-Actions Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dacd4186-125f-479b-aaf4-cf4c4eaf1a89 · outbound
Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning for decentralized multi-robot exploration with macro actions
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70379299-e73c-4cc0-9a69-730e3e2b758f · outbound
Meta-learning how to Share Credit among Macro-Actions Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cd121df8-6703-47da-97dd-62af7a387c5c · outbound
Meta-learning how to Share Credit among Macro-Actions Unlocking new strategies: Intrinsic exploration for evolving macro and micro actions
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08dc6753-272d-4a80-9ee9-df5b85750ee3 · outbound
Meta-learning how to Share Credit among Macro-Actions Reusability and Transferability of Macro Actions for Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ebba5e75-b913-477f-9a69-37e3efbe2511 · outbound
Meta-learning how to Share Credit among Macro-Actions Efficient Black-Box Planning Using Macro-Actions with Focused Effects
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fa82daaf-83d0-41d2-a3f4-192602222500 · outbound
Meta-learning how to Share Credit among Macro-Actions Learning macro-actions for arbitrary planners and domains
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5866fb1-179f-47e0-9711-492435eb9a86 · outbound
Meta-learning how to Share Credit among Macro-Actions Modeling and planning with macro-actions in decentralized pomdps
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 33e48015-0851-484f-b36e-d3619ed0a3b5 · outbound
Meta-learning how to Share Credit among Macro-Actions MAGIC: Learning Macro-Actions for Online POMDP Planning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8b605d1-db0b-4be6-8d9c-0fecf3c76f89 · outbound
Meta-learning how to Share Credit among Macro-Actions Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b78c45f-3d47-4264-aa80-9701fb320388 · outbound
Meta-learning how to Share Credit among Macro-Actions Human-level control through deep reinforcement learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2dd9f7fc-0919-41ec-8840-80b3e8c53152 · outbound
Meta-learning how to Share Credit among Macro-Actions Bellemare, Will Dabney, and Rémi Munos
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3938a167-c961-4f21-bd21-b3c6b175e0e0 · outbound
Meta-learning how to Share Credit among Macro-Actions Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e52ef023-d402-4c67-9239-969339b33889 · outbound
Meta-learning how to Share Credit among Macro-Actions Prioritized Experience Replay
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c92e8c-3033-433a-b610-33cdbee5ce27 · outbound
Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning with double q-learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32133ac0-1f1e-44bb-ac77-14e07dfc6df8 · outbound
Meta-learning how to Share Credit among Macro-Actions Dueling network architectures for deep reinforcement learning
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da0d07e9-62a7-43c6-8d8d-cbae40b1b60c · outbound
Meta-learning how to Share Credit among Macro-Actions Noisy Networks for Exploration
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783e554d-9d9f-441d-8ac5-76e368f69599 · outbound
Meta-learning how to Share Credit among Macro-Actions Meta-gradient reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73968742-7555-49a2-9de1-b2ba74f34a83 · outbound
Meta-learning how to Share Credit among Macro-Actions Universal value function approxima- tors
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 288709fa-0c4b-49f7-b2d3-d811e4c5cf4f · outbound
Meta-learning how to Share Credit among Macro-Actions The arcade learning environment: An evaluation platform for general agents
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a756d7de-1916-4c1b-a497-9b7424264729 · outbound
Meta-learning how to Share Credit among Macro-Actions OpenAI Gym
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a619404-64c9-49f3-990e-b2e22582241b · outbound
Meta-learning how to Share Credit among Macro-Actions The Atari Grand Challenge Dataset
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 75bbc518-cc62-4693-b51f-183392014552 · outbound
Meta-learning how to Share Credit among Macro-Actions Gym-minigrid: Minimalistic gridworld environment for openai gym
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bfe23e56-c531-443b-a03e-b43f469d379a · outbound
Meta-learning how to Share Credit among Macro-Actions - Perform a standard TD update with the MASP penalty, updating θ → θ′ using Σ fixed
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e3f33879-12df-4e55-8281-c15b9750d088 · outbound
Meta-learning how to Share Credit among Macro-Actions - Evaluate the performance of the updated θ′ using a meta-objective (the standard TD loss)
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.