Pith. sign in

Paper Citation Record · LEDGER

Meta-learning how to Share Credit among Macro-Actions

As of 7 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.13690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.13690 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:33:25.016687Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact2
  • verified fuzzy18
  • unresolved17
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c032ebbc-9c2e-46ee-a361-fe1d88d5b00b · outbound

This paper cites Mas- tering the game of go with deep neural networks and tree search.

Meta-learning how to Share Credit among Macro-Actions Mas- tering the game of go with deep neural networks and tree search

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.784818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.784818Z digest=sha256:5822ad9d7191ef290df285fa8750d2695c77947ac7d757f886cb3ff99fc8dfca

Observation 5293bb1b-0cb6-4997-8349-c0e0c202749c · outbound

This paper cites Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H.

Meta-learning how to Share Credit among Macro-Actions Czarnecki, Michaël Mathieu, Andrew Dudzik, Junyoung Chung, David H

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.815439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.815439Z digest=sha256:ebe1ad5ae82c111dcc9147930c2e1049b5a65d0605390b0433972cedb8a557be

Observation efaa988e-376d-495e-b48f-edcb56b3070a · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Meta-learning how to Share Credit among Macro-Actions Dota 2 with Large Scale Deep Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.820678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.820678Z digest=sha256:0a5783877986a2c905fec9beac3997d6b74c3fd87e363a069fd509b3d0f3d0a6

Observation 1550dea7-2e6d-4b8a-8c3e-723b6dcf7f3e · outbound

This paper cites Autonomous navigation of stratospheric balloons using reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Autonomous navigation of stratospheric balloons using reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.860074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.824827Z digest=sha256:f523b1c2f74126846354c5bd5308db17f45269c845ab87a043361eed372d550a

Observation e134d38b-7f0c-48d9-8251-0e86ca64c058 · outbound

This paper cites Magnetic control of tokamak plasmas through deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Magnetic control of tokamak plasmas through deep reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.618915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.834321Z digest=sha256:eb467d12c60ad757bff1fcaa1a08f226176e97d6ae5f1cad98f985ae4ed6c48c

Observation b97099db-f115-465c-869a-d62877af7863 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:28.416804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.838471Z digest=sha256:cdfcb88ef921b875d363afe6fc3d664c9de4897ad64c6fb4ce3ce7318c9b66a7

Observation 89f20de9-1747-46ab-b6b9-641b87c887f6 · outbound

This paper cites Hierarchical solution of markov decision processes using macro-actions.

Meta-learning how to Share Credit among Macro-Actions Hierarchical solution of markov decision processes using macro-actions

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:28.120610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.847801Z digest=sha256:6d1d73c80b6637e43a87c3713331bdd3f14bed9d03f2174df89d6555d3ef6b40

Observation d911cf51-915f-40ee-a8f0-d1eec8604439 · outbound

This paper cites Fikes and Nils J.

Meta-learning how to Share Credit among Macro-Actions Fikes and Nils J

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.851769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.851769Z digest=sha256:4db58de6487df863cc9e07fba1bc7bec1d920f2e57d766a037bcd5cb31c614ef

Observation 6d934cc5-c425-4162-9b9f-d21b743f5ae3 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:28.020322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.855868Z digest=sha256:083bb3d177dd8c4fd1e8dfdb5c19f2d6051294665364219ceb9a2bba338b514d

Observation 10a3d0ce-5480-461f-abac-770f3db7d5b3 · outbound

This paper cites Durugkar, Clemens Rosenbaum, Stefan Dernbach, and Sridhar Mahadevan.

Meta-learning how to Share Credit among Macro-Actions Durugkar, Clemens Rosenbaum, Stefan Dernbach, and Sridhar Mahadevan

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.845972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.859868Z digest=sha256:5b50504134a54f331be87b96e2e65855674650d225460daab39a8e949b3217fc

Observation bd294195-f6c9-4baa-9985-dd1a27100afa · outbound

This paper cites Rainbow: Combining improve- ments in deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Rainbow: Combining improve- ments in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.642436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.863690Z digest=sha256:5bb1184befca863a8169e5d6d5b9acfc2fa75398a26313be53937daf1a158cfe

Observation 43c626f0-e847-4ef6-b22b-87cd80632f26 · outbound

This paper cites Learning macro-actions in reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Learning macro-actions in reinforcement learning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.346950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.867588Z digest=sha256:5cb6fccfb9cd122e96e23eba1a3237cf1d6bc077edd0dfc291e5a0cf691f72f8

Observation f83086f3-342b-4f35-96e1-961be60390b2 · outbound

This paper cites Macro-actions in reinforcement learning: An empirical analysis.

Meta-learning how to Share Credit among Macro-Actions Macro-actions in reinforcement learning: An empirical analysis

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:27.134155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.871801Z digest=sha256:5d3b70cfefcb1875d9190547c6945754afb9500b3c4bfecd27e835cea92f5a9d

Observation aff8175a-875f-48c2-adaf-321be739ae24 · outbound

This paper cites Meta learning shared hierarchies.

Meta-learning how to Share Credit among Macro-Actions Meta learning shared hierarchies

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.896757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.876353Z digest=sha256:79464905689ae0fe80f69a69c0c95e899929bff4881341b7a151b92c15c10974

Observation e1bf3579-d282-4194-b486-5fac2345c4ce · outbound

This paper cites Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery.

Meta-learning how to Share Credit among Macro-Actions Hierarchical Meta-Reinforcement Learning via Automated Macro-Action Discovery

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.881222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.881222Z digest=sha256:824249d35d8dc9a53eb2234f4fbd2dad00863e45d8ec160e0df24b7eb289acc4

Observation dacd4186-125f-479b-aaf4-cf4c4eaf1a89 · outbound

This paper cites Deep reinforcement learning for decentralized multi-robot exploration with macro actions.

Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning for decentralized multi-robot exploration with macro actions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.767158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.885488Z digest=sha256:627ce356d0b9ad20beafe8d562445a4315f9e49792f896b2be561d0ad1dcf8d0

Observation 70379299-e73c-4cc0-9a69-730e3e2b758f · outbound

This paper cites Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability.

Meta-learning how to Share Credit among Macro-Actions Macro-Action-Based Multi-Agent/Robot Deep Reinforcement Learning under Partial Observability

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.537499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.890230Z digest=sha256:c0b7df367e5772e38d5f989166ef86a635f83185516fb538b665c37a500992b2

Observation cd121df8-6703-47da-97dd-62af7a387c5c · outbound

This paper cites Unlocking new strategies: Intrinsic exploration for evolving macro and micro actions.

Meta-learning how to Share Credit among Macro-Actions Unlocking new strategies: Intrinsic exploration for evolving macro and micro actions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.359326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.893857Z digest=sha256:1f8e8bed252e2b630f1968f5b4e3bc8c2b27781899e4d5c6c2a0686c753ead4d

Observation 08dc6753-272d-4a80-9ee9-df5b85750ee3 · outbound

This paper cites Reusability and Transferability of Macro Actions for Reinforcement Learning.

Meta-learning how to Share Credit among Macro-Actions Reusability and Transferability of Macro Actions for Reinforcement Learning

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:33:25.228396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.897577Z digest=sha256:69eaedfa72ee1a767639793cf8dde21df364c16058c75a9ba9a51391449fc106

Observation ebba5e75-b913-477f-9a69-37e3efbe2511 · outbound

This paper cites Efficient Black-Box Planning Using Macro-Actions with Focused Effects.

Meta-learning how to Share Credit among Macro-Actions Efficient Black-Box Planning Using Macro-Actions with Focused Effects

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:33:25.197756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.901782Z digest=sha256:caa053ec9582d6607e3277503993cb71366ea9f29b5760d832f6c460d457a817

Observation fa82daaf-83d0-41d2-a3f4-192602222500 · outbound

This paper cites Learning macro-actions for arbitrary planners and domains.

Meta-learning how to Share Credit among Macro-Actions Learning macro-actions for arbitrary planners and domains

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.275840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.906030Z digest=sha256:82ff2449b10287056e441b8e744b0e4f83e14273dd5770a820ecb306bc80f4c1

Observation e5866fb1-179f-47e0-9711-492435eb9a86 · outbound

This paper cites Modeling and planning with macro-actions in decentralized pomdps.

Meta-learning how to Share Credit among Macro-Actions Modeling and planning with macro-actions in decentralized pomdps

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.202469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.909481Z digest=sha256:61e742304e537170a4818bb5d667990006a31db0a42cc6d0ba2e48e8d186bbf7

Observation 33e48015-0851-484f-b36e-d3619ed0a3b5 · outbound

This paper cites MAGIC: Learning Macro-Actions for Online POMDP Planning.

Meta-learning how to Share Credit among Macro-Actions MAGIC: Learning Macro-Actions for Online POMDP Planning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.917700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.917700Z digest=sha256:78635889aa7249667ff3855bfc1ad78046e68a0db5c40ef62640b269fff534a0

Observation d8b605d1-db0b-4be6-8d9c-0fecf3c76f89 · outbound

This paper cites Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps.

Meta-learning how to Share Credit among Macro-Actions Deep Reinforcement Learning Based Navigation with Macro Actions and Topological Maps

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.924438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.924438Z digest=sha256:c5d6a28b3b73ae12a0654b94a212ce2edebb1a2b6d6b3bb594eb07cf0e42b5e6

Observation 7b78c45f-3d47-4264-aa80-9701fb320388 · outbound

This paper cites Human-level control through deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Human-level control through deep reinforcement learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.928615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.928615Z digest=sha256:ebc7c8ee0a411926ddf3250a3fe233e40c80dea44bf054393537212309183cfc

Observation 2dd9f7fc-0919-41ec-8840-80b3e8c53152 · outbound

This paper cites Bellemare, Will Dabney, and Rémi Munos.

Meta-learning how to Share Credit among Macro-Actions Bellemare, Will Dabney, and Rémi Munos

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:26.074752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.932262Z digest=sha256:b420e2456d93f00624172cf56007bbb3c4edcb1775e9e3dd47e4da7ca7f52621

Observation 3938a167-c961-4f21-bd21-b3c6b175e0e0 · outbound

This paper cites an unresolved cited work.

Meta-learning how to Share Credit among Macro-Actions Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T00:33:25.914748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.936341Z digest=sha256:aec5727a6c7edf22774a4cc5a769d306f3918929bf2676a9c7e2a6fe70330802

Observation e52ef023-d402-4c67-9239-969339b33889 · outbound

This paper cites Prioritized Experience Replay.

Meta-learning how to Share Credit among Macro-Actions Prioritized Experience Replay

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.940102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.940102Z digest=sha256:7ee5074a03c7216b2a7c8ae04fa63273f64db7580cd5286479b1effc43430ba8

Observation 23c92e8c-3033-433a-b610-33cdbee5ce27 · outbound

This paper cites Deep reinforcement learning with double q-learning.

Meta-learning how to Share Credit among Macro-Actions Deep reinforcement learning with double q-learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.944182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.944182Z digest=sha256:4e91750e6ea1bbcb1841497ceb7bbf3b664d95a0209b5190e4d8e35338c88959

Observation 32133ac0-1f1e-44bb-ac77-14e07dfc6df8 · outbound

This paper cites Dueling network architectures for deep reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Dueling network architectures for deep reinforcement learning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.948639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.948639Z digest=sha256:5afef57cab0bedf24a6fc1f1710e6f8e8fa179ed359414c12f45d2504a88cd1f

Observation da0d07e9-62a7-43c6-8d8d-cbae40b1b60c · outbound

This paper cites Noisy Networks for Exploration.

Meta-learning how to Share Credit among Macro-Actions Noisy Networks for Exploration

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.952368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.952368Z digest=sha256:6a0bc8625f803f1c2a85eb2cb6620880590a675f271ea3a1433cd74373c0a796

Observation 783e554d-9d9f-441d-8ac5-76e368f69599 · outbound

This paper cites Meta-gradient reinforcement learning.

Meta-learning how to Share Credit among Macro-Actions Meta-gradient reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.665399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.956996Z digest=sha256:f8362d4d319113f83c65b52a30091b99ef1355e7808e272de5990d81d1402bc1

Observation 73968742-7555-49a2-9de1-b2ba74f34a83 · outbound

This paper cites Universal value function approxima- tors.

Meta-learning how to Share Credit among Macro-Actions Universal value function approxima- tors

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.559712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.967341Z digest=sha256:56fa7d7f8c70b855c3b65b23a3ce63059d41d2a1b7c557374418e49a4491f4f7

Observation 288709fa-0c4b-49f7-b2d3-d811e4c5cf4f · outbound

This paper cites The arcade learning environment: An evaluation platform for general agents.

Meta-learning how to Share Credit among Macro-Actions The arcade learning environment: An evaluation platform for general agents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.972757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.972757Z digest=sha256:bf731a25dff9426b78cbf486fed3a3a9829397c0274c86a60aa30b0443016d58

Observation a756d7de-1916-4c1b-a497-9b7424264729 · outbound

This paper cites OpenAI Gym.

Meta-learning how to Share Credit among Macro-Actions OpenAI Gym

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:24.980907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:24.980907Z digest=sha256:2c67cef36fd5b022ed6364cc4f6b6c43d55781517171079b0227a9682166a417

Observation 4a619404-64c9-49f3-990e-b2e22582241b · outbound

This paper cites The Atari Grand Challenge Dataset.

Meta-learning how to Share Credit among Macro-Actions The Atari Grand Challenge Dataset

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:33:25.057740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:24.996820Z digest=sha256:1c67461e408037470c107e2f53c51ff60667397f0e61b60e413897cc4b469412

Observation 75bbc518-cc62-4693-b51f-183392014552 · outbound

This paper cites Gym-minigrid: Minimalistic gridworld environment for openai gym.

Meta-learning how to Share Credit among Macro-Actions Gym-minigrid: Minimalistic gridworld environment for openai gym

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.504282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:25.001724Z digest=sha256:88427ff896d23d31fa3bb0c4f769c5dcb596fa871cc87f98703e905542948942

Observation bfe23e56-c531-443b-a03e-b43f469d379a · outbound

This paper cites - Perform a standard TD update with the MASP penalty, updating θ → θ′ using Σ fixed.

Meta-learning how to Share Credit among Macro-Actions - Perform a standard TD update with the MASP penalty, updating θ → θ′ using Σ fixed

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:33:25.489509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:25.005793Z digest=sha256:4e6d7c6f4c5dcb110b15f5371a4b04d58d08ba0166a19940f7d142d2414a018e

Observation e3f33879-12df-4e55-8281-c15b9750d088 · outbound

This paper cites - Evaluate the performance of the updated θ′ using a meta-objective (the standard TD loss).

Meta-learning how to Share Credit among Macro-Actions - Evaluate the performance of the updated θ′ using a meta-objective (the standard TD loss)

Reference 39

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:33:25.449892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T00:33:25.016687Z digest=sha256:33f4dc179edeb71ab4b4ac15d504242dc2a5832e9c5018352b68c95146b9ce68

Pith citing papers

No inbound Pith citation observations are available.