Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning with Targeted Causal Interventions

As of 7 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2507.04373.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04373 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:01.602504Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved8
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bcd355d3-c603-4b23-a5b9-f8a40174012e · outbound

This paper cites For” loop, it records a trajectory τj. The “While.

Hierarchical Reinforcement Learning with Targeted Causal Interventions For” loop, it records a trajectory τj. The “While

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:06.021733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.566592Z digest=sha256:ad9870d202af24e181d3a4ebcae1623d6065aa4be6fd2dd3ab570277fa587f86

Observation 95ea9249-855c-411c-a23e-662050c6afee · outbound

This paper cites As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n.

Hierarchical Reinforcement Learning with Targeted Causal Interventions As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.944378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.618392Z digest=sha256:303b04143903dbc25f466f99db8214601f847dc6faba1895b9654386c2f99a62

Observation 69954592-9938-4d56-bf2e-2c8dc644aadb · outbound

This paper cites At this point, we have constructed Ijr.

Hierarchical Reinforcement Learning with Targeted Causal Interventions At this point, we have constructed Ijr

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.518464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.825900Z digest=sha256:dab57d5f7a20d732fc562a4730c56b5700f99b95269d662a88c427008fbd13d8

Observation 3c607cc7-9191-47d2-aed3-1feccd0c5f42 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.057351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.670188Z digest=sha256:8a4a64af4f76ffbc35fd62e0fbf7f6cbaa23be80a4af03c33fd0ee87cfd76511

Observation edde5922-6480-4b98-ab29-786f37d13290 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.828947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.661189Z digest=sha256:6e0dc42c0f073b4b71d8b5730b7fea85295f0805b047c5978eb82cfe5a12915e

Observation 42e29f81-69a3-43a8-b12a-03053e05786f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.693361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.724253Z digest=sha256:291d9d3d467156f1cd37cba3069c809820c1934a1251f87a667c70f7205d5d55

Observation fa4077de-a00f-4075-9185-2e15929efd7f · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.217088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:58:59.911931Z digest=sha256:3243737afec73af0bf4007a67176ad363d9e4ad569aaaa77c89c5bddbd0c782c

Observation 9e8f781a-6a32-4a70-8290-de481f5cb137 · outbound

This paper cites • If gi is the goal, backtrack to construct the path.

Hierarchical Reinforcement Learning with Targeted Causal Interventions • If gi is the goal, backtrack to construct the path

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.993906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.025603Z digest=sha256:968fb9821422754f63809250e8c19c53ef87492280636c88a8470efe80c1d6dd

Observation 0fdd8088-47d8-400b-8192-c94b8819f64e · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.830033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.125185Z digest=sha256:9f1d13d071d750e1c556587dcde72acb59f5ff69bcc2a9f9d4804f88928594b9

Observation 1d7564ee-af70-4be3-a534-1afc40a97600 · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.636736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.267033Z digest=sha256:485b52a6cdc5fe9e727575d7b0988d122fde2fa396b21f29fde3c55d0e8bbd83

Observation 85b7b82b-c5fa-4cc4-a2ce-39d18e44cbd0 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.475273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.396920Z digest=sha256:fad8e382d8cbcfe50b8780d2413c452bd7f47c988552025e37a0ceab0dbe2e01

Observation 6a16f74e-6055-4dfa-af95-128e97919592 · outbound

This paper cites – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.279653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.547872Z digest=sha256:4980f8ddb947c2c79d02754e5a5805c1be3e52c01651d0f41718bf07948d6ee2

Observation 8eb58f2e-b6e1-4739-badf-4d54b33deb9d · outbound

This paper cites This implies that in all possible assignments, at least one parent of gi is always 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This implies that in all possible assignments, at least one parent of gi is always 1

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.743170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.788274Z digest=sha256:5893d5b86ae8128257d671f9f48c2e5600edd81062eeb05703226af193873bca

Observation a35dfc16-e6ec-4fd6-b17a-0a9593aa35ec · outbound

This paper cites This means that in the collected data, if Xj is one, at least of some other parent Xk is also one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that in the collected data, if Xj is one, at least of some other parent Xk is also one

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.450693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:00.909216Z digest=sha256:23512bcd07f2e0f8027f7860668043f5bd82d1512b9bc202d5c6e1ebe86e29f3

Observation d5db77dc-54b7-4b3e-bc67-5d6e4c1cdc69 · outbound

This paper cites Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.115544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.009013Z digest=sha256:3168f4a06cf776cbe22245755a20e67fb1951cafb7422cdfbfb11e8f7ee17181

Observation a059b745-1826-4392-b327-b64a4db82a01 · outbound

This paper cites This means that if Xj is zero, at least of some other parent Xk is one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that if Xj is zero, at least of some other parent Xk is one

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:02.908195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.140753Z digest=sha256:ec47e973e964669d97b280dd5db3ac69339fc04e761f971587145ad576d39aa1

Observation 5cc87883-af28-4119-a045-0c75be5b4e8f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 19

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.610404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.268863Z digest=sha256:5cde6d1362a95983e6e0bc12ddff7111981f712ee1017beb77e8aa198e839619

Observation 8419e1d6-20e4-4c15-b4a2-176bb38a79d9 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:02.369210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.396496Z digest=sha256:672fdf4516255a7bb5bee1c0e790f7fcea32a118aca0dee2c063014eb6bcf680

Observation 6bfd4648-1748-4764-9424-c2d5a7829dae · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 21

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.158132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.490528Z digest=sha256:1613402bb8bd896719765c36ae95929aad100f5bb1a94cb3ee00e1d37717a348

Observation c33044bd-59e3-4cb9-aa43-89263fdf337d · outbound

This paper cites oracle goal space.

Hierarchical Reinforcement Learning with Targeted Causal Interventions oracle goal space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:01.882505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T19:59:01.602504Z digest=sha256:f86236b8bb2ba3bb21f5cdd7e07eb988e85deb02d507bb87fc867313c1dcd0a6

Observation 69459973-4e23-4b37-8620-6369ffb101a3 · outbound

This paper cites MineRL: A Large-Scale Dataset of Minecraft Demonstrations.

Hierarchical Reinforcement Learning with Targeted Causal Interventions MineRL: A Large-Scale Dataset of Minecraft Demonstrations

Reference 1528

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.435876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.435876Z digest=sha256:2774b4ae318662220d85e0f23dcc3a125d7d11495cdc827ba00a8fe28255d75c

Observation 9faebd17-940c-4b85-a192-68f2b8d61b45 · outbound

This paper cites Learning Neural Causal Models from Unknown Interventions.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Learning Neural Causal Models from Unknown Interventions

Reference 3119

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.513260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.513260Z digest=sha256:b65fb08e6da9576cc7d1b457e8f12ff5958bbcf3dfcc2b427375522e49676d5e

Pith citing papers

No inbound Pith citation observations are available.