Pith. sign in

Paper Citation Record · LEDGER

Hierarchical Reinforcement Learning with Targeted Causal Interventions

As of 7 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2507.04373.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04373 v1

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:59:01.602504Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved8
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bcd355d3-c603-4b23-a5b9-f8a40174012e · outbound

This paper cites For” loop, it records a trajectory τj. The “While.

Hierarchical Reinforcement Learning with Targeted Causal Interventions For” loop, it records a trajectory τj. The “While

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:06.021733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.566592Z digest=sha256:37cd528bebd89431bcf782a43cfd651d08ee3002340e38509f14cddaab1de2a8

Observation 95ea9249-855c-411c-a23e-662050c6afee · outbound

This paper cites As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n.

Hierarchical Reinforcement Learning with Targeted Causal Interventions As we mentioned earlier, q1 is the upper bound on the expected number of the nodes that are ancestors of the node n

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.944378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.618392Z digest=sha256:deffa63fc12052f8e7d81225d527b9044e34568d691ceb62ed53134680a49675

Observation 69954592-9938-4d56-bf2e-2c8dc644aadb · outbound

This paper cites At this point, we have constructed Ijr.

Hierarchical Reinforcement Learning with Targeted Causal Interventions At this point, we have constructed Ijr

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.518464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.825900Z digest=sha256:b6d8e1e12bcf30ef3b44ffd321dab1f2411bfdf1124c39227ea5247d4d5b66bb

Observation 3c607cc7-9191-47d2-aed3-1feccd0c5f42 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.057351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.670188Z digest=sha256:0a3e541d3661f6b883f50812c97cd36f1ae78fb9294e05bc93fc1343b22b4bfc

Observation edde5922-6480-4b98-ab29-786f37d13290 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.828947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.661189Z digest=sha256:1b263d3375d18a7bb01720987d3bed541c6cabe7e503585ed604ee8f5b025762

Observation 42e29f81-69a3-43a8-b12a-03053e05786f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:05.693361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.724253Z digest=sha256:cd3e1865e11c91d7bf269048db6f8b089d3a42b19b0dc243707cbdf55c6fbecb

Observation fa4077de-a00f-4075-9185-2e15929efd7f · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:05.217088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:58:59.911931Z digest=sha256:c9bc6af86e9c4a58084766aee8db180671775752842e5f268e5b403195d56b03

Observation 9e8f781a-6a32-4a70-8290-de481f5cb137 · outbound

This paper cites • If gi is the goal, backtrack to construct the path.

Hierarchical Reinforcement Learning with Targeted Causal Interventions • If gi is the goal, backtrack to construct the path

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.993906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.025603Z digest=sha256:e8dcb51f94963f9b7879829060e0bbf66bb58ee13e83980a5862e89fc220fc2e

Observation 0fdd8088-47d8-400b-8192-c94b8819f64e · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.830033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.125185Z digest=sha256:af0fcf0e73f8b8a5d3403aff7240e443e72711461c7f7169364c1b1802abb774

Observation 1d7564ee-af70-4be3-a534-1afc40a97600 · outbound

This paper cites We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi.

Hierarchical Reinforcement Learning with Targeted Causal Interventions We define a function parent pointer : Φ → Φ, which assigns the corresponding parent of a node gi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.636736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.267033Z digest=sha256:bf698e0558e78d848f1fca7c3cedb76e13a71bb28074d564810dc7762021528d

Observation 85b7b82b-c5fa-4cc4-a2ce-39d18e44cbd0 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:04.475273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.396920Z digest=sha256:25686938176e5fd465e9fce84cd1e64a42d0035a13da3e5891980516d6085a9c

Observation 6a16f74e-6055-4dfa-af95-128e97919592 · outbound

This paper cites – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions – For each neighbor gv of gu in ˆGt: * Set w(gu, gv) = |C HˆGt gu \ back track(gu)| + 1

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:04.279653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.547872Z digest=sha256:66dfe94bb09736a4886be91afa8e33b33dfdf93e38986c1a18c045471a8b038f

Observation 8eb58f2e-b6e1-4739-badf-4d54b33deb9d · outbound

This paper cites This implies that in all possible assignments, at least one parent of gi is always 1.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This implies that in all possible assignments, at least one parent of gi is always 1

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.743170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.788274Z digest=sha256:e29ebd3d53b0f473e045d2c1c0ce6612d7d9f4cc8d3ce51e0544bb4b159bb16e

Observation a35dfc16-e6ec-4fd6-b17a-0a9593aa35ec · outbound

This paper cites This means that in the collected data, if Xj is one, at least of some other parent Xk is also one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that in the collected data, if Xj is one, at least of some other parent Xk is also one

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.450693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:00.909216Z digest=sha256:cb9d56a8cff32768ee066b5fe475397c4338905dc48c6b84ab95efc60034916a

Observation d5db77dc-54b7-4b3e-bc67-5d6e4c1cdc69 · outbound

This paper cites Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Thus, the AND operator equals to 0: θi(Xt) = ^ gj ∈PAgi X t j = 0

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:03.115544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.009013Z digest=sha256:2510e56414b61ac8515f7acb4f6ea5e8d77573d865180125e8b3c5f27dead1c2

Observation a059b745-1826-4392-b327-b64a4db82a01 · outbound

This paper cites This means that if Xj is zero, at least of some other parent Xk is one.

Hierarchical Reinforcement Learning with Targeted Causal Interventions This means that if Xj is zero, at least of some other parent Xk is one

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:02.908195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.140753Z digest=sha256:ac7a4138ba49ecd499c9bcf0e5923510d84616f4ca0dc33a656edc62f3cd811f

Observation 5cc87883-af28-4119-a045-0c75be5b4e8f · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 19

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.610404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.268863Z digest=sha256:6661b2c70d4a9a6df9ad8ebdf22b35a395d67248b2bc7935e7bb8a6abda9f0e1

Observation 8419e1d6-20e4-4c15-b4a2-176bb38a79d9 · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:59:02.369210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.396496Z digest=sha256:709ea899e60f030bdb86383dbdaab816737ac9b572b5643a4a569268df95ef34

Observation 6bfd4648-1748-4764-9424-c2d5a7829dae · outbound

This paper cites an unresolved cited work.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Unresolved cited work

Reference 21

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T19:59:02.158132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.490528Z digest=sha256:8ed8640a1cad4f0f44d559d94df5eb6fb905cedba1cb25c08619fb4e4cb616d2

Observation c33044bd-59e3-4cb9-aa43-89263fdf337d · outbound

This paper cites oracle goal space.

Hierarchical Reinforcement Learning with Targeted Causal Interventions oracle goal space

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:59:01.882505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T19:59:01.602504Z digest=sha256:c2f1de676e225765b278483fc33f3de1b4b5ca68e0d325907bfbc364de750489

Observation 69459973-4e23-4b37-8620-6369ffb101a3 · outbound

This paper cites MineRL: A Large-Scale Dataset of Minecraft Demonstrations.

Hierarchical Reinforcement Learning with Targeted Causal Interventions MineRL: A Large-Scale Dataset of Minecraft Demonstrations

Reference 1528

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.435876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.435876Z digest=sha256:2774b4ae318662220d85e0f23dcc3a125d7d11495cdc827ba00a8fe28255d75c

Observation 9faebd17-940c-4b85-a192-68f2b8d61b45 · outbound

This paper cites Learning Neural Causal Models from Unknown Interventions.

Hierarchical Reinforcement Learning with Targeted Causal Interventions Learning Neural Causal Models from Unknown Interventions

Reference 3119

Resolution
unresolved
no resolver link, observed 2026-08-06T19:58:59.513260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:58:59.513260Z digest=sha256:3366e629e394720b4230592aae491398bcad2cd4685c65c21eb92b6407fd62be

Pith citing papers

No inbound Pith citation observations are available.