Pith. sign in

Paper Citation Record · LEDGER

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies

As of 20 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2507.14901.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.14901 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:51:37.495684Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T01:13:11.483599Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:38:55.871103Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy34
  • unresolved6
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c996a0ed-63c1-43bc-acea-b116e46d4fae · outbound

This paper cites Mastering the game of Go with deep neural networks and tree search.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Mastering the game of Go with deep neural networks and tree search

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.350707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.266533Z digest=sha256:86a215bb21c24f035f32ee5e3f5ec21ae4c73bf1f272cc94ed014e98d794f680

Observation d3a92e27-364d-4403-8015-737a8d3a7b7d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Playing Atari with Deep Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.272028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.272028Z digest=sha256:3dcd04e4bbc7ba3ef4778bd2836dd8635c2cdb5bc6a65f1c73b49bc1b8434461

Observation 13efe497-c48a-4eb8-b7f2-aa9cd992a0a9 · outbound

This paper cites Reinforcement learning in robotics: A survey.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning in robotics: A survey

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.330692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.278064Z digest=sha256:3f2ae7858f45692f91a725357128c8070f81fee751997021e9e9512ecafdcd4d

Observation 9c5ddcdc-76b8-4b30-b285-98c9dcfb7fc1 · outbound

This paper cites Resource management with deep reinforcement learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Resource management with deep reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.302120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.283822Z digest=sha256:39a55a74c4f1b7646582475b502341bc29416981dd1c9b3e6e95312eda151f42

Observation f6bc40c8-296c-41b7-831f-8d63d4b72b9c · outbound

This paper cites End to End Learning for Self-Driving Cars.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies End to End Learning for Self-Driving Cars

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.289123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.289123Z digest=sha256:a68e287b72344cf3bf4dcc52c8b3ebd131f13df741ebb33a0079149dd6982e40

Observation 5f3ee64d-9fcd-4e88-934d-8000b51c2f17 · outbound

This paper cites Reinforcement learning based recommender systems: A survey.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning based recommender systems: A survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.277061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.294637Z digest=sha256:e159b20fafb85c09b6c71b0436798ff96c7630737d0f626f62354e724adcb217

Observation b8c337f6-bd16-4313-b1f4-08e289837d3c · outbound

This paper cites A review on reinforcement learning: Introduction and applications in industrial process control.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies A review on reinforcement learning: Introduction and applications in industrial process control

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.259262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.301545Z digest=sha256:0aff9b26b7b33770a36959276d52537d3e9f10552f48c97f6832b9145bcb777d

Observation d5e13382-5942-4d95-a0a0-e8ae69e58dcd · outbound

This paper cites Targeted Reduction of Causal Models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Targeted Reduction of Causal Models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.237252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.306817Z digest=sha256:31562f22617b76fa6356e5e906c46cf0a9cd6815a882c8b4a26170bcfa55e660

Observation dc779504-5c2d-4899-9bac-a3eba4581158 · outbound

This paper cites Causality.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causality

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.192843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.316979Z digest=sha256:a8d3bbcc8cb51a35566f8aac58f967fd8dba3772f8b4c1d82e9937cba36ceaa7

Observation 3d7f21a3-a9c1-4a92-96a6-5a3cba8a0666 · outbound

This paper cites Elements of Causal Inference – Foundations and Learning Algorithms.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Elements of Causal Inference – Foundations and Learning Algorithms

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.171955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.321785Z digest=sha256:6cf4fd9124342dd7502518fd11c369a275526617e9827de61113b3dd60d95c78

Observation 1842fbfd-2e17-4c27-baa7-6860a213115d · outbound

This paper cites Abstracting Causal Models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Abstracting Causal Models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.151157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.327967Z digest=sha256:d93491fa7c7579a665c13464861557e80b4446d6159b4f451b8e4e29e4050032

Observation 39bc8457-3cd2-4e20-bcb0-5b1bd2e8210a · outbound

This paper cites Approximate Causal Abstractions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Approximate Causal Abstractions

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.128006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.333042Z digest=sha256:ae12f62a894ee544d8ca225ee972eb234ccaaead572be8e6d7e0ddc60b7bd442

Observation 7694d6ea-c01d-4faa-aaf3-59cc967fae51 · outbound

This paper cites Causal Abstraction with Soft Interventions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction with Soft Interventions

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.108359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.337908Z digest=sha256:52f4eb1a086d0e697e9eb34e32f233a004424e135608a5a3a42ec5bd17477bb8

Observation e58c0533-2ce7-4186-8d35-9ec41c874277 · outbound

This paper cites Compositional abstraction error and a category of causal models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Compositional abstraction error and a category of causal models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.086804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.343403Z digest=sha256:ecbbacad5b863efcb4cb364586e064c7d6fe11010a59bc2797093e1dcea91ad4

Observation 4ecdeea4-ca1e-4230-8825-1855bc0ca1db · outbound

This paper cites Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.348803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.348803Z digest=sha256:ef571df99f6b21b10e30d195e50dcfee1cabd595335705879f4c190343a5d4a1

Observation ef440e19-8954-40df-8a91-650817e81f34 · outbound

This paper cites Visual causal feature learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visual causal feature learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.064441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.354099Z digest=sha256:053968f4e98768f7805717fcc51585ace9edf3c817b77aff8c204a256d8c5705

Observation 442984d5-a9f9-4ade-b3e0-cd46dd24725e · outbound

This paper cites Causal consistency of structural equation models.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal consistency of structural equation models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.042834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.358642Z digest=sha256:ace7cd6c249e10809c3c03a4b51c17bfd6f9951d8c0c65638fbfb06d5b556a9c

Observation 78cf3f97-fb04-412d-ad77-fc8b2fef3ff5 · outbound

This paper cites Homomor- phism Autoencoder–Learning Group Structured Representations from Observed Transitions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Homomor- phism Autoencoder–Learning Group Structured Representations from Observed Transitions

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.024586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.363208Z digest=sha256:dc2a58fe9e5b339e6fbd0fc48420bb52a7957ae5d2b29387865c8b50b4dac0a9

Observation cfe8954d-212a-4e89-a73b-2bbd02d0cc79 · outbound

This paper cites Gymnasium: A Standard Interface for Reinforcement Learning Environments.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Gymnasium: A Standard Interface for Reinforcement Learning Environments

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.368467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.368467Z digest=sha256:4a5f39e3990155e3c6846a2bf645c818a3556366ba7d7899a05a9508a49292c5

Observation 8191d131-a17a-4d61-8fad-91056cfc21e6 · outbound

This paper cites Learning to play table tennis from scratch using muscular robots.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning to play table tennis from scratch using muscular robots

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:38.000926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.373740Z digest=sha256:f0c9b20bec9cee930667c4aa7377ee75e51e19641897f120409cc2773d8ae0e4

Observation 94589efd-bd16-4145-8b33-daccf0f594cf · outbound

This paper cites Safe & accurate at speed with tendons: A robot arm for exploring dynamic motion.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Safe & accurate at speed with tendons: A robot arm for exploring dynamic motion

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.983185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.378670Z digest=sha256:bc41429ee3a8c0d49b1aa1163a82c89221e85b2e6a6114efa05fe2a55ad3f1d0

Observation 9cee4fc6-7cd1-419b-ab02-b0d7cbe61986 · outbound

This paper cites Explainable reinforcement learning: A survey and comparative review.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning: A survey and comparative review

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.967569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.383546Z digest=sha256:506320cd6d37c5473d71c5921d8f7546bb3f89418e589429a7ca8fe0968ae7e9

Observation 073134fe-be44-444b-a248-1261899fbdd7 · outbound

This paper cites Visualizing and understanding Atari agents.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visualizing and understanding Atari agents

Reference 23

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.950030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.389396Z digest=sha256:a3b85a560bacfe50ce8a830dcb2770be0a021a585e1b430dfc1b8d0c097f459b

Observation 422b1cd2-0d11-49dc-ae39-b79406e8d39f · outbound

This paper cites Transparency and explanation in deep reinforcement learning neural networks.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Transparency and explanation in deep reinforcement learning neural networks

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.933941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.394262Z digest=sha256:a84439566b2bc075edbe43f2029de1297cb786dbc980e5aad906786d6a7f0d50

Observation 0d0c8e15-c119-4931-85b4-07b1195709af · outbound

This paper cites Towards interpretable reinforcement learning using attention augmented agents.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Towards interpretable reinforcement learning using attention augmented agents

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.918084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.400049Z digest=sha256:f8150f9dcf373513012fc5a22cf5714771ffc44eda5f7792c98b4aeefb3cceff

Observation 2f817359-2cf9-4f5b-b3ed-8c2149af47a4 · outbound

This paper cites Explainable robotic systems: Under- standing goal-driven actions in a reinforcement learning scenario.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable robotic systems: Under- standing goal-driven actions in a reinforcement learning scenario

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.900836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.404901Z digest=sha256:19a6b1394156349ee8233063a246f5a15277a8e4ff8e85ea1d74809e16e91ee8

Observation 0731a134-e957-444a-bc9f-6ead1652a7bf · outbound

This paper cites Explaining reinforcement learning to mere mortals: An empirical study.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explaining reinforcement learning to mere mortals: An empirical study

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.883975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.410288Z digest=sha256:94966cc2c7a8ea700e653b5c027218db606c766c39f467af028566e840c94267

Observation 72e5ddf2-be85-4494-929f-e2850e358765 · outbound

This paper cites Learning "what-if" explanations for sequential decision-making.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning "what-if" explanations for sequential decision-making

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.867244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.415048Z digest=sha256:25298bf8be15d4188edb7b88504b06883e30accf6ec09b44bc3cae57916a4919

Observation a54b121b-5235-4e93-8eee-1d2dfbd287c8 · outbound

This paper cites Graying the black box: Understanding DQNs.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Graying the black box: Understanding DQNs

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.849813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.419818Z digest=sha256:1826bdaf6dfb6e2573d3a7abbccce9f1704a9511dde30f8961d6dd38f2622df2

Observation ee84d02c-791a-4b0f-a01d-1ece076cd3c5 · outbound

This paper cites Generation of policy-level explanations for reinforcement learning.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Generation of policy-level explanations for reinforcement learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.833618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.425178Z digest=sha256:c782664d26af02b86c17f03e0ad9fa87198a39a6b7bff49b6a4310d0ac71c3c8

Observation 9267c37f-385a-4c34-a9c2-cec05a521d51 · outbound

This paper cites TLdR: Policy summarization for factored SSP problems using temporal abstractions.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies TLdR: Policy summarization for factored SSP problems using temporal abstractions

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.812888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.430381Z digest=sha256:41a760ac101b2e22e79c585032618d44e83f3583f174b9f3c0bec499a026f8eb

Observation f290b596-ced0-4c63-b473-c29dd69035af · outbound

This paper cites Explainable reinforcement learning through a causal lens.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning through a causal lens

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.795008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.436342Z digest=sha256:167b03bee03548e5bef564f4a735ba4a6d9cacf6ed6b89802db540d09d297ca8

Observation d0371ab7-c37f-4c11-9fb1-b4cd23cf7617 · outbound

This paper cites Causal abstractions of neural networks.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal abstractions of neural networks

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.777692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.441148Z digest=sha256:68e24595bc7c84f2bf76ae51091b81034790515d6abb715e16030e88cd1c211f

Observation 32c7e9ef-373c-46ad-a735-46ac73f5df31 · outbound

This paper cites Segment anything.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Segment anything

Reference 34

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.760111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.446132Z digest=sha256:9e81ca83d2b02e3de172028048cbe4f0f93c025838407de3d7f128e2f3ee0369

Observation c6d279c7-9ac4-4e9c-96c7-7fdf2563759d · outbound

This paper cites Grad-CAM: Visual explanations from deep networks via gradient-based localization.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Grad-CAM: Visual explanations from deep networks via gradient-based localization

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.744252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.453979Z digest=sha256:b4f02126c5953116a70e33920e356969e6c5b2619edcb36ecffdfa96f00a08eb

Observation 506935a4-89b3-441b-8d66-ae28f18e397e · outbound

This paper cites Foundations of structural causal models with cycles and latent variables.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Foundations of structural causal models with cycles and latent variables

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.727134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.459902Z digest=sha256:c24258fcd283a16aff7c81ab8fb1be442fa6f747b3623bfb8e781605f8b54383

Observation 0f986416-cc18-4aaf-b1d5-2d6a3fa1cbeb · outbound

This paper cites Dependence, correlation and gaussianity in independent component analysis.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Dependence, correlation and gaussianity in independent component analysis

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.709254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.466360Z digest=sha256:88a9dbda15465ce399fe19f24d3c5e29a9fae95f900ab29028c024c6e960750e

Observation 1fcdcd2d-dcc3-4c41-9ec7-25fe42b0c189 · outbound

This paper cites Stable-Baselines 3: Reliable reinforcement learning implementations.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Stable-Baselines 3: Reliable reinforcement learning implementations

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.692501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.471536Z digest=sha256:f3478963c2461543d99e69f0523d94218ac6dd9b1624dce8752c422c3f76be96

Observation 814cd2cb-92ff-4d17-a93c-72f28dd5e93e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Proximal Policy Optimization Algorithms

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T15:51:37.477154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:51:37.477154Z digest=sha256:3bb186a66d188ec6112a37022887e27d8e3c4547fef526e2b83eb47602b795cf

Observation bf472ef0-6772-40db-b29b-1bce7b2d9b99 · outbound

This paper cites The high-level model has (n+1) endogenous variables {Y, Z1,.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The high-level model has (n+1) endogenous variables {Y, Z1,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.673668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.482934Z digest=sha256:5aa36270edae806dcb45bf648894c13ed07313c5f54d9be1e186298de089f6c5

Observation c7b768c2-5b5a-4aed-b5e3-e492691afd4f · outbound

This paper cites The exogenous variables {W0, W1,.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The exogenous variables {W0, W1,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:51:37.653791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.489483Z digest=sha256:89ff6c56ea2d0b7de636671ce6439a60a8335ccada5199a598f7b5ad5432f407

Observation b4eb86eb-1987-477e-bc8b-154b844b785d · outbound

This paper cites (id − f1)−1.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies (id − f1)−1

Reference 43

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T15:51:37.634033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.495684Z digest=sha256:70ad089f64952feadc4dbe7b622cc5aedab719841019d284235a8a07613981cc

Observation 824cb148-9f02-49f3-90c1-c643ed51f9a8 · outbound

This paper cites an unresolved cited work.

Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:51:38.211855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T15:51:37.312001Z digest=sha256:20586b11cfa49d9c1f6ef8cf9b79f275e048cf2e77a3bd50132f0befc31bdef5

Pith citing papers

Observation 750427c3-6156-4f82-80be-4ce14ec2e6ed · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:38:55.872495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:ee952ec95729509fcc69b002136dc4fe1c3847e9a0c1a46c65d84f19fb38a015