Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:51:37.495684Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2507.14901.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:51:37.495684Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-27T01:13:11.483599Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T20:38:55.871103Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c996a0ed-63c1-43bc-acea-b116e46d4fae · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Mastering the game of Go with deep neural networks and tree search
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d3a92e27-364d-4403-8015-737a8d3a7b7d · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Playing Atari with Deep Reinforcement Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13efe497-c48a-4eb8-b7f2-aa9cd992a0a9 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning in robotics: A survey
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9c5ddcdc-76b8-4b30-b285-98c9dcfb7fc1 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Resource management with deep reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f6bc40c8-296c-41b7-831f-8d63d4b72b9c · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies End to End Learning for Self-Driving Cars
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f3ee64d-9fcd-4e88-934d-8000b51c2f17 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Reinforcement learning based recommender systems: A survey
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b8c337f6-bd16-4313-b1f4-08e289837d3c · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies A review on reinforcement learning: Introduction and applications in industrial process control
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d5e13382-5942-4d95-a0a0-e8ae69e58dcd · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Targeted Reduction of Causal Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dc779504-5c2d-4899-9bac-a3eba4581158 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causality
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3d7f21a3-a9c1-4a92-96a6-5a3cba8a0666 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Elements of Causal Inference – Foundations and Learning Algorithms
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1842fbfd-2e17-4c27-baa7-6860a213115d · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Abstracting Causal Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 39bc8457-3cd2-4e20-bcb0-5b1bd2e8210a · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Approximate Causal Abstractions
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7694d6ea-c01d-4faa-aaf3-59cc967fae51 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction with Soft Interventions
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e58c0533-2ce7-4186-8d35-9ec41c874277 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Compositional abstraction error and a category of causal models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ecdeea4-ca1e-4230-8825-1855bc0ca1db · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal Abstraction: A Theoretical Foundation for Mechanistic Interpretability
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef440e19-8954-40df-8a91-650817e81f34 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visual causal feature learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 442984d5-a9f9-4ade-b3e0-cd46dd24725e · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal consistency of structural equation models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 78cf3f97-fb04-412d-ad77-fc8b2fef3ff5 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Homomor- phism Autoencoder–Learning Group Structured Representations from Observed Transitions
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cfe8954d-212a-4e89-a73b-2bbd02d0cc79 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8191d131-a17a-4d61-8fad-91056cfc21e6 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning to play table tennis from scratch using muscular robots
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 94589efd-bd16-4145-8b33-daccf0f594cf · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Safe & accurate at speed with tendons: A robot arm for exploring dynamic motion
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9cee4fc6-7cd1-419b-ab02-b0d7cbe61986 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning: A survey and comparative review
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 073134fe-be44-444b-a248-1261899fbdd7 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Visualizing and understanding Atari agents
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 422b1cd2-0d11-49dc-ae39-b79406e8d39f · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Transparency and explanation in deep reinforcement learning neural networks
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0d0c8e15-c119-4931-85b4-07b1195709af · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Towards interpretable reinforcement learning using attention augmented agents
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2f817359-2cf9-4f5b-b3ed-8c2149af47a4 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable robotic systems: Under- standing goal-driven actions in a reinforcement learning scenario
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0731a134-e957-444a-bc9f-6ead1652a7bf · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explaining reinforcement learning to mere mortals: An empirical study
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72e5ddf2-be85-4494-929f-e2850e358765 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Learning "what-if" explanations for sequential decision-making
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a54b121b-5235-4e93-8eee-1d2dfbd287c8 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Graying the black box: Understanding DQNs
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee84d02c-791a-4b0f-a01d-1ece076cd3c5 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Generation of policy-level explanations for reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9267c37f-385a-4c34-a9c2-cec05a521d51 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies TLdR: Policy summarization for factored SSP problems using temporal abstractions
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f290b596-ced0-4c63-b473-c29dd69035af · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Explainable reinforcement learning through a causal lens
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d0371ab7-c37f-4c11-9fb1-b4cd23cf7617 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Causal abstractions of neural networks
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 32c7e9ef-373c-46ad-a735-46ac73f5df31 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Segment anything
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c6d279c7-9ac4-4e9c-96c7-7fdf2563759d · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Grad-CAM: Visual explanations from deep networks via gradient-based localization
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 506935a4-89b3-441b-8d66-ae28f18e397e · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Foundations of structural causal models with cycles and latent variables
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f986416-cc18-4aaf-b1d5-2d6a3fa1cbeb · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Dependence, correlation and gaussianity in independent component analysis
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1fcdcd2d-dcc3-4c41-9ec7-25fe42b0c189 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Stable-Baselines 3: Reliable reinforcement learning implementations
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 814cd2cb-92ff-4d17-a93c-72f28dd5e93e · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Proximal Policy Optimization Algorithms
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf472ef0-6772-40db-b29b-1bce7b2d9b99 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The high-level model has (n+1) endogenous variables {Y, Z1,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7b768c2-5b5a-4aed-b5e3-e492691afd4f · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies The exogenous variables {W0, W1,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b4eb86eb-1987-477e-bc8b-154b844b785d · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies (id − f1)−1
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 824cb148-9f02-49f3-90c1-c643ed51f9a8 · outbound
Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 750427c3-6156-4f82-80be-4ce14ec2e6ed · inbound
From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Learning Nonlinear Causal Reductions to Explain Reinforcement Learning Policies
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.