Pith. sign in

Paper Citation Record · LEDGER

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic

As of 8 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.05445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05445 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:01.072056Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9d48139d-8836-41e3-ae6b-2f8ff8f667bd · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.279110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.605339Z digest=sha256:31fd3f918954a28e8c4fcd38934eb934a363d9dc5c2222ab57dbc3b7ec17abab

Observation ebfad1ab-7e77-4f31-908c-554a0d59c0ca · outbound

This paper cites and Pearl, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Pearl, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.227641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.665123Z digest=sha256:1a26da3a777fb6fc3282719de40ad35af87bdb0528756db6d98647dfe2e65844

Observation d45b2319-f964-4499-bb6c-e92a8eae836f · outbound

This paper cites OpenAI Gym.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:56.756738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:56.756738Z digest=sha256:d8abb1bd0e033ed5403a48366c92251df436fb3d40b118fd0fa1fdd70e2de045

Observation c9389d4d-dee5-477b-bc34-d3fbf3a36cdc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.172607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.823275Z digest=sha256:9552d0b15a849b93525cd014deba8d54e0eafc7ecd48eeb6e31f982fb8a3fe99

Observation 269ca6ea-c4e5-459a-b44b-bdd72d06a58d · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.126760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.896354Z digest=sha256:6edc7376e84a4e12a7ba2de09a9de2bc183f2fef7ebd6e381f593e88aec5878f

Observation 0246fd85-67da-4d61-be80-6ef066d20f43 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.066815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.970646Z digest=sha256:d5221982d63864aa5d8ea37d140a9cb148f56bf2a93748d287b18be69f4a1146

Observation 54e0fd50-c7ff-48b2-92bc-d12fdeea3cd7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.004765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.045401Z digest=sha256:1bf866b5a07dfc8ceb328a83980d8d9dc0e773f5ae2dbd9be64b1de6155ae94b

Observation bcaa6341-0c28-4e2d-b9e9-124faad92207 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.944649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.100420Z digest=sha256:7aa5f8c67f8051e287ce43c4609f39984fec5e1e03aa01fbfeed061f82d541f2

Observation 0db23e0e-4f18-4ba0-9a9a-f05a486d13d7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.877110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.151008Z digest=sha256:793a9ac9ceaddea2a9dd07ea55397901164af0d29a8e9e9b2da63530d943db53

Observation 19c2d0ed-9252-4fb4-a7f4-1523034c1e75 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.794972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.207458Z digest=sha256:bbcef96cf4b82d9d6b6ea60678754aa2dca6a0eec98a9268af9791dab28d0ef3

Observation 93cededb-216f-4353-96b7-e98981c165ab · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.744172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.263934Z digest=sha256:286689de57de055f790544339e59d8ddcf204867ac78dc74e3c1681d7baf9ed5

Observation a3c08d5c-115f-451d-aa63-d8a87d64f85c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.686487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.336074Z digest=sha256:927eae1fe464d20133b9e5da0d0caccb80f38874296f67451e6ff428c65fea31

Observation e26ac464-1600-40d0-a5c7-8eac7cf39a4a · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.606814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.402866Z digest=sha256:89b715d61102583e843e9bac78d6131d96395649b4b3f771e0844dd0a1d58d31

Observation 1bb3ecff-3de4-4781-a77d-9f9581352569 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.529392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.466460Z digest=sha256:7860c1e95ad427a13dd23acb20a0f3b42ede1a99ac6f894f02b59b424bd26be3

Observation 1c900d36-b428-43cd-bbd6-7f6eb859d864 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.519671Z digest=sha256:244925180de5d03fabac6c2b6d7706587432b9d04d998c0e92d3dc59d9e8642a

Observation a5195288-9cf4-4d74-8860-25e085d20bd6 · outbound

This paper cites P., Hunt, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic P., Hunt, J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.392073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.574314Z digest=sha256:9ca75d86ac46681e9abac16771e4a5b718546303f95b6a3cb95c35eb748dcd1c

Observation 1f25f871-a3d6-4f61-b15e-c6f07e9ebbbd · outbound

This paper cites Continuous control with deep reinforcement learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Continuous control with deep reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:57.644181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:57.644181Z digest=sha256:a783f98068cc684452a952ca55f8a97b18c67bd1e2b343754804712dcf032e3d

Observation 94d44d67-9b6d-4e78-b265-4a61b205d29e · outbound

This paper cites and Krishnamurthy, A.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Krishnamurthy, A

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.330213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.711279Z digest=sha256:7d0701da0a1b06e4b2db58d5fdf165402b543937329f9cee16842608e214982a

Observation 254e7504-a210-4e33-9413-903c850d747a · outbound

This paper cites Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:36:01.307595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.781252Z digest=sha256:9b98bb4041ccf589e2fd80f4a6b7c7b748b164e9817a66592f1f507a6aee3a68

Observation 100ddbe6-78c6-4a76-95b9-2d421e05afc1 · outbound

This paper cites V., Sima, K., and Leong, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Sima, K., and Leong, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.277446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.852453Z digest=sha256:c611b092b1f0cdee585b89b95dbd540580f53e65f222f85b0efdceccf5f8240b

Observation cd705998-f34d-428b-b67f-219cefb0ea64 · outbound

This paper cites V., Fu, D., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Fu, D., and Leong, T.-Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.223838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.918663Z digest=sha256:656dcecdbdcdc93ab49c08ca422c365dc5efd2246158fd5b79a04174555e0c5c

Observation 169f980f-2cc9-41a2-b6cc-4eb0f12180f6 · outbound

This paper cites V., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., and Leong, T.-Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.201758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.008523Z digest=sha256:daa81aa36063527ed12f5042641f6da89144eb9449f0a21f7b6b774f4cb293b1

Observation 08ee4ad2-093f-4500-914e-d302f1fdeed7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.173753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.116971Z digest=sha256:7b09120a2cfbe1aba2f7589c58987be20a6bdd14da924a210ca562f800ff05d3

Observation cf3d0505-6724-4442-8426-b9eb59d5a31c · outbound

This paper cites and Sontag, D.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Sontag, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.138332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.205073Z digest=sha256:8c380535a6e00e07a152f5553c0417042851c339240187663817201af657becb

Observation 14a94d8c-10b4-4183-af31-0c2e57d9b490 · outbound

This paper cites o lkopf, B., R \.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic o lkopf, B., R \

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.099722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.290756Z digest=sha256:59321759154bb76bdb08fbc7039480d3a996340ce0d899e5bdab5d68f270c312

Observation 581578b9-8bdf-4b87-9b1e-d95a612d4adc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.072534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.377118Z digest=sha256:5fc1f7ac823541414be70c381df162e3664c73321032f50253255ab4353c5ee7

Observation e968b5eb-b467-4199-881c-78b8a88897fa · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.045301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.488977Z digest=sha256:db96fa71cd9f027cfbc81908578f29a20310e31a0911b2fb9d2c86173f514708

Observation 217fafc5-d441-4cde-a5b1-50b4c0ab6376 · outbound

This paper cites Robust Policy Optimization in Deep Reinforcement Learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Robust Policy Optimization in Deep Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.559435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.559435Z digest=sha256:c341fbefd579a101620a7c9bf1f8995fb17b8eb4665b892627523c705497ff01

Observation 4fc2e2ad-4e12-49c6-800c-3f5284aa57b3 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.023254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.624485Z digest=sha256:c1f21019e8579c0f439e395329507264e90a31999c2377f4d58335850c2e5f69

Observation 9d0de39e-ea15-431c-b06e-bd4542495984 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.000514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.710725Z digest=sha256:1dce249d277a9e58a15c3a6177cd6f5e97e0459251a3558c79fc66ba903a97b7

Observation 402f2e2e-4070-4651-aad8-3c837732275a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.897292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.897292Z digest=sha256:2323a5380c8a361d028cc5573a177c3da5728e1f50a2eaa42eef7f3eb3fcd833

Observation 48daab15-c73f-4d61-a90a-335b1c953714 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.979816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.021496Z digest=sha256:59ae2e2090823f21880062b4138f5cc82300ea95b6f98aa19f9dfaacf860f08c

Observation 8daf1743-76ff-4305-9b2c-f5d5373ddb91 · outbound

This paper cites A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.958176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.121778Z digest=sha256:1b50b5eff4fb86d2e5418d68748bec9f9dfa28cd5511673490e1551e7eeea8fa

Observation c5b5fe2e-7771-4e7e-9717-807af026f026 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.244632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.244632Z digest=sha256:1ae082d11a205e8c817862d43daedfac44bb7c531ea6cc49da4af31e24cac947

Observation 3352881e-78dc-4e8d-93d7-ad8bbee579a2 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.864149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.407242Z digest=sha256:38545a7cd1c0cf39fffef6c42dd7a6e4da2d0142c3bf028c5282ea98b8ec9ea5

Observation 4fac4cf2-e0b3-47da-87f2-e820812baf33 · outbound

This paper cites V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.723703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.524679Z digest=sha256:c517d3ccd63cd50b2009fdfcb1ded4c2448405ed58d501aaa26a4b3f30fad213

Observation d8dd8711-3e83-4881-b14d-2ab66946ff93 · outbound

This paper cites V., Lee, Y., Hoang, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Lee, Y., Hoang, T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.388402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.653506Z digest=sha256:d517589d8ee50f461c1e4f48b7e24308088e9ef3bb1b3fa514b9bb994fea9db2

Observation bfa7acdc-9448-4e86-9ce1-1a13d2e9f8a9 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.121133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.776828Z digest=sha256:554ae9bba4797d8f6ebd79416f30d596e35a0efe11a88632caf437fff0e09e98

Observation 1af0e510-f2e3-4913-8113-8e41f75d71a1 · outbound

This paper cites T., and Athey, S.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic T., and Athey, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.919306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.919306Z digest=sha256:17bf4ca25334a0e1bdf0d389f819b7808b0254827cf00cd757ecfb7fdbf7b766

Observation db3c2665-14ee-4a57-9ddb-ebd1843af036 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.970054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.076368Z digest=sha256:482b31127da23ec52e3f3b34e318ace7d69ef6cf1feba6de95dab9e062995a95

Observation c0da3255-67e7-4351-81f7-c722802bac01 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.787674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.191439Z digest=sha256:f787453136474e2b35c208c3fcfc392c112785803cbccbc5b7f5835e564f8a29

Observation 32c8258f-d0a5-4952-8da4-b8b519314caa · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.590333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.346427Z digest=sha256:71cbdaf3ad9fb3f247edf0097484db784a3e6afb2c65eeb3cd74bb8b60825bfd

Observation c21a2795-aca7-4916-b332-982f024c8b3f · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.404749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.501633Z digest=sha256:99a21d55bbf1e83a51bde87ffe35d95a6e6b83c002bb550a49bcfcdb9c0eb3bb

Observation 68da7043-9bbd-49a3-8a32-de12c8a35dec · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.249365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.671471Z digest=sha256:762d602e91132aee641afd397f5b452664514680ebd7c4dcc73af12498e96807

Observation 23686f90-8263-4ecb-a308-ad80d9cc31a7 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.052297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.792691Z digest=sha256:7b253a1a45ce86d84eab009b82a95d1a0977ec07127641acccf52bf1a7b7d8ae

Observation 0adad830-802a-444e-aef3-221fd80f7c2c · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:01.892224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.884611Z digest=sha256:4b00e49d8af934eb984263322954d6d8fcf47f4b0c55006bd0d1bbf1561509fa

Observation 30a6359d-8f35-45dc-9155-a15bfcea025c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.747435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.981793Z digest=sha256:a1d306528bf8802b65c1b87538d482580c9e13e63266573804d709335d2f2ead

Observation d1c23697-be22-4220-a151-691c6b834a31 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.551087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T10:36:01.072056Z digest=sha256:3052c4ff5df02fb3c9eb8896b00601c3540485cd34c2daa77bfb5438ef5e3fe0

Pith citing papers

No inbound Pith citation observations are available.