Pith. sign in

Paper Citation Record · LEDGER

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic

As of 9 August 2026, this Paper Citation Record lists 48 of 48 outbound references and 0 inbound Pith citation observations for arXiv:2506.05445.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05445 v1

Coverage vector

measured 48 of 48 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:36:01.072056Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

48 of 48 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9d48139d-8836-41e3-ae6b-2f8ff8f667bd · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.279110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.605339Z digest=sha256:389c4922ce3930df6818b9d98afb86e80f10560e799de7e9e31ba19a3bf66f82

Observation ebfad1ab-7e77-4f31-908c-554a0d59c0ca · outbound

This paper cites and Pearl, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Pearl, J

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.227641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.665123Z digest=sha256:8da742e5613a5df8c834d138b9dfd603557705a0cfb4871be6ecedb8a0eb70c0

Observation d45b2319-f964-4499-bb6c-e92a8eae836f · outbound

This paper cites OpenAI Gym.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic OpenAI Gym

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:56.756738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:56.756738Z digest=sha256:d8abb1bd0e033ed5403a48366c92251df436fb3d40b118fd0fa1fdd70e2de045

Observation c9389d4d-dee5-477b-bc34-d3fbf3a36cdc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.172607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.823275Z digest=sha256:7c5c6c13d927372c12f37608d549c8d6a8b5ec284ae8e883091d2e011ef72c17

Observation 269ca6ea-c4e5-459a-b44b-bdd72d06a58d · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:05.126760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.896354Z digest=sha256:ec4158366ef3bdd5d384bc316cb7ca0e6e310cc062001fbef5b7cff938e7e6c3

Observation 0246fd85-67da-4d61-be80-6ef066d20f43 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.066815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:56.970646Z digest=sha256:0fac8f403803ba82a82a748a0b1d2dd18cc0cd68b22c5f8185633ab678f87d7b

Observation 54e0fd50-c7ff-48b2-92bc-d12fdeea3cd7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:05.004765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.045401Z digest=sha256:dc62cbd1441af16c4cafef09a3c386a36bb75684ec49e1a23cd110570f3b22cd

Observation bcaa6341-0c28-4e2d-b9e9-124faad92207 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.944649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.100420Z digest=sha256:16008d69ae19d5e49cf0ccbafe17edf4e97c5c38c9971cc460772e75d439a4a1

Observation 0db23e0e-4f18-4ba0-9a9a-f05a486d13d7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.877110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.151008Z digest=sha256:329322846b1df62cf32f52fb7159980773f31a5d917ee0b0d80483e54c3c02ae

Observation 19c2d0ed-9252-4fb4-a7f4-1523034c1e75 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.794972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.207458Z digest=sha256:b10f10725842763f6aeeb4444631d5692b6f1242613253344b94ebf74519bf4e

Observation 93cededb-216f-4353-96b7-e98981c165ab · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.744172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.263934Z digest=sha256:8f81f08f1743452b39ea135ee1e3a44c635b64855977e4e83f686fd5ca5ada22

Observation a3c08d5c-115f-451d-aa63-d8a87d64f85c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.686487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.336074Z digest=sha256:4143a93fe14c513a5a1f49c0cc88350f58dbe0150466802b07036aebb17ca143

Observation e26ac464-1600-40d0-a5c7-8eac7cf39a4a · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.606814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.402866Z digest=sha256:10e7c6e395a63295caeb05ad3d19f861a8ba0448bbb66449562626aba6d268a0

Observation 1bb3ecff-3de4-4781-a77d-9f9581352569 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.529392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.466460Z digest=sha256:9621470abb53293a2ee53c621ef8ca784742c4d0688aa81aef6169b369c78635

Observation 1c900d36-b428-43cd-bbd6-7f6eb859d864 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.470713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.519671Z digest=sha256:32ebe9dc33480337244386ae6e6362195a65b035bb82ea33947f641efbcf3ee2

Observation a5195288-9cf4-4d74-8860-25e085d20bd6 · outbound

This paper cites P., Hunt, J.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic P., Hunt, J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.392073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.574314Z digest=sha256:18a1b2295c8778b8b4510e0b789758a15c8d20fe8a0c64b6b84772e63f95cf55

Observation 1f25f871-a3d6-4f61-b15e-c6f07e9ebbbd · outbound

This paper cites Continuous control with deep reinforcement learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Continuous control with deep reinforcement learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:57.644181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:57.644181Z digest=sha256:a783f98068cc684452a952ca55f8a97b18c67bd1e2b343754804712dcf032e3d

Observation 94d44d67-9b6d-4e78-b265-4a61b205d29e · outbound

This paper cites and Krishnamurthy, A.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Krishnamurthy, A

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.330213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.711279Z digest=sha256:2fb13ec49cf7c1ef6f84e3f1ed0a20cf47ccfad7d5248a7abe06458b8fd26126

Observation 254e7504-a210-4e33-9413-903c850d747a · outbound

This paper cites Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Rewriting History with Inverse RL: Hindsight Inference for Policy Improvement

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:36:01.307595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.781252Z digest=sha256:1870e10058b9fec4df2d9adb8f160c2c2767ac4dd25f9f3a23bcbb5f7b4b9cbf

Observation 100ddbe6-78c6-4a76-95b9-2d421e05afc1 · outbound

This paper cites V., Sima, K., and Leong, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Sima, K., and Leong, T

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.277446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.852453Z digest=sha256:823ab947c879fab8fb1eb4bca693d4c8b5dcc08459088e00c29f8b9374f8a601

Observation cd705998-f34d-428b-b67f-219cefb0ea64 · outbound

This paper cites V., Fu, D., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Fu, D., and Leong, T.-Y

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.223838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:57.918663Z digest=sha256:82f4e858c5b450bf6dd0ba5ec501eaf481981c2c4789c37a7481c705c164e904

Observation 169f980f-2cc9-41a2-b6cc-4eb0f12180f6 · outbound

This paper cites V., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., and Leong, T.-Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.201758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.008523Z digest=sha256:f4f14afe5eaf7dbcc70770407e638d6a905590c4ecf50a8c80eee05e8079ffb5

Observation 08ee4ad2-093f-4500-914e-d302f1fdeed7 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.173753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.116971Z digest=sha256:0296040ae8da671aead945339c8127cbd2ea518946cbd38fde9d9eb1207519ea

Observation cf3d0505-6724-4442-8426-b9eb59d5a31c · outbound

This paper cites and Sontag, D.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Sontag, D

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.138332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.205073Z digest=sha256:dc98c3279658790b9b6fbc3b9ef9a6ad03ab6c0bdb7115158ac216f086587b09

Observation 14a94d8c-10b4-4183-af31-0c2e57d9b490 · outbound

This paper cites o lkopf, B., R \.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic o lkopf, B., R \

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:04.099722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.290756Z digest=sha256:741a011fea39bc190ca38398f9255943aa3f07a2d143ee925dee5015150915e7

Observation 581578b9-8bdf-4b87-9b1e-d95a612d4adc · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.072534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.377118Z digest=sha256:69c35b14e48148d7bab659a16e90111bd63743761e27373e31605051b8e9fc71

Observation e968b5eb-b467-4199-881c-78b8a88897fa · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.045301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.488977Z digest=sha256:a58982a1209cbe23556693d534ff05eb207fa24aa006c157493245ed19d2b2fa

Observation 217fafc5-d441-4cde-a5b1-50b4c0ab6376 · outbound

This paper cites Robust Policy Optimization in Deep Reinforcement Learning.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Robust Policy Optimization in Deep Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.559435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.559435Z digest=sha256:c341fbefd579a101620a7c9bf1f8995fb17b8eb4665b892627523c705497ff01

Observation 4fc2e2ad-4e12-49c6-800c-3f5284aa57b3 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.023254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.624485Z digest=sha256:4fde3f8d6a50fc47c23a9b4a388293a15583e8aba2c6550cdcbbbe623dc1b46d

Observation 9d0de39e-ea15-431c-b06e-bd4542495984 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:04.000514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:58.710725Z digest=sha256:db46d9028fbe11b4effd9a44a2e34fe8ef14a35199a7892dd722b878d4ff40ef

Observation 402f2e2e-4070-4651-aad8-3c837732275a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Proximal Policy Optimization Algorithms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:58.897292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:58.897292Z digest=sha256:2323a5380c8a361d028cc5573a177c3da5728e1f50a2eaa42eef7f3eb3fcd833

Observation 48daab15-c73f-4d61-a90a-335b1c953714 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.979816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.021496Z digest=sha256:d3055f56c6defc538601af3e53855cb4c97956f6e8d136dd206a7d6f64d1a203

Observation 8daf1743-76ff-4305-9b2c-f5d5373ddb91 · outbound

This paper cites A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic A., Mehrjou, A., Itti, L., and Sch \"o lkopf, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.958176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.121778Z digest=sha256:d9885d07938d53c02e09cd16891fb2cb6a3a9474a049e14294e5f0871713f8c2

Observation c5b5fe2e-7771-4e7e-9717-807af026f026 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.244632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.244632Z digest=sha256:1ae082d11a205e8c817862d43daedfac44bb7c531ea6cc49da4af31e24cac947

Observation 3352881e-78dc-4e8d-93d7-ad8bbee579a2 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.864149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.407242Z digest=sha256:c1842ae304b02468c4a48ef2f2bd26031397e02bd820c608ea03481a1cacc71f

Observation 4fac4cf2-e0b3-47da-87f2-e820812baf33 · outbound

This paper cites V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Bhattacharyya, A., Lee, Y., and Leong, T.-Y

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.723703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.524679Z digest=sha256:69b0084462972c1583a066bcc8a2b3079fb0118ccbfb9855508325d6ef8a1abf

Observation d8dd8711-3e83-4881-b14d-2ab66946ff93 · outbound

This paper cites V., Lee, Y., Hoang, T.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic V., Lee, Y., Hoang, T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:03.388402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.653506Z digest=sha256:7aebaeb5bc3e77d590784c62d87b217ec3fd65994d783e18b97df200f8f2ae31

Observation bfa7acdc-9448-4e86-9ce1-1a13d2e9f8a9 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:03.121133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:35:59.776828Z digest=sha256:9c591707767b0f9a9311ff548c209785883e8efca8857ea64fdf28cc4582c644

Observation 1af0e510-f2e3-4913-8113-8e41f75d71a1 · outbound

This paper cites T., and Athey, S.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic T., and Athey, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:59.919306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:35:59.919306Z digest=sha256:17bf4ca25334a0e1bdf0d389f819b7808b0254827cf00cd757ecfb7fdbf7b766

Observation db3c2665-14ee-4a57-9ddb-ebd1843af036 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.970054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.076368Z digest=sha256:260c34c85963797cf5292241f6e57b2794b05a3afdb1e123c1a8e954b4058bc0

Observation c0da3255-67e7-4351-81f7-c722802bac01 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:02.787674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.191439Z digest=sha256:f0a8a4af0698b1867a656dd64025dd60ad6230cbb3d596e64106210d0f0d4446

Observation 32c8258f-d0a5-4952-8da4-b8b519314caa · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.590333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.346427Z digest=sha256:b640a5416beb7d45a77500564d538329acc64a2edbd3b0db61aa39abfdac87b3

Observation c21a2795-aca7-4916-b332-982f024c8b3f · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.404749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.501633Z digest=sha256:38eb604023a9b88f0f9b2fdd366752f0e8033e8c4c539c94d0eff4f222ae052f

Observation 68da7043-9bbd-49a3-8a32-de12c8a35dec · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.249365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.671471Z digest=sha256:cb6b441cc3f5ac36c974d13fa0a10bacb1e065146a215279c0adaa04b4ba745c

Observation 23686f90-8263-4ecb-a308-ad80d9cc31a7 · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:02.052297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.792691Z digest=sha256:6e5bf7047f799b45d3f2f6f44d7ed2dff1d0a14ee556f44819d49012d56185f3

Observation 0adad830-802a-444e-aef3-221fd80f7c2c · outbound

This paper cites and Bareinboim, E.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic and Bareinboim, E

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:36:01.892224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.884611Z digest=sha256:50ab7e5da665f79960ab4a9f572fa4045646cf3f037b6fc57844d9545fa380d3

Observation 30a6359d-8f35-45dc-9155-a15bfcea025c · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.747435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:00.981793Z digest=sha256:d499c130e7835eeec47f5c67fad2e36455ef636daffbf724d3e83b56298909b4

Observation d1c23697-be22-4220-a151-691c6b834a31 · outbound

This paper cites an unresolved cited work.

Causal Policy Learning in Reinforcement Learning: Backdoor-Adjusted Soft Actor-Critic Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:36:01.551087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T10:36:01.072056Z digest=sha256:7617c4b05bf4e4b4da9c8034ade022609d803b0be950ae8963800f57071ab980

Pith citing papers

No inbound Pith citation observations are available.