Pith. sign in

Paper Citation Record · LEDGER

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning

As of 23 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2412.18946.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.18946 v2

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:23:19.738484Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 27ac39b2-4344-498b-869d-c2a70008ce27 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.503252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.273231Z digest=sha256:89810994c739ddd6561c8d125e8c876bfb8d1028a26376a5d0d6c9d5163d940d

Observation bcea5a82-ca5e-4b41-84f2-9d321c0fe76b · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.491329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.277840Z digest=sha256:5408bbb23231eaabbb4dcfce357504d3bacfdad662e8ee4dac9bbb3e3af29abf

Observation dacf33f3-1b1a-47b8-9d7b-38c4706e7244 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.480292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.281882Z digest=sha256:cf87ddcc265766240b935a83ae303f531ffe6a0304e75531feb7dcaae020d124

Observation 8e7d8700-58f4-47f8-ae73-2886fb769cf7 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.417813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.286004Z digest=sha256:2cdb326da82b6893b2a529875e5e39e85405c1dffcd6191a1418d03a5fbb91ff

Observation e6f69185-28f6-44d5-b151-38fb0a2aaf98 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.207250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.290750Z digest=sha256:8b6213e9eaa31580165b8c9008fcb439fb9ad4a3cfa5f307a8ddccb6a80ed7f3

Observation ffaa8f7b-2541-4473-b77b-a14d557be8c3 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.167644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.294719Z digest=sha256:f0e9f34c5532c4ad66388fac29ce4cc5d1862f8837126f41c63c1b6a3a150ca5

Observation 720e00de-65d9-4523-b6c3-3a270e2bfbad · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.154949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.299142Z digest=sha256:35bd94adeb679d8b9d49facba14dddb7f1e1cda850dc91b9fdae5de09d18928d

Observation f610db2f-cd6d-4d6f-80ff-74fd8214b0b1 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.142843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.302836Z digest=sha256:a45ef8904a0450341e6a2c12b318a388f9906e97d239e745f3b785f44365b5d1

Observation 86a28032-3764-4d56-8b40-02e6c2896dfb · outbound

This paper cites Soft Actor-Critic Algorithms and Applications.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Soft Actor-Critic Algorithms and Applications

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.306252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.306252Z digest=sha256:bfbc36ad54b3e92ead8d3f2c0f03de929f5cee956cd177930a98dd88f28f8439

Observation 50a3f756-9219-4661-9ad0-01ea54925904 · outbound

This paper cites IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.309973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.309973Z digest=sha256:7fbe7a9574836e9e226bcbf94534be6a6bb2b3a0114c1977fab918f9c24e539a

Observation 3affd819-eeeb-4b18-9bb8-657b76ad1b5a · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.131504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.314769Z digest=sha256:14c28b7313a48e1b98b61fafc25445466a338948e37259cd32df4a616d02a70a

Observation 8fb621b4-8a68-43b6-962f-73b0813e9a97 · outbound

This paper cites H.; Ferguson, C.; Lapedriza, A.; Jones, N.; Gu, S.; and Picard, R.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning H.; Ferguson, C.; Lapedriza, A.; Jones, N.; Gu, S.; and Picard, R

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:23:21.121890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.318514Z digest=sha256:292d3cda5b1141f8dbfff2720e3789747b2cdbfff1d8d3d6bed5a83d64d7b329

Observation 49aaf9ea-5b63-4c6b-8c67-44296920fcc6 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.109516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.321999Z digest=sha256:efd2f9ad55bf5db1337b84513e4ba23d9fc1d8e69f08dbb61a5631bc8a502c17

Observation 277a3048-e88c-4cc1-8be7-93cab3b5584a · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.096307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.325142Z digest=sha256:cc2e71e9fe1c0744bb5a93990d4e23e65b3a4c91ac35b6a4d5e8d8793cab44ca

Observation 22b44148-c0d2-44af-81da-7602be2344bf · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.084402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.329113Z digest=sha256:9834f3a703280a3da610a5b12b4d286cc544b350b904da88e7401b475a670ec9

Observation ccc0b1ff-69cd-45da-8104-402df3abc69d · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.072072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.332277Z digest=sha256:c0a37ac42457c59c0b8b3b643e8e1e459d1def85e9c003d32a4cfb864a53c2ff

Observation 38596570-af66-4a90-b887-c02921ad0021 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:21.057268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.335855Z digest=sha256:0fd74cb35f7d2f5fee39b2f2367262bfe4ebab7a5936815d013877d3d2f7d170

Observation 4fa13806-999d-4bde-8b58-a4f1c3bd6a88 · outbound

This paper cites J.; Heess, N.; Precup, D.; Kim, K.-E.; and Guez, A.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning J.; Heess, N.; Precup, D.; Kim, K.-E.; and Guez, A

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:23:20.936927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.338817Z digest=sha256:de4b62e21be2a516afc3f21f00153f55499da78677153cf7701dc8bc2a0a7507

Observation 5d52f9e5-6cbe-4387-8501-c6aeb02f69fc · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.361615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.361615Z digest=sha256:e4276f16b6d3503d5458b477d1d2bbc0203625d00adbe38da00b7c31feee5ff7

Observation 7b99d604-a883-48f2-8f7e-9ca32a7cf590 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.674737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.447986Z digest=sha256:02379cccfb8d8c131b77f50a6f79ed9f6bdc24c22100c4a0f86f1a85fa87c146

Observation 321c3bef-755d-4428-9654-c71765b8a39c · outbound

This paper cites Deep Reinforcement Learning: An Overview.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Deep Reinforcement Learning: An Overview

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.508152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.508152Z digest=sha256:e129955fcad1e8ee505e937eeff58327d64a585304169c697c48884cb60d37bc

Observation 54023e3f-4198-48a5-8b83-56adbcef4d0d · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.662956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.550161Z digest=sha256:bcc2e12581223cd7ff04fdc7d14fda4a134849767859187917203510e2363e7b

Observation 5d5f2d17-0da5-4732-8a77-b087c51737df · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.651304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.566103Z digest=sha256:0b0464cc5e859f1edc3fcdeaed96999ceedbf5b0c51df96f6dd37f92984c978e

Observation 70d9383a-6fac-4c63-89ea-a16981dc2d7a · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.638739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.614034Z digest=sha256:3eb6b9d008c68e14e7ad8bda41c89ad0e404d3de4d8a52c5bcfbf0b9dc891c1f

Observation ad6d6f27-da06-41ee-9659-4a410cba6941 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.626619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.665758Z digest=sha256:8f31b4f8101b5a7f248570b2cc5e4d3a671ac7a823ab16e0ce22581471c6e4bf

Observation e5185fae-1aeb-4836-9dc4-30c3403cb186 · outbound

This paper cites A.; Veness, J.; Bellemare, M.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning A.; Veness, J.; Bellemare, M

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.675812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.675812Z digest=sha256:74b131fdba0e72dc9521028750342ca0adb3a70651ac4f60ff0b7a4d96279fdb

Observation 1cc6bd60-2579-4f14-ab1d-e9df5c5936f3 · outbound

This paper cites C.; Fiterau, M.; and Jagannath, J.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning C.; Fiterau, M.; and Jagannath, J

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:23:20.405316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.683954Z digest=sha256:56c0ddb1ce1777e7c8d40060fedb4f63531e096caa9e69c749d6f80e42fd0e09

Observation c51d4465-e89a-44f6-9583-746d3397ec41 · outbound

This paper cites Benchmarking Batch Deep Reinforcement Learning Algorithms.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Benchmarking Batch Deep Reinforcement Learning Algorithms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.687726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.687726Z digest=sha256:da21af50514298550bae3fa1cf215740e2549b3725d09c67bc85b6853e0d182f

Observation d620086e-5a5b-40f9-9ec2-418f34fa9dbc · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.331827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.691470Z digest=sha256:bc01beb55df9c3ed2351b1499898adec62365670d985a770f2dc996d9ceede26

Observation f4729089-cd4c-4763-bcaf-83286fdcc247 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.695164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.695164Z digest=sha256:da100633035c3845deb979f66374d1f6dbe7ee30293a2bccba092f30773b7a47

Observation ea40c241-e076-43f7-90ab-d695452c849a · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.312487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.698536Z digest=sha256:120fbbe5c4149ba48e2d68e77f9804b54491a45f205adb84ab288cbe52e9495b

Observation 5aaf7a36-fd67-4b25-84cb-79827e3b60b8 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.300214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.701904Z digest=sha256:4317ab9c027c177dbf15b702cb0409a0c0c82063962427d328b96e221df1d176

Observation 45b18044-9dc7-452d-b966-8c68bfae3b28 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.258001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.705343Z digest=sha256:526b9a4161146929821b26633909776c8346e94a5ebb0e0e76b80ccfe31c283b

Observation 1f505e17-f923-4a88-a588-9f89c5ad7b8f · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:20.059232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.708943Z digest=sha256:d3c82627a3482a24d477f5e0a32fd213262c15ff68d65767b8db2c86d94cbb0d

Observation 296b75f0-eff9-427c-b45d-e722464228d3 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:19.961695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.712519Z digest=sha256:0235b62befc7b696496c360f45c8bd479461abb14237acdf91a5d60a9f77f490

Observation 4970c9a3-0f4d-427b-9438-03b6f9eb7e1e · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:19.947758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.716481Z digest=sha256:ff72cb91a2083451886d8bb50d06066d8bbe07cad8170a5ae237b2b36dfd2613

Observation 1affd802-6f26-4dba-84db-575ecdd7ce90 · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:19.934935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.719862Z digest=sha256:38186c452bf6ae5cb0458860b6f630b59fa9c1717e12c2fe1a9d20e36ab098ee

Observation 3a2688dd-6cd6-4e8a-8eec-1e171aba4123 · outbound

This paper cites Y.; Levine, S.; Finn, C.; and Ma, T.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Y.; Levine, S.; Finn, C.; and Ma, T

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:23:19.922287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.723661Z digest=sha256:b71cbb5f27ce68fc233728c735fc613f131f9e8d10224d16e1cbee5031127b1e

Observation 29f42a07-7564-4122-9805-5a0acd96bd3e · outbound

This paper cites an unresolved cited work.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-11T04:23:19.909922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.727060Z digest=sha256:ff42cb21ba12b062089ba353e97ed8b2c05608980587db163236e05e23046cc2

Observation bc69dcd4-1292-418f-8c4b-93277e85ff4f · outbound

This paper cites E.; Zhan, X.; and Liu, J.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning E.; Zhan, X.; and Liu, J

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:23:19.876907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-11T04:23:19.730357Z digest=sha256:530cb1e271a5516c3f7ab0afa343d5ae1ee48fe2e85bf122949ba328846ec7c8

Observation e4000856-f2f7-4ca9-aa90-dacd0444dcc6 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning , " * write output.state after.block = add.period write newline

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.734001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.734001Z digest=sha256:a23b1aab1ccefaf736c712cf93c5033ef30a027c4eac38d5cd85264680874c1c

Observation 5298eb7d-a12c-42cf-8c1b-def54c197495 · outbound

This paper cites write newline.

Constraint-Adaptive Policy Switching for Offline Safe Reinforcement Learning write newline

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T04:23:19.738484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T04:23:19.738484Z digest=sha256:d3edb9962414e13fc135925ea70bc9082e6e789e30358415cd539b4f3c01b234

Pith citing papers

No inbound Pith citation observations are available.