Pith. sign in

Paper Citation Record · LEDGER

Datasets and Benchmarks for Offline Safe Reinforcement Learning

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2306.09303.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.09303 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:20:07.620593Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:28.404423Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 33bf6911-2cbc-4a33-befc-e5bf0747f0d2 · inbound

FAWAC: Feasibility Informed Advantage Weighted Regression for Persistent Safety in Offline Reinforcement Learning cites this paper.

FAWAC: Feasibility Informed Advantage Weighted Regression for Persistent Safety in Offline Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T17:33:37.634783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:33:37.634783Z digest=sha256:f1ef72b26d095a015c769049af855a80ef42b50d43c33dc29c9e9f32ffae5aeb

Observation 68046617-dae9-4b99-92c5-85eef071965c · inbound

Offline Safe Reinforcement Learning Using Trajectory Classification cites this paper.

Offline Safe Reinforcement Learning Using Trajectory Classification Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T11:30:06.691008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:30:06.691008Z digest=sha256:87daf1dd1b003ddca68479f7140b39bac3d65e9f2e25aeaf5837c89d20374128

Observation a3b4e8e4-3f9d-43d7-93f4-6ba9370317c8 · inbound

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control cites this paper.

From Uncertain to Safe: Conformal Adaptation of Diffusion Models for Safe PDE Control Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T13:03:17.975720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T13:03:17.975720Z digest=sha256:327b279c90156bfb3504ae6aeaeea6f62e3346d4be688013a8932f4dedd1cb68

Observation 97092806-0863-493d-be61-7e200cb4bf3e · inbound

Skill Expansion and Composition in Parameter Space cites this paper.

Skill Expansion and Composition in Parameter Space Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:09.304205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:09.304205Z digest=sha256:7d7eacdba2da9635ef06c3007fa52efb8d2ee89e06531c52fa5907df4cd6ed5e

Observation 1a8ba416-5a3b-41c9-8587-6202ec5da25c · inbound

Skill-based Safe Reinforcement Learning with Risk Planning cites this paper.

Skill-based Safe Reinforcement Learning with Risk Planning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:20:07.620593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:20:07.620593Z digest=sha256:01c7fa305aae019069b855a884eb0cda9682cf16449a3ceea4354f58cc21d3e5

Observation 55ce338c-309c-42ad-aca1-30454c431300 · inbound

PyTupli: A Scalable Infrastructure for Collaborative Offline Reinforcement Learning Projects cites this paper.

PyTupli: A Scalable Infrastructure for Collaborative Offline Reinforcement Learning Projects Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:58:01.250621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:58:01.250621Z digest=sha256:32cf14bfee310bce85637d12921e98a30711fe60fb8daecd77ebfefadb356c68

Observation f0751b9c-ed33-46ac-bf04-49112ec2ab7d · inbound

A Provable Approach for End-to-End Safe Reinforcement Learning cites this paper.

A Provable Approach for End-to-End Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:30:23.525259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:30:23.525259Z digest=sha256:2e06eca8912b5fd2488a5ddcfb80a16fc6a51a56aa12139caf1e298f901bd303

Observation 1057e0d8-bc57-497c-8d62-2ba660bb9523 · inbound

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning cites this paper.

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:09.839172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T15:44:36.262834Z digest=sha256:a919ee2e5f00219d474725e82c65c738c21dc1f9537299af6bd3626dfc51d369

Observation 93f73896-d807-493a-820c-1f05ce9435a7 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.732442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:c3415408130cdf8209cd9b0abf121ebf139a79a9217767c2d65b853ceb5bea5b

Observation d966ac6b-1eb3-4855-adb9-0b41eda9a314 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.827692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:227c8e4f16879ac65490539d54c39972776655c2576bdbb71c9da3fb68f2ec31

Observation ff058c93-975a-4621-9f80-3e6dfff1420c · inbound

Safe-RULE: Safe Reinforcement UnLEarning cites this paper.

Safe-RULE: Safe Reinforcement UnLEarning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.405955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T17:28:09.686163Z digest=sha256:58e832823da55fe307b60e79d21f60aaad3f81ab488feebbb837fbe9e188d523

Observation 141933e7-b4dd-4989-87f1-cd399241c878 · inbound

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation cites this paper.

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:45:43.023784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-01T05:05:51.441647Z digest=sha256:a71332bfea56565c740379c7ae13d94ce6a3c6a741c9c723a7712d7c68d19502

Observation 3c08cf8e-55d0-414e-b187-1e2135b5b822 · inbound

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning cites this paper.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:b5de1aadee460fff159d9d64674bbdc66ec01741e3f41e1055e97b58eef5b737