Pith. sign in

Paper Citation Record · LEDGER

Datasets and Benchmarks for Offline Safe Reinforcement Learning

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2306.09303.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.09303 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:08:40.265657Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:28.404423Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1057e0d8-bc57-497c-8d62-2ba660bb9523 · inbound

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning cites this paper.

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:09.839172Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T15:44:36.262834Z digest=sha256:bc264939c09813761bf3e90a00d541e358a42ae85f0e9bb65d4810f3417526f8

Observation 93f73896-d807-493a-820c-1f05ce9435a7 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:39:40.732442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-21T05:37:29.919862Z digest=sha256:df914daa5644bf006b50473e2d85796619a532bb875289e3e180a76d4fe48351

Observation d966ac6b-1eb3-4855-adb9-0b41eda9a314 · inbound

Behavior-Consistent Deep Reinforcement Learning cites this paper.

Behavior-Consistent Deep Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:06:21.827692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-22T10:06:04.478006Z digest=sha256:19ed6a81cc3b1c0d590fadfd6b4c536db1babc8c26e35c80113c63eedd383081

Observation ff058c93-975a-4621-9f80-3e6dfff1420c · inbound

Safe-RULE: Safe Reinforcement UnLEarning cites this paper.

Safe-RULE: Safe Reinforcement UnLEarning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.405955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T17:28:09.686163Z digest=sha256:0def63707b7a8c143c9df06234ec3253ec76f43ea8bf59dc03c3bf7e0845baea

Observation 141933e7-b4dd-4989-87f1-cd399241c878 · inbound

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation cites this paper.

OopsieVerse: A Safety Benchmark with Damage-Aware Simulation for Robot Manipulation Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:45:43.023784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T05:05:51.441647Z digest=sha256:4f89db64228784c27fbb6faa5b9826214feacb6aa1e196722ed576988f29aaf8

Observation 3c08cf8e-55d0-414e-b187-1e2135b5b822 · inbound

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning cites this paper.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning Datasets and Benchmarks for Offline Safe Reinforcement Learning

Reference 68

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:5f2930c1a15f6da0992bf287c260f3deb0218e75f948ff6ed268d4273a999069