Pith. sign in

Paper Citation Record · LEDGER

COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2204.08957.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2204.08957 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-11T23:08:40.265657Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:07:28.393682Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 76cb31d5-6693-467c-b868-ca0fa6697526 · inbound

Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting cites this paper.

Fitted $Q$ Evaluation Without Bellman Completeness via Stationary Weighting COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:08:19.145438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:05:34.417424Z digest=sha256:d79edb939d01fc12872148aebf3678a93e8759815655e28e2d06db32e1faaa9e

Observation 9d06994d-7169-417d-8974-917ee61ed2d0 · inbound

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration cites this paper.

Stationary Reweighting Yields Local Convergence of Soft Fitted Q-Iteration COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:58:23.629440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T19:53:23.616808Z digest=sha256:e0599bab12fd9124b7f84554c9388c8e43fc67a53c51e7b72369d10dbe8ce7db

Observation f9404a5b-6301-41be-9f85-ad4edf620d85 · inbound

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning cites this paper.

Decoupled Guidance Diffusion for Adaptive Offline Safe Reinforcement Learning COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:36:09.890678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T15:44:36.262834Z digest=sha256:600b2180f945adbc2fd2bcb9cab553ac3cd521158b2e5a58539dc41e3e987031

Observation eba810d0-f44a-44a6-8b2c-33336c479716 · inbound

Safe-RULE: Safe Reinforcement UnLEarning cites this paper.

Safe-RULE: Safe Reinforcement UnLEarning COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:07:28.395023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T17:28:09.686163Z digest=sha256:fe18ec8b28f2dc85bdb4e724850b4b54696687d9e26632acd9f4c5b4a7b79d20

Observation a707e98a-f7d7-4455-9fec-02180bf40a8b · inbound

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning cites this paper.

CDCP: Conditional Diffusion Model with Contextual Prompts for Multi-task Offline Safe Reinforcement Learning COptiDICE: Offline Constrained Reinforcement Learning via Stationary Distribution Correction Estimation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T23:08:40.265657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:08:40.265657Z digest=sha256:63959353da7d7e13461c34c76b7bffc23bd4753a219a3d0e8195b563212337ae