Pith. sign in

Paper Citation Record · LEDGER

Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.20555.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.20555 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:49.396954Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T04:17:36.912565Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ca30f5ef-759b-47ad-97b6-76f6c65d8374 · inbound

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning cites this paper.

FlowQ: Energy-Guided Flow Policies for Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:49.396954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:49.396954Z digest=sha256:fe221609fea9fcf179fe9d3aba3834947ad5ecf61f2e9268a1a945ab5458fe96

Observation 385977de-291c-443e-9bf2-9b99ae46f638 · inbound

EXPO: Stable Reinforcement Learning with Expressive Policies cites this paper.

EXPO: Stable Reinforcement Learning with Expressive Policies Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:05.237520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T05:09:02.111308Z digest=sha256:60456785498a7a3f5c7b34e1826981f1f7c7069a9e92c2f59fd5827bd262b059

Observation fd393471-b746-42dc-8718-6db43ac05ee8 · inbound

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning cites this paper.

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:41:49.076925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-12T02:17:25.783688Z digest=sha256:8cde0a3afa7e9a465cac8d42939ca330608718336720db4dec5aaab52cf6fff9

Observation 96c4c744-305e-43fd-85f9-7dba25b841cb · inbound

Reinforcement Learning for Flow-Matching Policies with Density Transport cites this paper.

Reinforcement Learning for Flow-Matching Policies with Density Transport Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:27:25.842430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T18:55:02.040180Z digest=sha256:7bc25fbe15d7fda81dc3d79d0f6bbf4debebb09c81a34d6bd3cc6ea1bbf704fd

Observation e71de311-483e-478e-a2ed-7fc2a0324cfb · inbound

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning cites this paper.

Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Diffusion Actor-Critic: Formulating Constrained Policy Iteration as Diffusion Noise Regression for Offline Reinforcement Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T04:17:36.914070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T14:05:01.073951Z digest=sha256:ff585405c986ba3383bcd707134615f2d97c2931c3b2c212a3ec641422b6aff4