Pith. sign in

Paper Citation Record · LEDGER

Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.19690.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.19690 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:39:25.521307Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8cfc379a-7beb-4dde-b478-966e54f3f4e9 · inbound

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning cites this paper.

Behavior-Regularized Diffusion Policy Optimization for Offline Reinforcement Learning Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T21:39:25.521307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:39:25.521307Z digest=sha256:f96d4d9be9fe0a1d472fdd05661a3985a132e21c2c199858ce805cf880ea552e

Observation 8b469ca3-e1da-43f2-8bf5-0f558bfde476 · inbound

Habitizing Diffusion Planning for Efficient and Effective Decision Making cites this paper.

Habitizing Diffusion Planning for Efficient and Effective Decision Making Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T15:40:15.862799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T15:40:15.862799Z digest=sha256:0a89a3b3774f832dbce66b3da9456093e743e9258df63b201aaa8f6128af26b0

Observation b957a1f2-5ba9-4570-9f48-7ebbc50661a0 · inbound

Steering Your Diffusion Policy with Latent Space Reinforcement Learning cites this paper.

Steering Your Diffusion Policy with Latent Space Reinforcement Learning Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:55:46.476363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T21:55:46.183007Z digest=sha256:fb37fee62504a7047a1ed5333d458b25e60193f416f74b03edf04ab9144adc97

Observation 30bc8fe0-d178-4f77-ac42-81ce3fe4e4d5 · inbound

Offline Reinforcement Learning with Penalized Action Noise Injection cites this paper.

Offline Reinforcement Learning with Penalized Action Noise Injection Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:38:32.141102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:38:32.141102Z digest=sha256:d5cdf9cc0fb61f4b19f305b930d6b0037288c2659fc563bbe26a8b7647e92f29

Observation 6cba423c-9758-4d87-9b04-8579813c1a5d · inbound

Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies cites this paper.

Behavioral Mode Discovery for Fine-tuning Multimodal Generative Policies Diffusion Policies creating a Trust Region for Offline Reinforcement Learning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:17:05.683658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T02:16:53.951387Z digest=sha256:a684b446d9fa52cf71923fd1bdda7e6002363b250234cb98ba7e4be0f565ed3f