Pith. sign in

Paper Citation Record · LEDGER

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

As of 14 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 7 inbound Pith citation observations for arXiv:2508.06804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06804 v1

Coverage vector

measured 4 of 4 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:37:38.873857Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:50:32.796809Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

4 of 4 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47af0d31-d22e-4a23-b863-c54299debc16 · outbound

This paper cites an unresolved cited work.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:37:38.937359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:37:38.870533Z digest=sha256:5b0c5fec0dd54d8e84f3d482aeef76135e10e9f50402a1c8fbf5299928e0be7c

Observation 8485382e-27fc-4013-a240-d932e75c77e5 · outbound

This paper cites Obs dim (State).

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Obs dim (State)

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T22:37:38.926673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T22:37:38.873857Z digest=sha256:56a0a589e8a14fbdd4e93c1e0a4bb6c803df808e9fc05a4059f9164b7822a09c

Observation e1c9652e-a846-42bb-b6cc-536c9be01cd3 · outbound

This paper cites Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.862346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.862346Z digest=sha256:018573482cf5eaa4a0a1609a961458419432b10d327c2da02abf5e84753b7ffc

Observation e420a4f3-8086-455a-98ff-da8b22e1a528 · outbound

This paper cites Denoising Diffusion Implicit Models.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Denoising Diffusion Implicit Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.866648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.866648Z digest=sha256:9ff8cbaea383ffdb223a5d9e43c867c5ab113f65c794ece248d87ba426aa2702

Pith citing papers

Observation 91e47940-5acb-4039-b67d-68e1ce95c787 · inbound

Generative Actor-Critic with Soft Bridge Policies cites this paper.

Generative Actor-Critic with Soft Bridge Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:28.347246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T04:01:08.896356Z digest=sha256:dbd8ab30ee902327a55e34afcf852f595e6cb31ff2d3f65e712ea5988541ba56

Observation 8f257ea9-f95f-44df-82b2-31a7586db65d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:24.032257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T11:57:36.200662Z digest=sha256:1de5c79d0a4aabbd1c4d3f3c563312feefd6270b72992f462edb9151fa7fa2b3

Observation c5857eb3-2e15-494e-beaa-806a16ba233d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-15T11:04:29.593725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:04:29.593725Z digest=sha256:7241dd62c8eb1a6e03f3d752acbbc8429cbcb3387a38364075a8fb161489f5e0

Observation 8c454ebb-7545-4570-ba9b-f2e777b59bbc · inbound

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making cites this paper.

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.972117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T22:25:23.391802Z digest=sha256:8ba6800f314545612380e1f895f97a0b474bd4485a6afd4602938a6fdb76a519

Observation 6f1066a6-1ca8-40b6-a273-ad19b3f0396d · inbound

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation cites this paper.

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.097365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T01:35:03.841581Z digest=sha256:46963a81e9771cee3c7dda2fa52f26dfb91c6175dcb33cdf3936f8e332994639

Observation 039fcb9d-5d4b-43ca-b5c0-989b55dfd01b · inbound

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies cites this paper.

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T05:45:25.222884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-01T05:44:43.682828Z digest=sha256:23bdd9ead0f190e732eb9f370913b0a8fd1f138e14ec7b37aa745e92718bd590

Observation f86f061b-1113-4748-8b5d-c12bbe3b94c6 · inbound

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control cites this paper.

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:50:32.796809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:50:32.796809Z digest=sha256:8ce9e5cc29d30484e3f95caadf892aaaf2fc3f659b9a02c3abaafddceeb8d060