Pith. sign in

Paper Citation Record · LEDGER

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 4 of 4 outbound references and 7 inbound Pith citation observations for arXiv:2508.06804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.06804 v1

Coverage vector

measured 4 of 4 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T22:37:38.873857Z

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T05:50:32.796809Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

4 of 4 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 47af0d31-d22e-4a23-b863-c54299debc16 · outbound

This paper cites an unresolved cited work.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-05T22:37:38.937359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:37:38.870533Z digest=sha256:911a79a2b891bf42a9487ff7a5461cd3f0cc33e595e13672c0caea33a9cf3266

Observation 8485382e-27fc-4013-a240-d932e75c77e5 · outbound

This paper cites Obs dim (State).

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Obs dim (State)

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-08-05T22:37:38.926673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-05T22:37:38.873857Z digest=sha256:7c1dd436deb4ee87c98e485a27fe97434c902e5e6c2f50d5681873213d6e4d1e

Observation e1c9652e-a846-42bb-b6cc-536c9be01cd3 · outbound

This paper cites Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.862346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.862346Z digest=sha256:19c923e35b81b3ba521812a80b253973d34afb355c6e1feb6f3e5bf1d779c2f5

Observation e420a4f3-8086-455a-98ff-da8b22e1a528 · outbound

This paper cites Denoising Diffusion Implicit Models.

D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Denoising Diffusion Implicit Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-05T22:37:38.866648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:37:38.866648Z digest=sha256:131027d7c2ede3334e0e792f24233c871c1a50ad302c7d33869877486ae65d64

Pith citing papers

Observation 91e47940-5acb-4039-b67d-68e1ce95c787 · inbound

Generative Actor-Critic with Soft Bridge Policies cites this paper.

Generative Actor-Critic with Soft Bridge Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:46:28.347246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:01:08.896356Z digest=sha256:ad89ddcfb432e1acf9bf5774f939f5c3c8c9be2ea399bd8ccf0b2534eea9b865

Observation 8f257ea9-f95f-44df-82b2-31a7586db65d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:24.032257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T11:57:36.200662Z digest=sha256:faf6cc05941a551f64d06023b290535391bf448933bdbbb1de999a1bf101a315

Observation c5857eb3-2e15-494e-beaa-806a16ba233d · inbound

SANTS: A State-Adaptive Scheduler for World Action Models cites this paper.

SANTS: A State-Adaptive Scheduler for World Action Models D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-15T11:04:29.593725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:04:29.593725Z digest=sha256:a14fa9f2df84375df1e9d96af4dfff9589db317275a6ce8d7ce9f651a5a52906

Observation 8c454ebb-7545-4570-ba9b-f2e777b59bbc · inbound

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making cites this paper.

On-Device Robotic Planning: Eliminating Inference Redundancy for Efficient Decision-Making D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:26:00.972117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T22:25:23.391802Z digest=sha256:4655647c71c9cd2be53b67548080c16053f9c95d2792bb824955d0b8a6b11f34

Observation 6f1066a6-1ca8-40b6-a273-ad19b3f0396d · inbound

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation cites this paper.

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:06:59.097365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T01:35:03.841581Z digest=sha256:26adc0d720708c20a331a2d638ade8b1b67dae1732b042310ad3c721f224e82e

Observation 039fcb9d-5d4b-43ca-b5c0-989b55dfd01b · inbound

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies cites this paper.

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T05:45:25.222884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T05:44:43.682828Z digest=sha256:c39b53349b596dc8964bec8e8f6927a8231b046558cafa151f6413cd0eef4d7c

Observation f86f061b-1113-4748-8b5d-c12bbe3b94c6 · inbound

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control cites this paper.

Learning When to Stop: Prefix-Optimal Dynamic Diffusion Policies for Continuous Control D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T05:50:32.796809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:50:32.796809Z digest=sha256:c8a3f1b894cd47a7ec3fc67552344467e70d82a857c73c263cf4bf83ef8b2daf