Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T07:06:28.426709Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2606.13795.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-27T07:06:28.426709Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 960b7e7c-b54b-4f7f-bbee-41bc7b37dc36 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a22fbede-cd96-4da3-a2c9-e8bb06fef7f0 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dfaa3aec-2983-4e79-b3b5-12ef839f308b · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart DFlash: Block Diffusion for Flash Speculative Decoding
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 107974ec-2f41-4bf2-a137-1ae970526c6f · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Training Verifiers to Solve Math Word Problems
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation aa89b0eb-d069-4902-a6f1-8c0dbf805984 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart EXPO: Stable Reinforcement Learning with Expressive Policies
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f34234b6-1274-43aa-baf0-8d82665e6cc1 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7016376a-1395-4f43-a273-9b4b3513c35c · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Imagen Video: High Definition Video Generation with Diffusion Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2b29a6f0-c9f7-4fef-beb7-ffc3689111d0 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Q-learning with Adjoint Matching
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c4e80b62-0b03-4115-ba8e-966dfc9001e4 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Reinforcement Learning with Action Chunking
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bf5b6a4e-37b3-4eb4-9983-2349a56664a5 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 644ab85c-d687-4f57-9b97-ce1248b4bbbe · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Flow Matching Policy Gradients
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 2b9f8ef1-dd90-43e7-bad7-8bb15d3e9c48 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Large Language Diffusion Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9f694d9c-246a-488f-b50c-04c35aba46ff · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Diffusion Policy Policy Optimization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 741582ff-527d-43e1-a9b0-30a634f53803 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Proximal Policy Optimization Algorithms
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7c3e0754-8852-4ef0-911e-0011faceacf9 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0490dc30-4fa1-46f9-9922-76ad4b0bb0d1 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Parrot: Data-Driven Behavioral Priors for Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c428be7f-c122-4f48-b4a8-844c4293810c · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Denoising Diffusion Implicit Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b47c0824-bb87-4ed6-b25d-44a113ac9537 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Seed Diffusion: A Large-Scale Diffusion Language Model with High-Speed Inference
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eb61d956-522b-4463-9064-0dda553ee36d · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart wd1: Weighted policy optimization for reasoning in diffusion language models.arXiv preprint arXiv:2507.08838
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7c67e3c6-2da7-41d7-bd19-863e76a4da89 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart SPG: Sandwiched Policy Gradient for Masked Diffusion Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 705f7fe2-ec23-4009-b52d-3b27e673a867 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Dream-Coder 7B: An Open Diffusion Language Model for Code
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 95866b58-e1d8-4d40-9303-159209c7b898 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart MMaDA: Multimodal Large Diffusion Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 32b5ea62-73c1-44c0-8797-b94df4ebc65b · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Flow policy gradients for robot control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0172605a-8bac-4417-a5f9-dfbc82e844bb · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f4a60010-8b8c-40e2-9770-f1d742b55638 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee9e2e89-74e5-42e6-b1c4-873b56624781 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart The environment transitions to state st+1 ∼P(·|s t, at), and gives the agent a reward rt =r(s t, at)
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e22934e3-26cb-4881-a589-4491db5d6669 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart SPG [Wang et al., 2025] obtains a practical EUBO surrogate for masked dLLMs by exploiting the special absorbing-mask forward process
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d6f7ac-cbfc-482d-a8f4-7f0b43bde1ca · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart This proves Equation (26)
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 340379b8-429d-46aa-937b-f9a255adaf79 · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Algorithm implementation.We implement FPO by updating the model parameters according to Equation (4), and SPG by updating the model parameters according to Equation (5)
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6a5c6ae-62b7-4327-a5df-1e088123639c · outbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Although SPG+DiPOD does not show a clear improvement at sequence lengths128 and 512, it is the only setting that reaches the near-100%regime in the zero-shot setting
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.