Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T08:48:14.776754Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 100 of 114 outbound references and 63 inbound Pith citation observations for arXiv:2409.00588.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T08:48:14.776754Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-03T19:03:03.126259Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T09:59:44.479162Z
100 of 114 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7fc3da64-cce0-4485-b8e8-fcd4bef622ea · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa2e04cf-b36c-48a7-a144-549378347b5e · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5c6c449-52ec-4354-bb77-d3ac64bdc76d · outbound
Diffusion Policy Policy Optimization Residual Reinforcement Learning from Demonstrations
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 13205ca0-4611-47f5-b046-f2e730159781 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 037f6897-8e31-4219-b4ce-3dabb7d8ebd5 · outbound
Diffusion Policy Policy Optimization Ankile, A
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 74c87f12-a7d2-44d5-a6ae-f284aaa2d218 · outbound
Diffusion Policy Policy Optimization From Imitation to Refinement -- Residual RL for Precise Assembly
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8e01698d-6a82-40f7-93f7-5329ffdc77dd · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0f0df64d-6db3-4e5e-b853-a4ace7053ec5 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89beb46f-a3fd-4a26-a0f2-2e66b5fd0bf5 · outbound
Diffusion Policy Policy Optimization Training Diffusion Models with Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4119c290-68af-456a-8dcf-fd4ca288b698 · outbound
Diffusion Policy Policy Optimization Block, A
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e65f2b57-bd45-499e-ac9d-0847cc47d7c1 · outbound
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b0bb4f68-9b28-4ddc-91c5-15129b1e15fe · outbound
Diffusion Policy Policy Optimization Brown, B
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1c8aa3e9-9d7f-4469-9a53-d60a1890d5ba · outbound
Diffusion Policy Policy Optimization Bruce, M
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 62b94b51-a62f-45a8-a2df-08784e9f42be · outbound
Diffusion Policy Policy Optimization Tutorial on Diffusion Models for Imaging and Vision
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f6303ee-95ea-4f23-b7ae-bcb49599f045 · outbound
Diffusion Policy Policy Optimization Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f6973f3-5d8a-48a4-bcfe-65f9796f460d · outbound
Diffusion Policy Policy Optimization Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b2db49b4-bb19-4625-8133-0cdf3387f210 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2da8068c-b370-4e02-96eb-99988b3b96ab · outbound
Diffusion Policy Policy Optimization Sequential Dexterity: Chaining Dexterous Policies for Long-Horizon Manipulation
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3722d88f-12eb-4f8b-a70b-40ebb751f6ad · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ce1b9978-aeab-4a09-b4c7-f8e851a084f8 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9388e5ab-fe5b-4e74-859d-4f6db40909d0 · outbound
Diffusion Policy Policy Optimization Directly Fine-Tuning Diffusion Models on Differentiable Rewards
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b5700202-7739-4162-9ac2-0a8b20910eda · outbound
Diffusion Policy Policy Optimization Consistency Models as a Rich and Efficient Policy Class for Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b6fd875f-1610-48a3-a25e-5fb03b57d5f1 · outbound
Diffusion Policy Policy Optimization Engstrom, A
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f64c0a8b-9824-4093-ade5-fc045b372b82 · outbound
Diffusion Policy Policy Optimization Optimizing DDPM Sampling with Shortcut Fine-Tuning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aca13e95-0d89-4088-b5cc-eb503d7cb804 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bfcfe2a8-0a6f-4dd2-90d5-a339d57c9847 · outbound
Diffusion Policy Policy Optimization Fefferman, S
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cf9196a7-3b82-43da-b8bf-349fa9112a12 · outbound
Diffusion Policy Policy Optimization Florence, L
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89c55f4c-1d9f-49da-a28c-a0353b44d3fb · outbound
Diffusion Policy Policy Optimization Florence, C
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5a6f5bcd-6534-4829-b82c-b488f681c20d · outbound
Diffusion Policy Policy Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c5e61fee-a515-4fbe-acc6-addb25fb0708 · outbound
Diffusion Policy Policy Optimization Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 373d9164-a265-41d8-bc18-049d89df86fd · outbound
Diffusion Policy Policy Optimization Know Your Boundaries: The Necessity of Explicit Behavioral Cloning in Offline RL
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation de2072f5-0ece-47fb-96a7-13ab1a87fafa · outbound
Diffusion Policy Policy Optimization Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7ff373f3-343a-4e3b-9590-1721b0eb3a3a · outbound
Diffusion Policy Policy Optimization Haarnoja, A
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5e66e647-3eb5-497e-9e62-6408778c974a · outbound
Diffusion Policy Policy Optimization Teach a Robot to FISH: Versatile Imitation from One Minute of Demonstrations
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 878d79bc-74fd-4246-b73e-ca6f977d5087 · outbound
Diffusion Policy Policy Optimization IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e5bbb4ca-f2b0-41e3-993f-5cf03bfac378 · outbound
Diffusion Policy Policy Optimization FurnitureBench: Reproducible Real-World Benchmark for Long-Horizon Complex Manipulation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d738a798-3bee-4f2a-b78e-a572ec5163a9 · outbound
Diffusion Policy Policy Optimization Hester, M
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ef94dfb-c94a-4e4b-831b-9aab0bc0d2f7 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b1e2ee15-62d4-4ef8-b3cf-6fe58860d580 · outbound
Diffusion Policy Policy Optimization Imagen Video: High Definition Video Generation with Diffusion Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e2f0e55a-ef98-4bc1-841c-b4476a8e8720 · outbound
Diffusion Policy Policy Optimization Imitation Bootstrapped Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3e2957bf-5e4d-4657-8beb-0dec2eddcacf · outbound
Diffusion Policy Policy Optimization Huang, R
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b4e7401-0600-4b2d-b788-612d3b0ecee6 · outbound
Diffusion Policy Policy Optimization Huang, L
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e4769549-a34a-4c94-8cf0-43f9a838bfda · outbound
Diffusion Policy Policy Optimization Hwangbo, J
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 207af9cd-d9ab-4bb8-8f6f-30f1af284488 · outbound
Diffusion Policy Policy Optimization Policy-Guided Diffusion
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 19cbe77a-fd06-4feb-a8a4-7841946094ef · outbound
Diffusion Policy Policy Optimization Planning with Diffusion for Flexible Behavior Synthesis
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9c6c5695-e309-4556-879d-e509aa1df0cd · outbound
Diffusion Policy Policy Optimization Towards Diverse Behaviors: A Benchmark for Imitation Learning with Human Demonstrations
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 897764b3-e58d-4ce1-82b0-fd7a1eca23ee · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6fdd6061-686d-4869-a20f-04de5dd0d5d3 · outbound
Diffusion Policy Policy Optimization Kaufmann, L
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3e21a21-e1bc-4461-82b7-ef25800d6f75 · outbound
Diffusion Policy Policy Optimization DiffWave: A Versatile Diffusion Model for Audio Synthesis
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 155b40ea-1331-4e01-91cb-ab31e395b473 · outbound
Diffusion Policy Policy Optimization Behavior Generation with Latent Actions
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1356edd8-6914-40d8-9d0b-b7e6692ce847 · outbound
Diffusion Policy Policy Optimization Uni-O4: Unifying Online and Offline Deep Reinforcement Learning with Multi-Step On-Policy Optimization
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 742849f3-951f-4432-a524-cfd741545fc4 · outbound
Diffusion Policy Policy Optimization Learning Active Task-Oriented Exploration Policies for Bridging the Sim-to-Real Gap
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b302e06-57de-4a1b-afc3-b4a893556f3f · outbound
Diffusion Policy Policy Optimization AdaptDiffuser: Diffusion Models as Adaptive Self-evolving Planners
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df6b903f-f4d1-481f-b74c-616082f4f03d · outbound
Diffusion Policy Policy Optimization Continuous control with deep reinforcement learning
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0982a0eb-cce3-41b3-8c02-48d8a9a31d43 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 604884e4-c53e-49c5-9bc8-199d48c873d9 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d991f1a3-6ec4-42a9-b978-4fe163b283cb · outbound
Diffusion Policy Policy Optimization SERL: A Software Suite for Sample-Efficient Robotic Reinforcement Learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f0e52589-5d3b-4192-bf35-4835a00ef1d4 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f15a744a-7a90-4d55-bc57-ef1cc7e6f697 · outbound
Diffusion Policy Policy Optimization Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa2403fc-9764-4ae2-89f9-24f751ce7ca1 · outbound
Diffusion Policy Policy Optimization What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5e7fcb95-9787-46ce-9e07-40e410eec516 · outbound
Diffusion Policy Policy Optimization AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 36a48b8c-c2df-41d0-a54b-62274b621b84 · outbound
Diffusion Policy Policy Optimization Nakamoto, S
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 01610e55-c2d7-408b-880e-25024554c854 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b5f3fc1f-9d8f-40ce-a304-bfc341f8a582 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d16ff11d-0094-492a-b86c-7a6160eb0b50 · outbound
Diffusion Policy Policy Optimization Ouyang, J
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba3952f7-0862-43f8-a735-ebde9901b913 · outbound
Diffusion Policy Policy Optimization Imitating Human Behaviour with Diffusion Models
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 428dd748-3e80-4a20-8a92-146606bf9833 · outbound
Diffusion Policy Policy Optimization Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a0b17cc-8989-40af-8bca-ee7a3b823343 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dc68bc04-70a0-4434-9196-8803b80c98ac · outbound
Diffusion Policy Policy Optimization Interpreting and Improving Diffusion Models from an Optimization Perspective
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b8f59e9e-3868-47c6-9d33-74f7baf4762d · outbound
Diffusion Policy Policy Optimization Peters and S
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0551ad93-1d15-41f8-acd3-b3b5e711fac2 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cd5b7c29-02b9-421b-b43f-410a2df7d7db · outbound
Diffusion Policy Policy Optimization DreamFusion: Text-to-3D using 2D Diffusion
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5de135df-59cb-4786-8bbe-515a5f0d00fe · outbound
Diffusion Policy Policy Optimization Popova, O
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 16c179a2-3336-4052-94c8-4f91798adf1a · outbound
Diffusion Policy Policy Optimization Learning a Diffusion Model Policy from Rewards via Q-Score Matching
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5e2d455d-2da9-46c3-9821-d7c39478237f · outbound
Diffusion Policy Policy Optimization Radford, J
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8bc30cf8-4c1a-4898-8881-acc51565a62e · outbound
Diffusion Policy Policy Optimization Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 499f6d8e-ba36-4e53-b0db-65f384514ed5 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a3805bbf-4098-4eca-b228-8771291a13d8 · outbound
Diffusion Policy Policy Optimization Goal-Conditioned Imitation Learning using Score-based Diffusion Policies
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 663e197e-a569-400a-8227-212c98c82aeb · outbound
Diffusion Policy Policy Optimization World Models via Policy-Guided Trajectory Diffusion
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 45bf2598-3577-4c36-a694-fb57716f0c1c · outbound
Diffusion Policy Policy Optimization Rombach, A
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 191b30f4-1b13-4b29-b0ad-f98b38730f58 · outbound
Diffusion Policy Policy Optimization Ronneberger, P
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3c90cbb-12ec-40fd-aade-ba21f527cd43 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7e7658b8-5ad6-4b95-98a5-b9ebda3c2b53 · outbound
Diffusion Policy Policy Optimization Simple and Effective Masked Diffusion Language Models
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f69d719f-a5cc-474a-afbd-a2b2ae833406 · outbound
Diffusion Policy Policy Optimization High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b0ec06c1-2c5b-4209-879e-aa7929218867 · outbound
Diffusion Policy Policy Optimization Proximal Policy Optimization Algorithms
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation adeff572-726d-4666-86a6-d6c7e1315a17 · outbound
Diffusion Policy Policy Optimization Sohl-Dickstein, E
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5da1fee1-02d7-40db-bc17-97e7664aae02 · outbound
Diffusion Policy Policy Optimization Denoising Diffusion Implicit Models
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4c0fb82d-3ff4-444d-9c07-425a9fa56ca4 · outbound
Diffusion Policy Policy Optimization Score-Based Generative Modeling through Stochastic Differential Equations
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5a73522d-5759-4371-b6b5-48190f546df5 · outbound
Diffusion Policy Policy Optimization NoMaD: Goal Masked Diffusion Policies for Navigation and Exploration
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 356ceedf-4fe0-46fe-8e5f-362fcd95820a · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 713d4dd9-23cf-4335-80dd-543914033f93 · outbound
Diffusion Policy Policy Optimization Unresolved cited work
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dfabe98e-6efe-48ad-8eef-0e80dba5264e · outbound
Diffusion Policy Policy Optimization Todorov, T
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ab7c209e-66ee-42e9-a4df-4d0b04c2c6fb · outbound
Diffusion Policy Policy Optimization Reconciling Reality through Simulation: A Real-to-Sim-to-Real Approach for Robust Manipulation
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2e5a19b2-9bcb-4236-9478-35b9fe1b0f1a · outbound
Diffusion Policy Policy Optimization Vaswani, N
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aba10201-d72d-4a14-8289-7536c0b95c42 · outbound
Diffusion Policy Policy Optimization Leveraging Demonstrations for Deep Reinforcement Learning on Robotics Problems with Sparse Rewards
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 90b479d3-48c4-489b-8680-f9d57dab6a0e · outbound
Diffusion Policy Policy Optimization Reasoning with Latent Diffusion in Offline Reinforcement Learning
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 47df5f88-d8fa-4abf-bd93-f6f80394c194 · outbound
Diffusion Policy Policy Optimization Diffusion Model Alignment Using Direct Preference Optimization
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 315847f5-1fdf-4e79-b817-e2a1c2e27614 · outbound
Diffusion Policy Policy Optimization Wang and E
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation df7a9bab-7142-4db7-b37d-82d3dc65899e · outbound
Diffusion Policy Policy Optimization PoCo: Policy Composition from and for Heterogeneous Robot Learning
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f6229189-5202-48aa-b716-dfd78823ce01 · outbound
Diffusion Policy Policy Optimization Diffusion Policies as an Expressive Policy Class for Offline Reinforcement Learning
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ba5ee6c1-f15f-4b12-aa3b-cca09a09d237 · inbound
DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization Diffusion Policy Policy Optimization
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0c4db814-78d9-41be-9964-269ad20edbc9 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control Diffusion Policy Policy Optimization
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 47bc4938-94c5-4022-aabc-3252f79a485d · inbound
ReCogDrive: A Reinforced Cognitive Framework for End-to-End Autonomous Driving Diffusion Policy Policy Optimization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 21565e65-1f4c-4528-880f-1c99c744359f · inbound
Steering Your Diffusion Policy with Latent Space Reinforcement Learning Diffusion Policy Policy Optimization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 7f6d4f1a-cac3-4bb2-a553-46a2aed6c269 · inbound
Reinforcement Learning with Action Chunking Diffusion Policy Policy Optimization
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a81dc09f-e3b4-45b2-b9b5-c3f9ace8201b · inbound
EXPO: Stable Reinforcement Learning with Expressive Policies Diffusion Policy Policy Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9ff6eb02-3981-4773-ab0e-31de15b9e121 · inbound
AnyPos: Automated Task-Agnostic Actions for Bimanual Manipulation Diffusion Policy Policy Optimization
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2826a469-44ff-4c82-a962-29060e3578a4 · inbound
Reinforcement Learning with Discrete Diffusion Policies for Combinatorial Action Spaces Diffusion Policy Policy Optimization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 578e1ed8-5be0-4d91-ae7e-10add8cfa854 · inbound
AID: Agent Intent from Diffusion for Multi-Agent Informative Path Planning Diffusion Policy Policy Optimization
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation cc1aa7b9-c440-4da9-8d13-fac8edc2e47e · inbound
Training Diffusion Policies via Prior-Mapping Co-Evolution Diffusion Policy Policy Optimization
Reference 2015
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2500e2c8-52eb-424e-b955-b2c9b97629c8 · inbound
SuperFlow: Training Flow Matching Models with RL on the Fly Diffusion Policy Policy Optimization
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11ceecf6-5757-4424-bb6f-8d9ac2889b82 · inbound
Self-Imitated Diffusion Policy for Efficient and Robust Visual Navigation Diffusion Policy Policy Optimization
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9c845c4-cb15-40fc-ac0f-722adfefde5a · inbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Diffusion Policy Policy Optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da956c23-11d5-4ca9-9175-67ddb8e53664 · inbound
ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training Diffusion Policy Policy Optimization
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f02eaebd-af78-4b47-a2bd-066653aa463b · inbound
RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Diffusion Policy Policy Optimization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 314312b9-a112-4e39-acc5-a801b507f8b3 · inbound
Space Syntax-guided Post-training for Residential Floor Plan Generation Diffusion Policy Policy Optimization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fd84c1b8-faff-46d5-8b92-c225d1c16a09 · inbound
What Does Flow Matching Bring To TD Learning? Diffusion Policy Policy Optimization
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e2a78e72-78c8-4cad-8e4e-acd0832556f6 · inbound
GeMPO: Generalized Measure Matching for Online Diffusion Reinforcement Learning Diffusion Policy Policy Optimization
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bde7aeb-4605-4e85-85ca-017a64f2e514 · inbound
From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning Diffusion Policy Policy Optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1849747a-3200-4c30-b7ed-693d3d74cb5e · inbound
You've Got a Golden Ticket: Improving Generative Robot Policies With A Single Noise Vector Diffusion Policy Policy Optimization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a000a8c-5c2a-4e34-91c8-833fcf93f83e · inbound
ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors Diffusion Policy Policy Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8a6824f1-9b00-495e-99b3-66203b1639fc · inbound
Redefining End-of-Life: Intelligent Automation for Electronics Remanufacturing Systems Diffusion Policy Policy Optimization
Reference 124
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2d37fa2e-fa61-4baa-8f87-e7ae3f2b2e0a · inbound
ScoRe-Flow: Complete Distributional Control via Score-Based Reinforcement Learning for Flow Matching Diffusion Policy Policy Optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c7199917-fcdf-46e4-a5e2-1629a42e72cb · inbound
StableIDM: Stabilizing Inverse Dynamics Model against Manipulator Truncation via Spatio-Temporal Refinement Diffusion Policy Policy Optimization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c8fbe686-b202-4159-8dc1-9a27a2657eda · inbound
One Step Forward and K Steps Back: Better Reasoning with Denoising Recursion Models Diffusion Policy Policy Optimization
Reference 209
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9ef6a3d8-ee54-40a8-9c76-ada06fbaa4cc · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Diffusion Policy Policy Optimization
Reference 171
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2af59ee1-3ad0-4010-935a-3ed5a7f4f754 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Diffusion Policy Policy Optimization
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 647b54c2-7905-4e05-a7ef-55c20243046c · inbound
ReflectDrive-2: Reinforcement-Learning-Aligned Self-Editing for Discrete Diffusion Driving Diffusion Policy Policy Optimization
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ec7fcb19-3c94-4d1e-a239-64f67249a527 · inbound
ReflectDrive-2: Reinforcement-Learning-Aligned Self-Editing for Discrete Diffusion Driving Diffusion Policy Policy Optimization
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 24450e5d-774f-45c1-a1a4-5e2f969605ef · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Diffusion Policy Policy Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ef41667e-d9f4-4a3f-aa51-e30b0168500d · inbound
Long-Horizon Q-Learning: Accurate Value Learning via n-Step Inequalities Diffusion Policy Policy Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation b63fad67-3602-400c-bb82-0aba5003c69d · inbound
BrickCraft: Visuomotor Skill Composition with Situated Manual Guidance for Long-Horizon Interlocking Brick Assembly Diffusion Policy Policy Optimization
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 14a2bfd0-4c76-4480-abf7-888cda1d2df1 · inbound
Driving Intents Amplify Planning-Oriented Reinforcement Learning Diffusion Policy Policy Optimization
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f8ad26a-967f-4dd1-b434-c5f8df1f7fd6 · inbound
Driving Intents Amplify Planning-Oriented Reinforcement Learning Diffusion Policy Policy Optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a866d89e-0557-4c40-bdbd-0ee324810524 · inbound
EponaV2: Driving World Model with Comprehensive Future Reasoning Diffusion Policy Policy Optimization
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9bd20706-5a50-47fd-bd4d-6997829c682f · inbound
Global Convergence of Sampling-Based Nonconvex Optimization through Diffusion-Style Smoothing Diffusion Policy Policy Optimization
Reference 180
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation fa0958ed-9273-4c4f-9083-cdd715b9a03c · inbound
Beyond Execution: Static-Analysis Rewards and Hint-Conditioned Diffusion RL for Code Generation Diffusion Policy Policy Optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9fe43d4b-6039-4d0b-9355-5fb987ec830c · inbound
NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control Diffusion Policy Policy Optimization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5b537296-f68f-4ea4-8ca3-56e1747d33fc · inbound
NaP-Control: Navigating Diffusion Prior for Versatile and Fast Character Control Diffusion Policy Policy Optimization
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a2c95a8-dfaa-4590-bfcb-9822ae1aaf70 · inbound
SCRIPT: Scalable Diffusion Policy with Multi-stage Training for Language-driven Physics-based Humanoid Control Diffusion Policy Policy Optimization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dfd762db-b4d8-4092-912e-22a5e0bae85f · inbound
Score-Based One-step MeanFlow Policy Optimization Diffusion Policy Policy Optimization
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6de1a311-b8a7-4a12-b72a-95fde042912b · inbound
Adversarial Dual On-Policy Distillation from Expressive Teacher Diffusion Policy Policy Optimization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4d781a1a-0f11-459a-bd6b-f7cc45f515cd · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models Diffusion Policy Policy Optimization
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 830c31ec-6ea0-45ac-8b4d-446e12a6f139 · inbound
Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance Diffusion Policy Policy Optimization
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation af852c71-a5dc-4c63-ac5f-924ee87058f9 · inbound
Lagrangian Perturbation Diffusion Steering: Latent Reinforcement Learning for Generative Policies Diffusion Policy Policy Optimization
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 99f79ee2-17f2-4be2-8652-c6e9cc4b1b3b · inbound
L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation Diffusion Policy Policy Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b8fe517-97fc-4df9-83de-54393ea24449 · inbound
GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios Diffusion Policy Policy Optimization
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1b3afeba-f5de-4e9c-ba21-a2b48e66f082 · inbound
Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Diffusion Policy Policy Optimization
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9f694d9c-246a-488f-b50c-04c35aba46ff · inbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Diffusion Policy Policy Optimization
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 89153b87-9389-418d-a987-6fc42e9c112d · inbound
Training and Evaluating Diffusion Policies with Long Context Lengths Diffusion Policy Policy Optimization
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd56ff9e-6c06-4a90-9f40-1ff194b3a13d · inbound
DF-ExpEnse: Diffusion Filtered Exploration for Sample Efficient Finetuning Diffusion Policy Policy Optimization
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ff475354-aa11-47bc-b37e-f498c9ceca88 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Diffusion Policy Policy Optimization
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8dcde12d-1cad-4f93-b565-a567e860bd7c · inbound
Support-Constrained RL Enables Real-World Policy Improvement without Real-World Experience Diffusion Policy Policy Optimization
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5bd0025f-7a96-4793-96fb-5b517d53088e · inbound
RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation Diffusion Policy Policy Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f29df593-48d9-405d-b8ed-45309a366282 · inbound
FAR: Failure-Aware Retry for Test-Time Recovery and Continual Policy Improvement Diffusion Policy Policy Optimization
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 27ba7173-50ca-4891-b91a-935619f026d8 · inbound
WorldSample: Closed-loop Real-robot RL with World Modelling Diffusion Policy Policy Optimization
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 0a61ad18-c644-4b46-b3eb-f0a62b6fc8de · inbound
PAC-ACT: Post-training Actor-Critic for Action Chunking Transformers Diffusion Policy Policy Optimization
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e527f006-6136-4ae1-b4ce-ed5878d95496 · inbound
Source-Lifted Flow Matching for Intervenable Multimodal Imitation Diffusion Policy Policy Optimization
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c643290a-e972-4984-97c1-b0b32927e796 · inbound
A Single Diffusion-Policy Controller for Multi-Task Block Pushing with Zero-Shot Sim-to-Real Transfer Diffusion Policy Policy Optimization
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a346701a-1427-4467-94a4-0992f343157d · inbound
Reinforcement Learning: From Algorithms To Foundation Models Diffusion Policy Policy Optimization
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61f6c394-e367-4202-bd4d-14ec0c057c0a · inbound
Twins: Learn to Predict Unified Representations with Focal Loss Diffusion Policy Policy Optimization
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a66dba47-35d2-44f2-9e09-d6587bece679 · inbound
Do You Really Need to Pretrain Q-Functions for Online RL Fine-Tuning? Diffusion Policy Policy Optimization
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f13788d-107f-4628-ba61-208e54ae897a · inbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Diffusion Policy Policy Optimization
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.