Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T07:55:31.706717Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2602.02924.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-16T07:55:31.706717Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-29T22:46:27.179341Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
32 of 32 outbound references displayed
External citation measurements
0
pith, observed 2026-08-05T02:28:24.338817Z
Observation d27449cf-6ef1-4339-a740-b8bda2ffa502 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Iterated Denoising Energy Matching for Sampling from Boltzmann Densities
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a1dfbca6-728b-40fd-b9f4-045f21b29428 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Solving Inverse Problems via Diffusion-Based Priors: An Approximation-Free Ensemble Sampling Approach
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5f35fc00-60ee-4106-8f46-195e62be1489 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Safe and stable control via lyapunov-guided diffusion models.arXiv preprint arXiv:2509.25375
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation de7f3371-2582-4361-b633-9094833e8c62 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Data-Driven Hamiltonian for Direct Construction of Safe Set from Trajectory Data
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3c1a766-0652-415a-a8f8-b41e117e5683 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Maximum Entropy Reinforcement Learning with Diffusion Policy
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3ccd2e71-4895-4cdb-bb92-e543eaf7ca3e · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a6031ea8-fbfa-4d29-9628-0b96b7c1edb2 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Planning with Diffusion for Flexible Behavior Synthesis
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3966d4ec-0b61-4528-962b-f38bcf1efb0d · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Model-based constrained reinforcement learning using generalized control barrier function
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 343df86c-0770-46d2-8ef5-e68026b1daa1 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Efficient Online Reinforcement Learning for Diffusion Policy
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c3abb07d-e78c-4e3c-9ca8-c884306e4681 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Flow Q-Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 47c4faa7-5cc4-4405-9729-7c73890d20fd · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Learning a Diffusion Model Policy from Rewards via Q-Score Matching
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 847e8c1e-8d51-47eb-99d3-9654f00b8c61 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Sablas: Learning safe control for black-box dynamical systems.IEEE Robotics and Automation Letters, 7(2):1928–1935
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e9c845c4-cb15-40fc-ac0f-722adfefde5a · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Diffusion Policy Policy Optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation ce658ca0-de2a-4bf2-a677-01a1da4f535d · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? High-Dimensional Statistics
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation d52239c0-fa27-433b-8848-6d136be14113 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Solving Stabilize-Avoid Optimal Control via Epigraph Form and Deep Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b451dd2-24f1-41b6-a0f5-becf0538f25e · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Score-Based Generative Modeling through Stochastic Differential Equations
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3beb648f-3f11-44c4-93f7-7ea03dd05792 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? DeepMind Control Suite
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 06d88804-9297-4977-ab65-33f3b10ca3e4 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Reward Constrained Policy Optimization
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 35e80717-c1a4-4fb7-9260-5350b2eba6ae · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? and Schwartz, A
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d519c0f-f77d-4afa-b5b7-a834da9d8e8d · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Understanding Reinforcement Learning-Based Fine-Tuning of Diffusion Models: A Tutorial and Review
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 77cd936f-939d-43e5-997f-1270f5770f7f · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Diffusion Policies as an Expressive Policy Class for Offline Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9d0ad8ed-6e32-4fcb-b826-9748be050b13 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Off-Policy Primal-Dual Safe Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dc6c532d-0623-4c1c-a351-397df8345958 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Constrained Diffusers for Safe Planning and Control
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f19839a0-4bd1-45a2-8b15-b24525f0c226 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Discrete GCBF Proximal Policy Optimization for Multi-agent Safe Optimal Control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 11a28bdb-dcf0-47f2-a01e-f4a1b9bcdefc · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 65b5515a-5638-4e24-b7cf-a47ebc146b86 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? dual variable) τdiffusion step 12 Augmented Lagrangian-Guided Diffusion Appendix Overview This appendix is organized into four main parts
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2c865bdc-9d5b-44e8-98d3-0017c5768256 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? However, these approaches are largely restricted to the offline setting
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1ab181eb-2975-4372-9b46-629a4fbd8561 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Despite recent progress, most existing diffusion-based approaches remain confined to the offline reinforcement learning setting
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation da782dae-1890-4c7c-90db-dd791b9f668d · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? In the following proposition, we present a method for estimating the exact score function for Lagrangian-guided diffusion under the VE SDE framework
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4a828c4a-d893-4c03-a81e-0c718ff3a041 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4ff6b2d8-4b40-43b5-bf91-8689f8bbef71 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? Z K 0 q dσ2(τ) dτ −1 × dσ2(τ) dτ ˜ϕA(s, aτ , τ)−ϕ ∗(s, aτ , τ) 2 dτ # = 1 2 Eπ0(a0|s)
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9abf9253-4ceb-4e06-a7af-f37c0cf79df4 · outbound
How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models? To rule out potential confounding effects, we evaluated the use of cost critic ensembles in the baseline methods, including SAC+Lag and CAL (originally proposed with ensembles)
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation dc9e0a22-3766-44bb-acd6-67d28e85a92e · inbound
SafeDiffusion-R1: Online Reward Steering for Safe Diffusion Post-Training How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models?
Reference 131
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f757a7ce-a53d-42c8-967a-03c4fb450188 · inbound
Scaling World-Model Reinforcement Learning Through Diffusion Policy Optimization How Does the Lagrangian Guide Safe Reinforcement Learning through Diffusion Models?
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.