Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:55:54.416203Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.23871.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:55:54.416203Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:15.443445Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:01:17.482077Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f07df930-7dab-4be3-9ff6-a941c0221bf2 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Ambient Diffusion Posterior Sampling: Solving Inverse Problems with Diffusion Models Trained on Corrupted Data
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6806c56b-80d0-4f93-bbda-f07b6239954e · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Offline Reinforcement Learning from Datasets with Structured Non-Stationarity
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ae3d6006-5349-4ead-8aad-7106da10dd1f · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Is conditional generative modeling all you need for decision making? In The Eleventh International Conference on Learning Representations, 2023
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4193ddd7-63bf-4de2-b501-5cff4344046b · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified q-ensemble
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8fc1e53c-519f-4b0e-996e-c04cbfd6fe86 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Pessimistic Bootstrapping for Uncertainty-Driven Offline Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0419cc4f-e479-4089-883d-e7c0c1657164 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Decision transformer: Reinforcement learning via sequence modeling
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b641ced-b476-4710-bf8c-e9cf78a61c58 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Exact policy recovery in offline rl with both heavy-tailed rewards and data corruption
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 901b35f6-a897-4991-b54a-fedcaa433656 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Consistent diffusion meets tweedie: Training exact ambient diffusion models with noisy data
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf4502ff-5025-487b-8de7-5dd4ed9c70ca · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c858ef78-624f-42f4-acee-de67074b791c · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning A minimalist approach to offline reinforcement learn- ing
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70525bd3-df61-4f38-8570-38c8e52a4f78 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Off-policy deep reinforcement learning with- out exploration
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ce47875b-e531-4665-b6ec-30fae9c7a8db · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7bb711e0-81f2-4b51-8742-16ae564a9b2c · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 044294f5-e2f2-4677-a284-ce6e4a52f95d · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Denoising diffusion probabilistic models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 71b1811b-c4c7-46d2-ad04-06838da71daf · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Planning with diffusion for flexible behavior synthesis
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9dedea89-7914-4726-8c53-7e240630a2fd · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Offline reinforcement learning as one big sequence modeling problem
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e84a4cd2-cafa-47e9-bfc4-3cb71585a727 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Neural stochastic differential equations for uncertainty-aware offline rl
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c4d21d72-cb97-4cb4-8a73-bdef2ef490e0 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51bdfcb9-5299-48f8-92cf-56ca80169906 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Conservative q-learning for offline reinforcement learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 506ed5af-a7b9-447f-ab85-59acdc70cfa7 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Offline reinforcement learning: Tutorial, review, and perspectives on open problems, 2020
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20eb3195-e3be-45c8-8419-0ff7be9c8d52 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Survival instinct in offline reinforcement learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e1eb5478-0d2b-4072-9b61-e5f6d10d193a · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning ROPO: Robust Preference Optimization for Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d4f3c54-dde8-45f3-bd3a-d8f6fc392673 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Adapt- diffuser: Diffusion models as adaptive self-evolving planners
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 81d71945-2dc4-486e-8ea6-8491f145900c · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Adaptive Advantage-Guided Policy Regularization for Offline Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2c5b186-0fa7-42c0-9625-32b4df67b044 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Towards Deep Learning Models Resistant to Adversarial Attacks
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc02cc5-db76-450b-8169-b5808703197b · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Robust re- inforcement learning using offline data
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4fc0fb9-2b0a-4640-86f5-4b5c9827f731 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning High-resolution image synthesis with latent diffusion models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d7ba50d-9111-434f-bda0-5b5d009993a2 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Distributionally robust model-based offline reinforcement learning with near-optimal sample complexity
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fa6ae2c6-7739-44b5-a371-55c774d36064 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03cb4af2-56f5-4eba-9fd3-1e8bf8881c8c · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Deep un- supervised learning using nonequilibrium thermodynamics
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d95c7342-e74a-4a3f-9126-2545d5b18d35 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Score-based generative modeling through stochastic differential equations
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e72e98e9-8ea3-4648-9020-99c453556ca2 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 397d09de-2df4-4cfc-b15f-04cefe7ba833 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning COPA: Certifying Robust Policies for Offline Reinforcement Learning against Poisoning Attacks
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae603eeb-7d7a-43b2-86bc-6f252b7d7d6a · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Tackling data corruption in offline reinforcement learning via sequence modeling
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bec7e5bc-db01-4841-ab8d-ae899f31e0dc · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Rorl: Robust offline reinforcement learning via conservative smoothing
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c2be624d-9675-487d-92de-65d134ad6429 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Regularizing Hidden States Enables Learning Generalizable Reward Model for LLMs
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc6e40a4-8c22-499d-916f-38b0606a70ab · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Towards robust offline reinforcement learning under diverse data corruption
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c23335ab-2839-49a6-ba43-d10898ab60cb · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Dmbp: Diffusion model-based predictor for robust offline rein- forcement learning against state observation perturbations
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 942f13a5-48d8-4637-8b4e-3dcb9e0b5a7a · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02a95ecd-77eb-4bc1-9265-743cc4e120cc · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Corruption-robust offline reinforcement learning with general function approximation
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9ecdb87-ee50-4795-8257-fb15e3942fd9 · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Robust Reinforcement Learning on State Observations with Learned Optimal Adversary
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5871579f-6a8d-4d36-901e-99fe7a7e3c0c · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning Robust deep reinforcement learning against adversarial perturbations on state observa- tions
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3eac0b00-c5a7-4c4f-bb3a-438918897dbf · outbound
ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning medium-replay- v2
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a4d3b03d-0631-4e30-a768-51c4b3a89431 · inbound
Ambient Diffusion Omni: Training Good Models with Bad Data ADG: Ambient Diffusion-Guided Dataset Recovery for Corruption-Robust Offline Reinforcement Learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.