Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:21:18.474336Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 0 inbound Pith citation observations for arXiv:2502.07279.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T13:21:18.474336Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 717804df-2c64-4226-83eb-f028019cddf1 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Deep reinforcement learning at the edge of the statistical precipice
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1186c289-b8cd-485a-8f62-13bd7c4d93da · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Tenenbaum, Tommi S
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52013a13-f6fb-4638-a35d-f2950327fed6 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion for World Modeling: Visual Details Matter in Atari
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c04bd1b-5546-4893-bbf7-dd6bb3d3540b · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Random polytopes, convex bodies, and approximation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b5f258a2-ebde-445b-a142-3e292e93ebe1 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Constrained Ensemble Exploration for Unsupervised Skill Discovery
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8f7e1583-b62b-47d7-81c5-d76e8d9c4e00 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Exploration by random network distillation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ce0c823-a0ac-4384-888a-ad82f96ec597 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Explore, discover and learn: Unsupervised discovery of state-covering skills
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f41f3bf-524c-4d7d-b5ec-4ce13f26c927 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning DIME:Diffusion-Based Maximum Entropy Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 749b2cf6-799b-4bcd-ad89-0bd4e41b92a5 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 193876ed-7802-427e-94db-c299f5ddd8df · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Simple Hierarchical Planning with Diffusion
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 249abdf1-5bc1-4cb2-b12b-91895894bbe9 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Offline reinforcement learning via high-fidelity generative behavior modeling
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b00d0d5a-5bcb-4ec0-81b1-cf47f5dbac27 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20988d13-2fcd-4b7a-9bad-507db77e9192 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion policy: Visuomotor policy learning via action diffusion
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8dca37e-e625-40fc-98d4-dffeb92cb3b0 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Posterior Sampling for General Noisy Inverse Problems
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1b9aff1-ba8b-4410-81e3-dbf0334efd1a · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion models beat gans on image synthesis
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bc01a3b-00f9-4a54-837c-210713bbf508 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion World Model: Future Modeling Beyond Step-by-Step Rollout for Offline Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cba315ae-fa8a-4bee-abeb-1f10bb41e019 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diversity is all you need: Learning skills without a reward function
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 54de6cbc-71f0-46b0-a13f-32ea678baa77 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning The information geometry of unsupervised reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e6ded4df-fa16-481d-ac4b-24d104174cae · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Reinforcement learning with deep energy-based policies
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8b9f12ab-cea0-4ef4-984c-bbdd091494d1 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1803d2cc-0e49-4e91-8a24-4828acd0e416 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d90fbec2-cc75-4562-a0d7-519a7e761017 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion model is an effective planner and data synthesizer for multi-task reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 65d6ffed-8431-4bc1-a236-6b8da9905ac9 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Denoising diffusion probabilistic models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09ba4fbd-d15c-448d-86d4-586560b38a47 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Langevin Soft Actor-Critic: Efficient Exploration through Uncertainty-Driven Critic Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64ea5b7d-a4d1-46a6-b3d7-bc072f27085f · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Planning with diffusion for flexible behavior synthesis
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c81c733-6952-4b95-b4ef-2d8cb7a0c569 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient diffusion policies for offline reinforcement learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3204f4ed-570d-404d-a18a-3d0ca8e4e2a1 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Unsupervised skill discovery with bottleneck option learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation edc71b78-a83a-4795-9c54-6c4638cda4b5 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Offline reinforcement learning with implicit q-learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b8f274-5633-4882-8931-202c37755d65 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Unsuper- vised reinforcement learning with contrastive intrinsic control
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cf259c37-eb98-47a3-ad23-3cb9c6052bb5 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Urlb: Unsupervised reinforcement learning benchmark
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4104a14-fb61-4278-a931-3d2289529bc0 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient Exploration via State Marginal Matching
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c091b6d-b0e9-499a-b58b-8156c8a0d9ec · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Hierarchical diffusion for offline decision making
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5d47d839-47e1-4e91-8bf6-b317946c851b · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Learning Multimodal Behaviors from Scratch with Diffusion Policy Gradient
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c2cff4f-45cc-4e34-84b6-a9c51d4e44ad · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Adaptdiffuser: Diffusion models as adaptive self-evolving planners
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bb26dcbb-b4dd-4744-8451-210dee826867 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Continuous control with deep reinforcement learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acacfe4a-278d-49ad-9b0e-3b4e0ab72907 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Aps: Active pretraining with successor features
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6e3ab296-d29d-45aa-8575-4f2845740779 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Behavior from the void: Unsupervised active pre-training
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 436e4d5e-93f0-458f-9091-e6ccdef8eb7a · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Energy-Guided Diffusion Sampling for Offline-to-Online Reinforcement Learning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43d3f187-783d-4d0c-bf6d-39429be652df · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d6dd9ac4-0c13-4e8e-ba70-a58255ea0506 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cab513be-a52c-4aa6-841e-378ec63bd8f9 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Synthetic experience replay.Advances in Neural Information Processing Systems, 36, 2024
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e1fcfedc-9fa5-464d-813c-f463941d62e6 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Efficient Online Reinforcement Learning for Diffusion Policy
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c5de78b-2d75-439a-83d5-9af270e71cc1 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad88db33-c4b8-4834-933b-e230cbf6614d · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Curiosity-driven exploration via latent bayesian surprise
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e36c7829-ae27-43f0-8d40-9e40197bdf24 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Lipschitz-constrained unsupervised skill discovery
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 25988992-dc49-4b2d-b2c8-54d202233e5e · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning METRA: Scalable Unsupervised RL with Metric-Aware Abstraction
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de555d34-5ec6-4044-b0ab-c4e20ecf47c2 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Curiosity-driven exploration by self-supervised prediction
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 360a91dc-9430-4d88-ad63-b63cf128d0fd · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Self-supervised exploration via disagreement
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b957b3af-87ce-4b1b-b3ef-1e63cc384755 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42042865-f0ce-403b-a6f5-260f47240610 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Learning a Diffusion Model Policy from Rewards via Q-Score Matching
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06931fde-8a40-45aa-87d0-7d1de4e1d46c · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Policy Policy Optimization
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2707d75e-063d-4aec-be49-31b33b7b274d · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Photorealistic text-to- image diffusion models with deep language understanding
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db3205d2-3abc-442f-97a3-48c031397268 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbbf8218-1a96-4e75-b11d-10589a6f930c · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Deep unsupervised learning using nonequilibrium thermodynamics
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1c4d90-8598-49f0-8f25-d4866102963a · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Denoising diffusion implicit models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 413bb471-9086-4afb-904a-9a6070a93c95 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Score-based generative modeling through stochastic differential equations
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b3437f2-eb5b-4c53-89d8-2a520ae2f4b7 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Reinforcement learning: An introduction
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55ee6be8-6836-40f9-a899-c5f22e3d4788 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning DeepMind Control Suite
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bce0a378-af0f-4c70-a587-9312f03ea012 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Prioritized Generative Replay
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6594c9c2-0ac2-47fb-9fb0-2a70427a63b8 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion policies as an expressive policy class for offline reinforcement learning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0198e1c6-8411-45ae-951a-24aca56b6b25 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning A problem in geometric probability
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5944f1a6-f8ea-44a9-90e7-e9a2249fdadb · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f30d08a-8ca5-447b-b349-8d05edb0bc5b · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Policy Representation via Diffusion Probability Model for Reinforcement Learning
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78c1c545-d060-40a5-be76-0c7b9837b1b8 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Behavior contrastive learning for unsupervised skill discovery
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a5ec26a0-6739-42d7-9d04-108188dd472f · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Peac: Unsupervised pre-training for cross-embodiment reinforcement learning
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6ba8d291-d21f-4541-8de4-667a1bc26e33 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Towards Safe Reinforcement Learning via Constraining Conditional Value-at-Risk
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5376eb95-3c8f-4f54-84ab-3ed8f27c2477 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Automatic intrinsic reward shaping for exploration in deep reinforcement learning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97a9345c-5642-4da6-94fc-5352cf84b381 · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning EUCLID: Towards Efficient Unsupervised Reinforcement Learning with Multi-choice Dynamics Model
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d7c98b1-be79-47c8-b37e-c988112b88fa · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning A mixture of surprises for unsupervised reinforcement learning
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a33c2e8f-595f-448c-ba50-fecddc257fad · outbound
Exploratory Diffusion Model for Unsupervised Reinforcement Learning Diffusion Models for Reinforcement Learning: A Survey
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.