Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:37:53.152448Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2507.15356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T15:37:53.152448Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
30 of 30 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a76a9418-e491-4598-86f4-58217c38bf00 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Is Conditional Generative Modeling all you need for Decision-Making?
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30860b91-1577-4040-985c-0b72487a98a8 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Decision transformer: Reinforcement learning via sequence modeling.Advances in neural information processing systems, 34:15084–15097, 2021
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5c97db9-1a8e-44de-80b8-ca512574d786 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Bail: Best-action imitation learning for batch deep reinforcement learning.Advances in Neural Information Processing Systems, 33:18353–18363, 2020
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 47113f23-51c2-46fe-9590-5fe20bba8a3c · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Semi-markov offline reinforcement learning for healthcare
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07fbdc61-538f-4b36-b0b2-7dcb7e3e48d8 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54ae3b4d-ee18-4c8a-b3c5-9c492e7bbd06 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1cc395fd-3a72-403a-8c89-f4f84ef09c21 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Rediffuser: Reliable decision-making using a diffuser with confidence estimation
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fb909b41-48b6-4d68-957f-6c019ca2aae3 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Denoising diffusion probabilistic models.Advances in neural information processing systems, 33:6840–6851, 2020
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 59d1e5b5-c75a-41dc-a1d9-9ee6e0b80539 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Planning with Diffusion for Flexible Behavior Synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f78e584-f1cb-4e89-a48b-2ac6943d9ab1 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Offline reinforcement learning as one big sequence modeling problem.Advances in neural information processing systems, 34:1273–1286, 2021
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85c565e6-3faf-4e2d-bf89-871b02cc75e7 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making MT-Opt: Continuous Multi-Task Robotic Reinforcement Learning at Scale
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40859e80-6de1-459e-a428-baf474a65320 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Morel: Model-based offline reinforcement learning.Advances in neural information processing systems, 33:21810–21823, 2020
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43a1d839-1ac2-437c-b95a-ae85c13f853e · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Adam: A Method for Stochastic Optimization
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1ba8dd4-94a7-4f26-b8c3-81f3c6cf6c5d · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Offline Reinforcement Learning with Implicit Q-Learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da7ba5ce-e97b-4969-8970-4aba80388222 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b257feac-cb78-4411-86fc-7217bf60151a · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Ceil: Generalized contextual imitation learning.Advances in Neural Information Processing Systems, 36:75491–75516, 2023
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48859532-af26-4812-8077-3c8dc2db1d14 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Synthetic experience replay
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce50624-81b0-4c45-bf06-4ee7df5b2a0c · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Double check your state before trusting it: Confidence- aware bidirectional offline model-based imagination.Advances in Neural Information Processing Systems, 35:38218–38231, 2022
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 911ad9aa-3fdb-4b84-92ad-a1fca77d468c · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb06dcd3-7321-4184-945a-729cc75c9bca · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making A survey on offline reinforcement learning: Taxonomy, review, and open problems.IEEE Transactions on Neural Networks and Learning Systems, 2023
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06fe5f51-8a19-4daa-80b3-f52fa83f3fff · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Offline Reinforcement Learning for Autonomous Driving with Safety and Exploration Enhancement
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d14a4f0e-edc8-4391-a6eb-f0d1e84a2d5a · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Deep unsuper- vised learning using nonequilibrium thermodynamics
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0a87f13-2e4f-42f8-b744-8985202f3e58 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Efficient exploration in continuous-time model-based reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03bdb8d9-ee11-4ff5-9c05-efa8396597c7 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Offline reinforcement learning with reverse model-based imagination.Advances in Neural Information Processing Systems, 34:29420–29432, 2021
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fef7d80d-7fe6-4637-91e5-4d6c7920b2ff · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Boot- strapped transformer for offline reinforcement learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d97484b4-55e2-4324-a4a2-df35d918ce21 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Critic regularized regression
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 373773cd-e2a3-4955-a4e8-aa13b83c08dc · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Combo: Conservative offline model-based policy optimization
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ad2f96c3-67bc-47dd-9753-81357462fe27 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Mopo: Model-based offline policy optimization.Advances in Neural Information Processing Systems, 33:14129–14142, 2020
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 147a32ca-1c0b-4248-983f-2c81921b4c16 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Uncertainty-driven trajectory truncation for data augmentation in offline reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 42708a2e-4d3b-4adb-b0ae-58386157e9d3 · outbound
RAD: Retrieval High-quality Demonstrations to Enhance Decision-making Decision stacks: Flexible reinforcement learning via modular generative models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.