Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T23:19:17.322890Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 100 inbound Pith citation observations for arXiv:2004.07219.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T23:19:17.322890Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T15:52:46.789736Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
24 of 24 outbound references displayed
External citation measurements
332
pith, observed 2026-08-05T02:28:24.338817Z
Observation a3c1607c-ae24-4103-a6d5-ac4d1bf911e0 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning On the Theory of Policy Gradient Methods: Optimality, Approximation, and Distribution Shift
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4ddb077d-4ab4-4e16-8f28-80bfed67cb41 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Scaling data-driven robotics with reward sketching and batch reinforcement learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e9534e1a-3b3b-4f71-b5cc-fe50080a65fb · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning End- to-end driving via conditional imitation learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 27625d21-bc61-4912-97bf-001f9bab866e · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Challenges of Real-World Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bfc646a4-47b4-4068-84fa-e9b213030d6c · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning An empirical investigation of the challenges of real-world reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 22f1d582-6358-4ddc-9f69-86b106be883b · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Off-Policy Deep Reinforcement Learning without Exploration
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f4bc3789-f7d7-4bf4-91aa-81d2fe885d58 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Deep reinforcement learning for robotic manipulation with asynchronous off-policy updates
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b0592225-ea90-47e1-bf98-8704da4577f9 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5c4e3e26-8a0c-4db9-a6d1-16be3f6ac593 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 68df719a-6080-45f7-8205-a8f327926b17 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning A Real-Time Model-Based Reinforcement Learning Architecture for Robot Control
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c6677e0b-b8ef-479f-83b7-a04ff281d78c · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning RecSim: A Configurable Simulation Platform for Recommender Systems
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8850fa17-b592-49b3-914c-3ca670d20d07 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Way Off-Policy Batch Deep Reinforcement Learning of Implicit Human Preferences in Dialog
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ce0e9a44-c3ff-492d-8b41-c1a1bd497106 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Stabilizing Off-Policy Q-Learning via Bootstrapping Error Reduction
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b1bc3922-d0e6-40bc-89f5-08672e92a0da · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dcb8eeaf-0b4c-4764-8eac-50538e18f1bf · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning AlgaeDICE: Policy Gradient from Arbitrary Experience
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 60ec4de0-a61c-4ee8-a7dc-fdc81d508179 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9ca6c1c0-191a-407d-9f18-79a353236244 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fb030031-43ef-4ba9-9115-f9d6f05640f3 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Deep Imitative Models for Flexible Inference, Planning, and Control
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5fc180cd-e726-4dbd-8194-476c6204ed10 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Mujoco: A physics engine for model-based control
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 50a3e41c-c4df-4057-923a-f163b7cf1b81 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Flow: A Modular Learning Framework for Mixed Autonomy Traffic
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f23487c3-36f7-43bf-8abd-e58104d9bd1f · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Behavior Regularized Offline Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a382beaf-5594-4e50-b63d-4ec349c568b9 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning carla-town
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9d668e8c-4184-4d06-ab12-a8f095779585 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning (2017) Random, Controller Franka Kitchen Gupta et al
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4fa517c2-2747-4d59-997a-b9a64c062f02 · outbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Training
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b03cbe9c-e8f5-4b3c-af90-3dc6ed420c31 · inbound
Decision Transformer: Reinforcement Learning via Sequence Modeling D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 96a97e61-ac58-4ab4-83c0-86c1fe1e232e · inbound
What Matters in Learning from Offline Human Demonstrations for Robot Manipulation D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f0990480-5133-486b-98f6-55088ccd70d6 · inbound
Offline Reinforcement Learning with Implicit Q-Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eb7e1477-4135-4d10-aa68-f0701e0b9b29 · inbound
A Generalist Agent D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 056b3df7-809a-4f66-9d87-30044f2e8677 · inbound
Diffusion Policies as an Expressive Policy Class for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2e383054-929e-49e3-aa6c-3220cb086042 · inbound
IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8121b253-57dc-4d88-9350-1a3797ac843d · inbound
Reinforced Self-Training (ReST) for Language Modeling D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6007e803-a731-4848-9346-735f01200ceb · inbound
Promptbreeder: Self-Referential Self-Improvement Via Prompt Evolution D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 115
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b443c3ce-7dc3-4900-abf3-2cee4db9d8d8 · inbound
CROP: Conservative Reward for Model-based Offline Policy Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5a6f5bcd-6534-4829-b82c-b488f681c20d · inbound
Diffusion Policy Policy Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 87d777e8-925b-4e41-a35a-3eb5d2dc0e9e · inbound
DINO-WM: World Models on Pre-trained Visual Features enable Zero-shot Planning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 768d8265-374c-461e-99ae-d608acbccdaa · inbound
SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b98c64b-e5f3-40c2-a30c-c8984397de76 · inbound
Self-Improving Skill Learning for Robust Skill-based Meta-Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f2cadc78-32f1-4156-8ad2-cbb1b08b6c3e · inbound
VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 082498cf-373b-4224-8c6f-e362e0bc2cbc · inbound
Using Ensemble Diffusion to Estimate Uncertainty for End-to-End Autonomous Driving D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 08e7d6ac-5f1e-46a2-a11f-46da1e673db4 · inbound
BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5aad7009-083e-488c-9af3-ed08460e21cd · inbound
Reinforcement Learning with Action Chunking D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation def2e6e4-6bb0-4153-ab55-74b75950ddb4 · inbound
EXPO: Stable Reinforcement Learning with Expressive Policies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a95ed71a-ccb4-4cca-8a15-e917a0b43254 · inbound
Re:Frame -- Retrieving Experience From Associative Memory D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d68f3ad-e4fc-40e6-be86-ce3a31ed6f15 · inbound
LLM-Driven Policy Diffusion: Enhancing Generalization in Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b32a97b-13a5-429d-8f21-7299b022d80f · inbound
Generative Auto-Bidding in Large-Scale Competitive Auctions via Diffusion Completer-Aligner D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 105b4d2d-95fe-4035-8093-997cfb7f6976 · inbound
RaC: Robot Learning for Long-Horizon Tasks by Scaling Recovery and Correction D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbd857a9-8bbf-4418-9d10-9f27dceac450 · inbound
Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31d12c3c-6aa9-4e46-b2e8-bcc925e672c5 · inbound
Geometric Analysis of Neural Regression Collapse via Intrinsic Dimension D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5456f005-3387-4bba-9222-69cef5bd08d8 · inbound
The Three Regimes of Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ff1f605-4b91-41a4-928e-48d7f8e0ade7 · inbound
Distributional Inverse Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3b4b8d2-937b-461b-9751-b200c71aea67 · inbound
Value Flows D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88ac4429-a9b8-48c8-b863-761caf8642c7 · inbound
When a Robot is More Capable than a Human: Learning from Constrained Demonstrators D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8b9c1b37-a5b4-4536-ad75-32bf82f10675 · inbound
From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5187225d-4763-413e-8056-f61c646e943b · inbound
From Static Constraints to Dynamic Adaptation: Sample-Level Constraint Relaxation for Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 35d8367b-ce4e-4e4a-b7b2-c7bb1c9c8b11 · inbound
HardFlow: Hard-Constrained Sampling for Flow-Matching Models via Trajectory Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4de086e7-0914-4087-887c-b83b1a45af85 · inbound
Enhancing Robustness of Offline Reinforcement Learning Under Data Corruption via Sharpness-Aware Minimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d823fb07-b89e-4f25-a906-c48b7bf19298 · inbound
Training Diffusion Policies via Prior-Mapping Co-Evolution D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc9e374-ce41-4c19-8a29-4bd2ca6cb10a · inbound
What Drives Success in Physical Planning with Joint-Embedding Predictive World Models? D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0097d759-404e-46a8-8509-b5575182b672 · inbound
Agile Reinforcement Learning through Separable Neural Architecture and Applications D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bdbc15a-247f-426f-9354-62caed138ed6 · inbound
Optimization and Generation in Aerodynamics Inverse Design D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2009
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8d2bc5c-62e5-4b66-bc8a-af6bf541dbee · inbound
VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a02efa69-4ede-40bb-a01d-524440790899 · inbound
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eb39c09e-d1d7-408d-b62c-421242145b72 · inbound
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9994c1e3-cfea-43ac-a448-9b2ec568ecc5 · inbound
Improving Diffusion Planners by Self-Supervised Action Gating with Energies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49fe07e8-0158-4b31-bc4c-500d4644040f · inbound
What Does Flow Matching Bring To TD Learning? D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0ebca3db-c17c-4070-8574-b1e1eef07451 · inbound
Offline Materials Optimization with CliqueFlowmer D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f5a947da-c221-46c4-b46b-19a01174cdf6 · inbound
WestWorld: A Knowledge-Encoded Scalable Trajectory World Model for Diverse Robotic Systems D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 76d88731-d32f-4c2f-a51e-9ac0178b4463 · inbound
OPRIDE: Offline Preference-based Reinforcement Learning via In-Dataset Exploration D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1205d63e-d996-4c7f-90e1-76b626f235c8 · inbound
Genuine pair density wave order on the kagome lattice D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9712c528-5430-493c-bb61-6410c676f04c · inbound
ReinVBC: A Model-based Reinforcement Learning Approach to Vehicle Braking Controller D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e3d4369c-2edd-4654-8f6a-7a530f96f858 · inbound
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 61f24b5e-0e21-4723-ab70-4d3bf8bdecc8 · inbound
ScoRe-Flow: Complete Distributional Control via Score-Based Reinforcement Learning for Flow Matching D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 190e4f2f-e278-4ee9-8593-9eace7de7c56 · inbound
Fisher Decorator: Refining Flow Policy via a Local Transport Map D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fdba2290-af2c-49d6-8d2c-e84a4fcd6e54 · inbound
DAG-STL: A Hierarchical Framework for Zero-Shot Trajectory Planning under Signal Temporal Logic Specifications D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c66bfe18-9a1c-4107-a25b-10e7dd7c5f43 · inbound
Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 464a0396-f629-4e6c-b26f-4b39347c925f · inbound
Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a84a47fc-3a2c-4d00-9698-2d3450947a1d · inbound
When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 18cabee3-709c-4f69-85c7-566a74ba41d8 · inbound
A Reward-Free Viewpoint on Multi-Objective Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 66f36f7c-97aa-4bc2-9122-46775a80aba1 · inbound
SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10136481-a4dd-4b07-910e-e847d1dade26 · inbound
Borrowed Geometry: Cross-Distribution Head-Importance Fingerprints of Frozen Pretrained Gemma 4 31B D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ca971361-1454-46ed-a56f-7e766c9ecb40 · inbound
Borrowed Geometry: Cross-Distribution Head-Importance Fingerprints of Frozen Pretrained Gemma 4 31B D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3bd84c05-6355-435e-9499-8b697f1c501c · inbound
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7fdfa13b-b218-4768-b952-242a14399073 · inbound
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c43bb4b8-d90c-46e0-9218-f097d8f54369 · inbound
QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c17af99f-deba-4ed7-a985-9d972bd8a6ca · inbound
AdamO: A Collapse-Suppressed Optimizer for Offline RL D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cf50f462-3ba9-4e54-9f39-99c5fb671ce3 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 145
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7f2b5ac5-f29a-402d-86da-8695f7888a87 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 24b3ba85-c884-47fe-85b1-f56cffe7dd96 · inbound
Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dc57dcd3-07fe-47ae-ae2a-e3e1efdcdcca · inbound
When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d828322f-2d29-4f44-a7f2-160889b7446b · inbound
When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b6b75675-be3b-4006-a910-0a84b22b2387 · inbound
Beyond Autoregressive RTG: Conditioning via Injection Outside Sequential Modeling in Decision Transformer D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ff817eb7-3b28-4405-98e4-6c28d6b6c3af · inbound
Hitting Time Isomorphism for Multi-Stage Planning with Foundation Policies D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d9ecb821-828c-4490-951c-054cff5e55d9 · inbound
Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 126dba32-4761-4544-8df1-4cb6b8e91b78 · inbound
Path-Coupled Bellman Flows for Distributional Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 674b60ee-f6a5-4472-9be6-b3ea984461eb · inbound
Muninn: Your Trajectory Diffusion Model But Faster D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5168a44a-e61d-4195-ace1-3e30d6303173 · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f1b1d3b8-6b7a-49ac-9ea9-261dd9604361 · inbound
RankQ: Offline-to-Online Reinforcement Learning via Self-Supervised Action Ranking D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 38465ba8-f265-4ffc-a194-0da14a4ed6a9 · inbound
Discrete Flow Matching for Offline-to-Online Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4d16ba0-f6d5-4fdd-a772-588d649e45a1 · inbound
Aligning Flow Map Policies with Optimal Q-Guidance D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a5acfc70-d5e7-43e2-b95c-0f1863985a87 · inbound
Bridging Domain Gaps with Target-Aligned Generation for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f7760670-0df6-4bb9-8cf4-69235e2804c6 · inbound
Trajectory-Level Data Augmentation for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation def7e9d9-865b-4986-8179-fd7c5d15688b · inbound
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation adfd5b49-d753-43f4-864b-8cb788035b6f · inbound
Q-Flow: Stable and Expressive Reinforcement Learning with Flow-Based Policy D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 237d37e1-954c-4458-9241-6c3cc4ed5702 · inbound
Matrix-Space Reinforcement Learning for Reusing Local Transition Geometry D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d55df0e7-3ba9-46ad-80ea-eb57cb4a6040 · inbound
ROAD: Adaptive Data Mixing for Offline-to-Online Reinforcement Learning via Bi-Level Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8198ffb8-9122-401a-bff6-c39122b2a403 · inbound
Peng's Q($\lambda$) for Conservative Value Estimation in Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3733f407-d294-46b9-9a2f-ebebce9c9633 · inbound
Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 289
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 43488997-d8b8-465b-b061-0582c4de41c9 · inbound
ISEP: Implicit Support Expansion for Offline Reinforcement Learning via Stochastic Policy Optimization D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 93b0f1b0-0f9c-4e52-9bae-199ba0d54173 · inbound
Planner-Admissible Graph-PDE Value Extensions for Sparse Goal-Conditioned Planning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 84e4ed2a-33f8-405d-ae91-ead714c2f975 · inbound
Mechanisms of Misgeneralization in Physical Sequence Modeling D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13bbe72d-86e0-46db-b225-55663de43657 · inbound
stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0083e051-41b1-47bf-adc2-b36f6b49bcf1 · inbound
Target-Aligned Bellman Backup for Cross-domain Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ef85e69b-e9fa-42b6-933e-aa3bc3ad74af · inbound
Goal-Conditioned Agents that Learn Everything All at Once D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0483ddab-512c-4619-a015-6c90d733a3e4 · inbound
Nano World Models: A Minimalist Implementation of Future Video Prediction D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d993b474-5cde-4561-9112-c7cfa0d22ab6 · inbound
Neuro-Inspired Inverse Learning for Planning and Control D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d57e4617-eefe-4682-bab0-65e4c02a40ac · inbound
On the Stability and Realizability of Recurrent Polynomial Surrogate Ternary Logic Gate Networks D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f8673c23-e7cc-4c9e-a7e1-0eb89ca582e6 · inbound
Cross-Domain Energy-Guided Diffusion Generation for Off-Dynamics Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 620cc88a-f8c3-4a09-b8de-8368248bf548 · inbound
Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9bb7d4a3-abc3-43de-b545-3c4ccdbc78d5 · inbound
How to Mitigate the Distribution Shift Problem in Robotics Control: A Robust and Adaptive Approach Based on Offline to Online Imitation Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13412e0f-17b7-4ab5-9482-8704fa492281 · inbound
Back to Parsimonious Latents: Learning Task-Centric World Models from Visual Foundations D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 79c8ee20-236e-4919-a362-eb5da8c599a2 · inbound
Latent Representation Alignment for Offline Goal-Conditioned Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2778117a-0dbd-4f99-87f7-bd8ec7e2894c · inbound
Aligning Few-Step Generative Models by Amortizing Sample-based Variational Inference D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 748f8126-2dba-4f68-b95e-629533aaf2f5 · inbound
SPAR: Support-Preserving Action Rectification D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dda4bdf0-2a28-435e-b842-2f9e7fc58c6a · inbound
Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.