Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:38:35.725134Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2507.02356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:38:35.725134Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1da2545c-de84-4043-9793-2d18e1ffe030 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Uncertainty-based offline reinforcement learning with diversified q-ensemble
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3e5d8264-1aa1-4027-9faf-c087bdbf7e9a · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Layer Normalization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dede0fb2-7726-4036-bc49-c3a9da87a142 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Offline Reinforcement Learning via High-Fidelity Generative Behavior Modeling
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7d77ab3-7008-4262-a7e8-eb0341d5621b · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Score Regularized Policy Optimization through Diffusion Behavior
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30bc8fe0-d178-4f77-ac42-81ce3fe4e4d5 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Diffusion Policies creating a Trust Region for Offline Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 270e3281-3349-405a-be63-20c824aff9ab · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Heavy-tailed denoising score matching
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cfd8ec7d-c3f3-4b95-a8e1-9eed79c10ea8 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2458abd8-0a18-49dc-8f1e-74484648aefc · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection A minimalist approach to offline reinforcement learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe4cf0f3-adb7-45c5-9de0-b0f5b52e3330 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Addressing function approximation error in actor-critic methods
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8f0a35-73b2-4bb6-91b2-d81f6d6e16cc · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Off-policy deep reinforcement learning without exploration
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bdfee5c-e361-4565-82e3-156ebaf484f5 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Calculus of variations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44a5caa0-2ef4-435c-bf9d-d4675b8f5884 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd83fcf-0547-4a0b-9dc1-ecabbe7bc4a1 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Estimation of non-normalized statistical models by score matching
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23d6c552-6feb-46af-97a5-0fbb323cb07f · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Understanding diffusion objectives as the elbo with simple data augmentation
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0601121a-7f5e-46fb-ab8b-600e70a417b3 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Adam: A Method for Stochastic Optimization
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dcf4a96-1f10-4f7d-8c6b-853a671b92f4 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Offline Reinforcement Learning with Implicit Q-Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e7a1eb-a7a6-4c39-b6e4-c53cf57c1574 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Stabilizing off-policy q-learning via bootstrapping error reduction
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcf8ceda-bcaa-428f-bb39-bc5937405986 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Conservative q-learning for offline reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 41b9557a-32c3-4a56-ba75-368c9b45e136 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Reinforcement learning with augmented data
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 55ec212d-70ea-418b-a4dd-af6133addaaa · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Batch reinforcement learning with hyperparameter gradients
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ec0ef95b-8d50-4a57-abaf-d3cabbd5cc84 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Learning energy-based models in high-dimensional spaces with multiscale denoising-score matching
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d3ce1b0b-1a87-4a90-96ad-568a994fe870 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Contrastive energy prediction for exact energy-guided diffusion sampling in offline reinforcement learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51d92faf-a390-4ba8-9e6e-a0c6d0d6d8af · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baf425d1-58da-4a6f-bb21-d574dbcd88a6 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Anti-exploration by random network distillation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b65795e4-72c4-4282-bfeb-70a1979aaa5b · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Heavy-Tailed Diffusion Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78e68a9c-3574-42a0-8af4-e66fd739dad3 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection DreamFusion: Text-to-3D using 2D Diffusion
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c922708f-0235-4672-9a6f-680c9919918d · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Efficient differentiable simulation of articulated bodies
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 995cbc0a-a10c-42ba-8e3e-cbb70fa0c5ac · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Offline reinforcement learning as anti-exploration
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation de5bffb6-547d-4f2c-a9f5-e3254fbdfef1 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection S4rl: Surprisingly simple self-supervision for offline reinforcement learning in robotics
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb38a4bf-2c58-4375-80ea-5a4aa65b7b1e · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Generative modeling by estimating gradients of the data distribution
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a2cce95-1e29-4969-ab39-8b4baee2e0c2 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Score-Based Generative Modeling through Stochastic Differential Equations
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af1d8a34-686b-4d5f-8006-e3459a95c306 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Revisiting the minimalist approach to offline reinforcement learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78e6b11b-7441-4184-b72d-1864ae3d9c6f · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection A connection between score matching and denoising autoencoders
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95dc4f19-86d8-4674-8e21-89a853191561 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Diffusion Policies as an Expressive Policy Class for Offline Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cfadee7-393b-42d9-8e70-a045f47d28c7 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection On scale mixtures of normal distributions
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 993d3e30-6b2c-4ac3-b75d-ea23486c921c · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Exploration and Anti-Exploration with Distributional Random Network Distillation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bea9b73e-a826-4506-ab75-db014269a869 · outbound
Offline Reinforcement Learning with Penalized Action Noise Injection Rorl: Robust offline reinforcement learning via conservative smoothing
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.