Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:05:34.935087Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 1 inbound Pith citation observation for arXiv:2506.14058.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:05:34.935087Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T20:52:08.062450Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T20:54:21.647943Z
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2deaca87-2bff-41fd-8f7e-860ed42a450d · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93bc56a0-3331-482d-9605-ce882d4f0f45 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning A survey on offline reinforcement learning: Taxonomy, re- view, and open problems,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 604314b4-3e27-4d90-8ca3-6f19d7d0d8a0 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Beyond uniform sampling: Offline rein- forcement learning with imbalanced datasets,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 236357e8-6ef3-4a9a-977a-73bbac6958dd · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning The Importance of Pessimism in Fixed-Dataset Policy Optimization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8e25237-783c-40f6-ba56-904d2be06f7b · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning A minimalist approach to offline reinforcement learning,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e73a3f8a-57a8-4dd5-9eab-d581031347f3 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Safe Offline Reinforcement Learning with Feasibility-Guided Diffusion Model
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b35471d9-ad49-46ca-9c13-3370f8286c4e · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Spectral normalization for lipschitz-constrained poli- cies on learning humanoid locomotion,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation eec35a7d-9d51-4cb5-bf0b-4ba930e96ecd · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Off-policy deep reinforcement learning without exploration,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c7862269-d9a3-4488-95fd-24c8c5f651ca · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Conser- vative q-learning for offline reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f5c72a33-f7f0-4801-b359-a790c5ca075e · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified Q-ensemble,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 204d28c3-6802-4fa4-b287-10233ce27d3a · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be6ea27-521a-474d-bce6-e720a79165cd · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning L2c2: Locally lipschitz continuous con- straint towards stable and smooth reinforcement learn- ing,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 965f766b-652e-41a8-aa8a-b01fac3cd49c · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Monotonic value function factorisation for deep multi-agent reinforcement learn- ing,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 9028cea8-d25d-45e6-aefa-462185e76c73 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Optnet: Differentiable opti- mization as a layer in neural networks,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c596da04-9ee7-4e72-9769-113a3ebad4c7 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning A theory of regularized Markov decision processes,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 29ba232e-c4c5-4f17-979a-8bad596bdc84 · outbound
Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afe72333-69ab-4fb1-9db2-b6f349c12b10 · inbound
Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets Implicit Constraint-Aware Off-Policy Correction for Offline Reinforcement Learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.