Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:46:08.077256Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.00503.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:46:08.077256Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
34 of 34 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9ef48b7a-14cd-479f-ab2e-d0d3145028b2 · outbound
Variational OOD State Correction for Offline Reinforcement Learning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99096310-b63b-46a7-a373-48b63db62b94 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Learning markov state abstractions for deep reinforcement learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 31367436-eaa5-4c53-b449-1d95c54811b9 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified q-ensemble
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation dd567200-1253-42c1-9a0c-313ef739d610 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bc4da37b-efec-4f24-8847-de23489ccacf · outbound
Variational OOD State Correction for Offline Reinforcement Learning Label-noise robust logistic regression and its applications
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 644e8d65-2e9b-424a-9f37-13d8059d9c18 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Importance Weighted Autoencoders
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c74fccd9-e3bd-4b09-a8ff-dffe82f11d8f · outbound
Variational OOD State Correction for Offline Reinforcement Learning Tutorial on Variational Autoencoders
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77265d7d-19f6-4a15-a776-b59715b6960a · outbound
Variational OOD State Correction for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e17f339a-c7d0-449b-8255-1e612dc0fda8 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Off-policy deep reinforcement learning without exploration
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation aee7d1f4-b677-4eb2-892e-53ce6e4efd57 · outbound
Variational OOD State Correction for Offline Reinforcement Learning A comprehensive survey on safe reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05514fed-c6af-47bf-bfc3-e0f6155478a2 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Estimation of non-normalized statistical models by score matching
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc48874b-b5b7-4575-aec1-3c9f4866fadb · outbound
Variational OOD State Correction for Offline Reinforcement Learning Planning with Diffusion for Flexible Behavior Synthesis
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3e7bd93-c663-4fc8-8a28-9a4b8ed1cabb · outbound
Variational OOD State Correction for Offline Reinforcement Learning Recovering from out-of-sample states via inverse dynamics in offline reinforcement learning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7a0d9678-f707-4c44-ba55-1a10b9a9fca4 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cc2807a4-8eb1-4fe9-9019-3ca614338c95 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Scalable deep reinforcement learning for vision-based robotic manipulation
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8fd905cf-c544-4936-8ada-8068da5e5e3c · outbound
Variational OOD State Correction for Offline Reinforcement Learning Lyapunov density models: Constraining distribution shift in learning-based control
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 23404f0f-5c81-45da-8c2f-ba52a23a7863 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Kingma and Max Welling
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85b7faeb-8dbe-47b5-8987-fdfc68a02a9c · outbound
Variational OOD State Correction for Offline Reinforcement Learning Conservative q-learning for offline reinforcement learning
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 397c65bb-9786-48bd-899d-038cbc8dda7c · outbound
Variational OOD State Correction for Offline Reinforcement Learning Batch reinforcement learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9ada6e4-f5ef-48db-9be9-6964eee85dc7 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Supported value regularization for offline reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0aae39b5-9e6a-454d-94cf-706f7adb7df4 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2e559b6e-9885-4097-9294-ffad781c7a30 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Human-level control through deep reinforcement learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092274a0-d39b-42ab-a467-a2bc360b27fc · outbound
Variational OOD State Correction for Offline Reinforcement Learning Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e98c70ba-ad82-4ed3-b7b0-01ac023280e3 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Mastering the game of go without human knowledge
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34534397-3fd8-466f-b144-d77d96c76e07 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1d9d1b-7911-43d5-bc75-7ddcca1d3607 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Robust distance metric learning in the presence of label noise
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 83c57b72-b982-4f67-bfa2-35f8d884eb2f · outbound
Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 43b9d0b8-28e4-46c7-aa6e-fba8f1034228 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13475ed8-4071-40a5-bbb8-dbfd4fe2687d · outbound
Variational OOD State Correction for Offline Reinforcement Learning Supported policy optimization for offline reinforcement learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1999d70c-3f12-4a5c-b282-7d01cecce048 · outbound
Variational OOD State Correction for Offline Reinforcement Learning RORL: Robust Offline Reinforcement Learning via Conservative Smoothing
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a879cd6a-38f7-4ccf-a223-a454c0989d89 · outbound
Variational OOD State Correction for Offline Reinforcement Learning An implicit trust region approach to behavior regularized offline reinforcement learning
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4ea89976-a352-4dfa-b40c-aba3d976a6fe · outbound
Variational OOD State Correction for Offline Reinforcement Learning State deviation correction for offline reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9840f7b5-c279-4ecf-aa8d-79d284f69b90 · outbound
Variational OOD State Correction for Offline Reinforcement Learning Constrained policy optimization with explicit behavior density for offline reinforcement learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 85c0bc90-d052-4f8d-b25f-85ed27f8bca7 · outbound
Variational OOD State Correction for Offline Reinforcement Learning write newline
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.