Pith. sign in

Paper Citation Record · LEDGER

Variational OOD State Correction for Offline Reinforcement Learning

As of 17 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2505.00503.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.00503 v3

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:46:08.077256Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9ef48b7a-14cd-479f-ab2e-d0d3145028b2 · outbound

This paper cites GPT-4 Technical Report.

Variational OOD State Correction for Offline Reinforcement Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.965220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.965220Z digest=sha256:1a5f4f63a45771197a08515526fd6f6e69eb1a426823006b65b083b5607a4b15

Observation 99096310-b63b-46a7-a373-48b63db62b94 · outbound

This paper cites Learning markov state abstractions for deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Learning markov state abstractions for deep reinforcement learning

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.400482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.969768Z digest=sha256:464d303c84c007c3661620181930b1276f8ae9e4f356074a442b44a244fb6304

Observation 31367436-eaa5-4c53-b449-1d95c54811b9 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Variational OOD State Correction for Offline Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.391492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.973100Z digest=sha256:3c96e5f9d74679bf6de20aac4c6df4b6b73b68f1a59d180fdd752caa9fc42da7

Observation dd567200-1253-42c1-9a0c-313ef739d610 · outbound

This paper cites Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Pessimistic bootstrapping for uncertainty-driven offline reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.382105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.976836Z digest=sha256:db2ec2ce427f409fbd67ca2d0b05c074536628f8c0491cf364b73fdbf35e391c

Observation bc4da37b-efec-4f24-8847-de23489ccacf · outbound

This paper cites Label-noise robust logistic regression and its applications.

Variational OOD State Correction for Offline Reinforcement Learning Label-noise robust logistic regression and its applications

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.372433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.980276Z digest=sha256:eda83ccee5c859aa14fb5731eea956601f9a89e362074955a431e15b0c8ba775

Observation 644e8d65-2e9b-424a-9f37-13d8059d9c18 · outbound

This paper cites Importance Weighted Autoencoders.

Variational OOD State Correction for Offline Reinforcement Learning Importance Weighted Autoencoders

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.983646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.983646Z digest=sha256:d90b8fd2ce16bf56832324ee9aa65821fcb303b32dec06eda2392e4d59f1615e

Observation c74fccd9-e3bd-4b09-a8ff-dffe82f11d8f · outbound

This paper cites Tutorial on Variational Autoencoders.

Variational OOD State Correction for Offline Reinforcement Learning Tutorial on Variational Autoencoders

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.987957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.987957Z digest=sha256:6e407e26b49bef45b5b314716f20bcc5f770296043e414d46832401e3144c8a1

Observation 77265d7d-19f6-4a15-a776-b59715b6960a · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Variational OOD State Correction for Offline Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.991451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.991451Z digest=sha256:1a5927a85697176f01d1835c1d83f1e414455917ffd64a2a845087777ebb72aa

Observation e17f339a-c7d0-449b-8255-1e612dc0fda8 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Variational OOD State Correction for Offline Reinforcement Learning Off-policy deep reinforcement learning without exploration

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.362697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:07.994764Z digest=sha256:02267c87a19bb578b33ce9252aa994167b0e44b62f36f668c30f1f25d5a149f7

Observation aee7d1f4-b677-4eb2-892e-53ce6e4efd57 · outbound

This paper cites A comprehensive survey on safe reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning A comprehensive survey on safe reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:07.998083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:07.998083Z digest=sha256:ac1350c5e3952fe3d24e856b3dcb99413742bae82a52491fc03160671b119562

Observation 05514fed-c6af-47bf-bfc3-e0f6155478a2 · outbound

This paper cites Estimation of non-normalized statistical models by score matching.

Variational OOD State Correction for Offline Reinforcement Learning Estimation of non-normalized statistical models by score matching

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.001245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.001245Z digest=sha256:bcb5b00af369cb9acc0d3b01a40a75d5da935c4cd8cd04d70b7dbbabe719cca3

Observation fc48874b-b5b7-4575-aec1-3c9f4866fadb · outbound

This paper cites Planning with Diffusion for Flexible Behavior Synthesis.

Variational OOD State Correction for Offline Reinforcement Learning Planning with Diffusion for Flexible Behavior Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.008559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.008559Z digest=sha256:b4ab1d0769e2ab4b2f35abbec2acdb4cbae0829398d02f80370e416f2419f1f7

Observation d3e7bd93-c663-4fc8-8a28-9a4b8ed1cabb · outbound

This paper cites Recovering from out-of-sample states via inverse dynamics in offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Recovering from out-of-sample states via inverse dynamics in offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.342922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.012362Z digest=sha256:95b2099a18cb937e889b1b149af8e32f776dc311c0656da7aca4480646f26493

Observation 7a0d9678-f707-4c44-ba55-1a10b9a9fca4 · outbound

This paper cites an unresolved cited work.

Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:46:08.332759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.015471Z digest=sha256:bbfa5d91aed3112d7f63125c14bf6f46f0c47e2b6294524498f07dd5435a9aec

Observation cc2807a4-8eb1-4fe9-9019-3ca614338c95 · outbound

This paper cites Scalable deep reinforcement learning for vision-based robotic manipulation.

Variational OOD State Correction for Offline Reinforcement Learning Scalable deep reinforcement learning for vision-based robotic manipulation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.323263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.018719Z digest=sha256:1af28cea3fa7593bd8a60fc562064f2c35a0e9b67f7dfa37db779044007656b0

Observation 8fd905cf-c544-4936-8ada-8068da5e5e3c · outbound

This paper cites Lyapunov density models: Constraining distribution shift in learning-based control.

Variational OOD State Correction for Offline Reinforcement Learning Lyapunov density models: Constraining distribution shift in learning-based control

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.313028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.021758Z digest=sha256:e75a147771d32cce7f6473da42cea689d400bc5a0368338ee104910d34f64ab7

Observation 23404f0f-5c81-45da-8c2f-ba52a23a7863 · outbound

This paper cites Kingma and Max Welling.

Variational OOD State Correction for Offline Reinforcement Learning Kingma and Max Welling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.024790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.024790Z digest=sha256:e214e3e1a51d7409475032dd211efaee7ecaaeb15eb1909475a86f37aece44d5

Observation 85b7faeb-8dbe-47b5-8987-fdfc68a02a9c · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Conservative q-learning for offline reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.297346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.027953Z digest=sha256:81884988745ea78335ae173d714077d890ec5165d56c6f7efc8b607f5420c2b9

Observation 397c65bb-9786-48bd-899d-038cbc8dda7c · outbound

This paper cites Batch reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Batch reinforcement learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.030931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.030931Z digest=sha256:69f4b55db5cca166ae39e3db5aa168bdbc7b1104d9e1173644d088a013330b83

Observation c9ada6e4-f5ef-48db-9be9-6964eee85dc7 · outbound

This paper cites Supported value regularization for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Supported value regularization for offline reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.280727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.034368Z digest=sha256:ee20666fa13db25964b365ed7e14c7a1663a372e10affe358cf9dd6c26a75b76

Observation 0aae39b5-9e6a-454d-94cf-706f7adb7df4 · outbound

This paper cites Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression.

Variational OOD State Correction for Offline Reinforcement Learning Offline Reinforcement Learning with OOD State Correction and OOD Action Suppression

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:46:08.144011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.037469Z digest=sha256:3f6d885638040dad6bb044e1cdda4e249acbcab445488526b291c48e6eaa24cc

Observation 2e559b6e-9885-4097-9294-ffad781c7a30 · outbound

This paper cites Human-level control through deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Human-level control through deep reinforcement learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.040956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.040956Z digest=sha256:e98ca37baf0177a29f99b78e0d0b526c46414fcdda3de0d21a9704732ba0379e

Observation 092274a0-d39b-42ab-a467-a2bc360b27fc · outbound

This paper cites Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Deeploco: Dynamic locomotion skills using hierarchical deep reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.265226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.043948Z digest=sha256:59b27ba4d3b626431ad833400725b93286eefc894a4dfd3211b9b6bce4fdeefb

Observation e98c70ba-ad82-4ed3-b7b0-01ac023280e3 · outbound

This paper cites Mastering the game of go without human knowledge.

Variational OOD State Correction for Offline Reinforcement Learning Mastering the game of go without human knowledge

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.046845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.046845Z digest=sha256:6db49ffc4ecf16d14d02294d8a126d4ad4d97663903ea11d93697ddaa6f2dcd0

Observation 34534397-3fd8-466f-b144-d77d96c76e07 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Variational OOD State Correction for Offline Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.049612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.049612Z digest=sha256:a395fc3191069a18257a39cf29728bf5b6df92bf1c6c36c26ac37c6f9bef2ff1

Observation 0d1d9d1b-7911-43d5-bc75-7ddcca1d3607 · outbound

This paper cites Robust distance metric learning in the presence of label noise.

Variational OOD State Correction for Offline Reinforcement Learning Robust distance metric learning in the presence of label noise

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.249511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.052689Z digest=sha256:878e221a3969cc118961024a0f8c16a2c834006bc5cf9ffc9e0181ac6cb40ad7

Observation 83c57b72-b982-4f67-bfa2-35f8d884eb2f · outbound

This paper cites an unresolved cited work.

Variational OOD State Correction for Offline Reinforcement Learning Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-16T04:46:08.238963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.055612Z digest=sha256:0b6bdde8fd00edde9416e5678a7a148f0634c3add3b78e3dc18427660f791808

Observation 43b9d0b8-28e4-46c7-aa6e-fba8f1034228 · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Variational OOD State Correction for Offline Reinforcement Learning Behavior Regularized Offline Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.058685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.058685Z digest=sha256:dcad0ef5acadc53ca95d86b4e1980baef72377edcf16c504928a3de03aaf2122

Observation 13475ed8-4071-40a5-bbb8-dbfd4fe2687d · outbound

This paper cites Supported policy optimization for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Supported policy optimization for offline reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.229077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.061838Z digest=sha256:e6028a4e2e2eb53d08a4d414b689b79c9715e69a1ffbf07cf3d39b37299d7372

Observation 1999d70c-3f12-4a5c-b282-7d01cecce048 · outbound

This paper cites RORL: Robust Offline Reinforcement Learning via Conservative Smoothing.

Variational OOD State Correction for Offline Reinforcement Learning RORL: Robust Offline Reinforcement Learning via Conservative Smoothing

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-16T04:46:08.113098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.064871Z digest=sha256:d5bba4a5efe2133b5b11d0b5d9abf79d28615b6101463085e09307aa0abd5ed1

Observation a879cd6a-38f7-4ccf-a223-a454c0989d89 · outbound

This paper cites An implicit trust region approach to behavior regularized offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning An implicit trust region approach to behavior regularized offline reinforcement learning

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.218677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.068085Z digest=sha256:3f2f21617f66c163fb36c5643da09bbcdf360923320569bbb26bd8b72ef74793

Observation 4ea89976-a352-4dfa-b40c-aba3d976a6fe · outbound

This paper cites State deviation correction for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning State deviation correction for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.208816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.071301Z digest=sha256:f78e9907f4ffa83286e2e2d84f0d2620d1a8f5afd16e4cb5f47e845cabd27a8a

Observation 9840f7b5-c279-4ecf-aa8d-79d284f69b90 · outbound

This paper cites Constrained policy optimization with explicit behavior density for offline reinforcement learning.

Variational OOD State Correction for Offline Reinforcement Learning Constrained policy optimization with explicit behavior density for offline reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T04:46:08.199344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-16T04:46:08.074192Z digest=sha256:bde0fb404e7c999d929846f5a6635e8d008a0875e7838eadb4d330d44005ed11

Observation 85c0bc90-d052-4f8d-b25f-85ed27f8bca7 · outbound

This paper cites write newline.

Variational OOD State Correction for Offline Reinforcement Learning write newline

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-16T04:46:08.077256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:46:08.077256Z digest=sha256:0eae7528065ac7e34328e035bc8228a50e2e239c34a179b1d6817bbf5835e22d

Pith citing papers

No inbound Pith citation observations are available.