Pith. sign in

Paper Citation Record · LEDGER

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood

As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.08417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08417 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:22:27.653919Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact5
  • verified fuzzy28
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2727ca4-f5f9-430c-bae1-d53ed89ade7b · outbound

This paper cites write newline.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.652759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.652759Z digest=sha256:e879946262bf0723e7b72302cdd151f327c95cfaabd403409bbbb217851100da

Observation 04f87bd3-0ea1-41e8-bf93-3fa6826f2fb7 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.687083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.687083Z digest=sha256:f518e9acc2b7934cec120799f9cb6baa5c5c6f346e94790c4c0fa0e7bfcf5657

Observation b62d4df3-f0fa-4a6b-9e96-2a1417d4dbb4 · outbound

This paper cites Near-optimal regret bounds for reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Near-optimal regret bounds for reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.626467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.746503Z digest=sha256:6870fa0d0926d18e65bc12ac4f112bacde0e3e35d41fc63f9e2e91a82cb3ea14

Observation 0012dfda-f904-4179-9f74-4956d0a95440 · outbound

This paper cites Manifold topology divergence: a framework for comparing data manifolds.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Manifold topology divergence: a framework for comparing data manifolds

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.611573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.776646Z digest=sha256:423debf0f7d361c98ec7137a75d3a7b152b195beea9307fef27fa18bd8e7b631

Observation 77245729-e20d-448a-8673-a1ab89bc83d5 · outbound

This paper cites Laplacian eigenmaps and spectral techniques for embedding and clustering.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Laplacian eigenmaps and spectral techniques for embedding and clustering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.848799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.848799Z digest=sha256:b518fc456d5a2318c5dbc2331bdc58dac96ba979f9957ee69fe50715b19c9120

Observation ad686d95-9820-434d-8f85-954a04ad2cb1 · outbound

This paper cites On the Inductive Bias of Neural Tangent Kernels.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the Inductive Bias of Neural Tangent Kernels

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.919649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.919649Z digest=sha256:efb5d59fbd35d35cf1c133eb786b744f880bad9865aeb227d9cdc846940caa99

Observation 77b6dbda-ac72-4049-bbe7-f8caf6c48683 · outbound

This paper cites Flows for simultaneous manifold learning and density estimation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Flows for simultaneous manifold learning and density estimation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.586125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.989185Z digest=sha256:39dd2869a96ca6eba148a2daf8afdebb075beb587344ae6c190eadbf9489031a

Observation 36370ff4-c1f4-4016-bb26-f32426d32543 · outbound

This paper cites OpenAI Gym.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood OpenAI Gym

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.040414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.040414Z digest=sha256:290c5be182ef44c468547f18ad3c96b876ab846b7c955612f5511429edcc56a0

Observation edfe8984-2ac4-41d4-864d-aa29aaf70d34 · outbound

This paper cites Bail: Best-action imitation learning for batch deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Bail: Best-action imitation learning for batch deep reinforcement learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.097958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.097958Z digest=sha256:c3acf359a93352d008490c6346972352e7c8cd012088e396796a88827c943d3e

Observation c13f2ac5-941f-4bbc-ae58-d7ee4c981374 · outbound

This paper cites Diffusion maps.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Diffusion maps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.561407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.181646Z digest=sha256:88a97da8443f763d2d9efeebce468b3fcd1efc86373b38ae16246b7e4c61a38e

Observation 46b1f972-6798-496f-a2c9-70fe477b9e2a · outbound

This paper cites Pink noise is all you need: Colored noise exploration in deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Pink noise is all you need: Colored noise exploration in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.547034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.256430Z digest=sha256:25f3151172d4ea2217a36045172d2275a71bfc1a6c51d06718a0d64e6abe80a1

Observation e57d3fa9-29c9-4044-89e8-69d03c6d1aef · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2021.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood D4rl: Datasets for deep data-driven reinforcement learning, 2021

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.311145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.311145Z digest=sha256:9c1e16bc0e477741e749c9e5a1b8dd140e22523f863e91f316e4f191e6b2ea34

Observation 7ae79ce4-115e-4677-9086-b1f6e9c8af6d · outbound

This paper cites A minimalist approach to offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A minimalist approach to offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.522372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.363317Z digest=sha256:5b34a4e59512fc243dacf8f6b175eb62ecae0cb0b88bb6784c35dd49376815c2

Observation 423edc72-3bbc-4e63-a4cb-c285062e1a18 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Off-policy deep reinforcement learning without exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.507852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.417323Z digest=sha256:291bdc0f8f06491810a8ec4d4b871d217c92c981fd9aa0220fd3b1dc030a9b11

Observation 01b87718-6770-47eb-be7a-7a426b0c7700 · outbound

This paper cites Learning rankings via convex hull separation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning rankings via convex hull separation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.493119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.490677Z digest=sha256:4a3588ae68daf4196fee06fb9f03a98f4c80c6602353f50be24af4483cf3edac

Observation 14dd046c-95d9-4c78-9362-9a3b7e400315 · outbound

This paper cites Extreme Q-Learning: MaxEnt RL without Entropy.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Extreme Q-Learning: MaxEnt RL without Entropy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.560434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.560434Z digest=sha256:094dee5e9061c69d58aacfda56ddd5e0ffae830ce30676d6969b2e2c426a2016

Observation 0762145d-4da6-4aea-8d80-90ebb6106a89 · outbound

This paper cites Improving Offline RL by Blending Heuristics.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving Offline RL by Blending Heuristics

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:28.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.609399Z digest=sha256:03aa9d6950508c6075b91217d18224b5b30a7d3ba7a2d5f52e56d4254b67e62b

Observation 5c018bd8-872b-4025-8e05-9254dd4f0ef2 · outbound

This paper cites Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.478149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.668798Z digest=sha256:125f9938cad17a1ab5361a71961a50c6de344f8d3122dcb3c16883addb218b8a

Observation 4e505ac3-d422-4752-b19e-2ae4a178802b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.748337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.748337Z digest=sha256:8eff622ce5bc629177ec81eabd97aa678b37d537dd93b840e5c6cc0d11bc2c7f

Observation 087f99af-5649-49de-bd47-2c2e0f950ff8 · outbound

This paper cites Random projections for manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Random projections for manifold learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.451972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.810006Z digest=sha256:644875e511312757af0dbde2145abae03397b384868efeb4996ef75f28cfed70

Observation f174c80b-b129-4129-8507-7b5ac451a2dd · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.436324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.868522Z digest=sha256:d7ecb341f98030819b1cdcc03264308c40b62ee6e69626a2d1ff1c874b694a65

Observation 7483e25c-d457-4a7d-85ad-9f0c26e589c7 · outbound

This paper cites Mild policy evaluation for offline actor--critic.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mild policy evaluation for offline actor--critic

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.421916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.941158Z digest=sha256:191ddb8202678d51f93cfcc0c8e199a94897c29c7772b1305054d165e1bb7f47

Observation bb277ef5-c440-4243-8801-ae2ca4e47a57 · outbound

This paper cites Neural Tangent Kernel: Convergence and Generalization in Neural Networks.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Neural Tangent Kernel: Convergence and Generalization in Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.087828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.087828Z digest=sha256:20ff568fdfd27527fb372f4447e166512229a75975cb7e9a1dbcbb1eb9f0b45a

Observation 0971bad8-7bec-440e-91f0-f6a9c4586f28 · outbound

This paper cites A convex hull-based data selection method for data driven models.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A convex hull-based data selection method for data driven models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.407761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.188248Z digest=sha256:9d8debb68b1ddfb65861ed75b4637d6a8dd94d9cba6ff60197335a77e05d0653

Observation b435e2c8-7c51-4615-a19b-45bfc93b67ce · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.300203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.300203Z digest=sha256:5a8b599be3470a3ad058aa98f6b7b68b1ea880ad385f4d3a8ce6a7732cf515d1

Observation b4ac79f4-d8d5-4a13-8266-8afc8aabe59c · outbound

This paper cites Offline reinforcement learning with fisher divergence critic regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline reinforcement learning with fisher divergence critic regularization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.392898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.361489Z digest=sha256:614beb54d600b77ac74015e94731313eca8c123635715764afcad151ba647c05

Observation 2796467c-3a3c-4708-a2c5-91d74746154e · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline Reinforcement Learning with Implicit Q-Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.460008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.460008Z digest=sha256:8372743802c4e563926c8fde19f55a370a58b978fb6a385a6db762afc408a308

Observation 950a4f7b-b3a5-4673-bf86-e4a380591bf4 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.377984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.549357Z digest=sha256:f8dfb9d3ca7bb3d72618c7bde03cddca440087393b853f9f1f38a5ed80500b19

Observation 7acef59d-6aa9-4981-9cec-af33d8400982 · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Conservative q-learning for offline reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.669895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.669895Z digest=sha256:81f41fba551a8ecb5734fcd84ac404456a354251c84c594918270f45130663de

Observation 8b551c13-6061-4f88-90c0-b6a1620cb95d · outbound

This paper cites Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.992371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.764096Z digest=sha256:3cb9167f14fb9582fcfde12e32fd346b8357c6f5e8eb4e19f780fd719f4a04d9

Observation 7a30ff04-8ad2-4ad5-a2df-3330feda9fd5 · outbound

This paper cites When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.971229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.844016Z digest=sha256:267f53ce0609aa1bf89903219bb0d9bbe58375374c16b4f59dd59d6b5eefe67b

Observation 258495b4-b0ff-4cb4-99e9-f73d6b6800dc · outbound

This paper cites Mildly conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mildly conservative q-learning for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.352784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.916094Z digest=sha256:88670bd29f362ecfd816ed140c1818ed4e86e8981ee0a9c14616e4b74a8555b1

Observation b067009a-5214-4801-85fc-419e7b097027 · outbound

This paper cites SEABO: A Simple Search-Based Method for Offline Imitation Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood SEABO: A Simple Search-Based Method for Offline Imitation Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.045813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.045813Z digest=sha256:ad7835391ac8a3dfcc3e9af635b770965e7cadeadf8e2eaab97481be3a31f11a

Observation 1196c970-65f4-4ac2-8829-4bdcd0bce33a · outbound

This paper cites On the role of general function approximation in offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the role of general function approximation in offline reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.337075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.115252Z digest=sha256:090034520d670e7ba6d5b2dff87c46661e86acfcbaf975cfa8c3212404e14378

Observation 3d6c7838-3c2c-4f35-874e-6992398433af · outbound

This paper cites Machine learning algorithm based on convex hull analysis.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Machine learning algorithm based on convex hull analysis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.322758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.203111Z digest=sha256:211cff6329d5b3aa3b8604c1d1c2a38e1d2041b0fa6fdbfd6da6f3ca39e0fc8e

Observation 44f94f89-08f2-41ca-be92-788441ea4c08 · outbound

This paper cites Why is Posterior Sampling Better than Optimism for Reinforcement Learning?.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why is Posterior Sampling Better than Optimism for Reinforcement Learning?

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.932706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.273819Z digest=sha256:627edd08fc2f68d94c97404a475c26f7f60850e6b3d5dc0b2451a1b118e73ffc

Observation e3c0c889-bdd8-436e-ba64-193c5f90571d · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.320316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.320316Z digest=sha256:5155c20a20b30a261abb8edb2cda2398345ee72c78e66d82fda48d8f1d2730e4

Observation 4fbc39b4-afb8-4167-9a6f-cd1fdce17504 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.393335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.393335Z digest=sha256:7aeb45524da811eb514be5ac7ba3a82f9a838c7e38af8667e2aa5610dd79af96

Observation 87e71b41-e966-446e-9f88-910d9bf76560 · outbound

This paper cites Policy regularization with dataset constraint for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Policy regularization with dataset constraint for offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.308226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.477113Z digest=sha256:95281ccb2192ed5c9fbc25aeef59878e493930da6b52f0b7abd3b2b07dc84148

Observation 1e868df6-a8bb-44bf-9b57-e2a19a62e438 · outbound

This paper cites Nonlinear dimensionality reduction by locally linear embedding.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Nonlinear dimensionality reduction by locally linear embedding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.293800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.530029Z digest=sha256:8d8aaf21e2165185de5ddcc615be7cb0b6e15c23fa4ded0efb3e26ea29752346

Observation e2562a35-92bd-4490-8447-166d9e165625 · outbound

This paper cites A dataset perspective on offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A dataset perspective on offline reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.279695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.572723Z digest=sha256:b0735faa279519dfc98fea5e5dcb6e53cb18a355b95fd52553e6b07794b5ce80

Observation 5c6daea8-bc9b-4774-9770-09b969caaafe · outbound

This paper cites Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.632916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.632916Z digest=sha256:742253b3211204ad2ee46d1d73f4db175d672ad9e282442cbad3739a9487604c

Observation be1e0c20-24d3-4c46-9125-6725309df06d · outbound

This paper cites Sutton and Andrew G.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Sutton and Andrew G

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.731486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.731486Z digest=sha256:b9ff5bd0c391302535398f09cd581ce6c7feb88cdadd474cf9b7863ebe6b5e8a

Observation bdf784a3-85fc-4a18-a1b8-10cce6db4427 · outbound

This paper cites A global geometric framework for nonlinear dimensionality reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A global geometric framework for nonlinear dimensionality reduction

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.254341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.777449Z digest=sha256:30261bb3d1dfaf6b84e926d92245e0de8d705a15a995ee664b15469f91583355

Observation c3550023-30a5-4f75-860e-b65b43c43a7e · outbound

This paper cites Mujoco: A physics engine for model-based control.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mujoco: A physics engine for model-based control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.818469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.818469Z digest=sha256:34d76b1c9a900816a9f343d6ec5c5c65c0ed3af88dd01a0245675d5533a29771

Observation 97666142-1f76-4ba5-8747-d477965e14a5 · outbound

This paper cites Adaptive manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adaptive manifold learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.237822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.909702Z digest=sha256:1898563fc95163c534445285c66b7abf7b59cd36d358c9b37bf1371b449933a3

Observation c32c712c-f7b1-4f43-826c-ec7805ae0790 · outbound

This paper cites Improving generalization in reinforcement learning with mixture regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving generalization in reinforcement learning with mixture regularization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.223654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.975185Z digest=sha256:974faa36ab17430ac6afa0cd1df449703d6f76587998215c8894490d4f135e81

Observation 6b0d68aa-5d91-45a0-872b-69df93598980 · outbound

This paper cites Exponentially weighted imitation learning for batched historical data.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Exponentially weighted imitation learning for batched historical data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.993358Z digest=sha256:1cad423bf6ae077fc249839b077ecb402df9b3afcaeece771ae711df5289ecc8

Observation 77c722db-48ec-49f6-b92f-dfe28046357c · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Behavior Regularized Offline Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.055708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.055708Z digest=sha256:4edc2b2415acd3c3e6d7cf858c45e546bcee1b9998349e844e0644e97d80eac6

Observation fbb913a5-e59d-418c-a15d-31470aad2381 · outbound

This paper cites Zeta hull pursuits: Learning nonconvex data hulls.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Zeta hull pursuits: Learning nonconvex data hulls

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.194510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.159787Z digest=sha256:6455508293effb381bbf3fd32ba0d84bdb95036b4d0daff8409748c6869c058c

Observation 45f17633-9104-4ffd-9565-f2fa0d64e7c0 · outbound

This paper cites Uncertainty svm active learning algorithm based on convex hull and sample distance.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty svm active learning algorithm based on convex hull and sample distance

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.179977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.217496Z digest=sha256:01d51a0027e98281c9e198fefe85fed4e308848845085bcfe255a0e7a4f728e3

Observation a73e2a5f-1b28-4238-8d50-5f4e3eb823c2 · outbound

This paper cites Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.299725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.299725Z digest=sha256:0fb8ce99ad94c5173a29f53adac435410a66c86c40fe54936d4c294abdca9bb2

Observation 29ef1edb-2f09-4290-8ace-a5f42f6e7ef8 · outbound

This paper cites Rorl: Robust offline reinforcement learning via conservative smoothing.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Rorl: Robust offline reinforcement learning via conservative smoothing

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.402534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.402534Z digest=sha256:1d88caf2d738c9a76640ea4598b5b0e2954de6c86a17785a42ed61e4d1b72360

Observation a42cc906-eef6-49be-9a4f-edf5155a047b · outbound

This paper cites Towards Robust Offline Reinforcement Learning under Diverse Data Corruption.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Towards Robust Offline Reinforcement Learning under Diverse Data Corruption

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.698977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.451131Z digest=sha256:15ec4174b52b19c573afbbeec35c558b4feaa12ce86cd1a44e091d7aee4682ca

Observation a0d3b574-9bee-46ee-8dc8-2945e3a3ccb4 · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood In-sample actor critic for offline reinforcement learning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.154600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.547421Z digest=sha256:747970d8177926f9fe260f18fe6cfea271a6972474148a1021c066ca850e5a43

Observation 2572a492-724e-4fa6-94af-d2a55b0b8b92 · outbound

This paper cites @esa (Ref.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood @esa (Ref

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.644839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.644839Z digest=sha256:d25484c4b5f73c54ce61e3ecd9112723614ba472348b1584a0a70cdb5d886741

Observation 772415cc-6c2a-46be-9dac-6ee889b10779 · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.649510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.649510Z digest=sha256:eb336c7890613d76740b005f5bc55fd4f0487e5d1868091da6c6b692941245d7

Observation 9b3481cc-3210-418a-a671-83b889d1099b · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.653919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.653919Z digest=sha256:fb606967d7b4e8c503bd0ab29bbc7cd54388014dd9e216a8a108331496e7b47e

Pith citing papers

No inbound Pith citation observations are available.