Pith. sign in

Paper Citation Record · LEDGER

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood

As of 8 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 0 inbound Pith citation observations for arXiv:2506.08417.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08417 v1

Coverage vector

measured 58 of 58 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:22:27.653919Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

58 of 58 outbound references displayed

  • verified exact5
  • verified fuzzy28
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b2727ca4-f5f9-430c-bae1-d53ed89ade7b · outbound

This paper cites write newline.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.652759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.652759Z digest=sha256:e879946262bf0723e7b72302cdd151f327c95cfaabd403409bbbb217851100da

Observation 04f87bd3-0ea1-41e8-bf93-3fa6826f2fb7 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified q-ensemble.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty-based offline reinforcement learning with diversified q-ensemble

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.687083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.687083Z digest=sha256:f518e9acc2b7934cec120799f9cb6baa5c5c6f346e94790c4c0fa0e7bfcf5657

Observation b62d4df3-f0fa-4a6b-9e96-2a1417d4dbb4 · outbound

This paper cites Near-optimal regret bounds for reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Near-optimal regret bounds for reinforcement learning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.626467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.746503Z digest=sha256:55ffec26d05e37c8739cd79440fc2b054fca108cb9a487c455e19eb0b3735959

Observation 0012dfda-f904-4179-9f74-4956d0a95440 · outbound

This paper cites Manifold topology divergence: a framework for comparing data manifolds.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Manifold topology divergence: a framework for comparing data manifolds

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.611573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.776646Z digest=sha256:cd2cfa1813c4d3a9aaa3460ec825025dc3d74878139e419adb4776d39d29f076

Observation 77245729-e20d-448a-8673-a1ab89bc83d5 · outbound

This paper cites Laplacian eigenmaps and spectral techniques for embedding and clustering.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Laplacian eigenmaps and spectral techniques for embedding and clustering

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.848799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.848799Z digest=sha256:b518fc456d5a2318c5dbc2331bdc58dac96ba979f9957ee69fe50715b19c9120

Observation ad686d95-9820-434d-8f85-954a04ad2cb1 · outbound

This paper cites On the Inductive Bias of Neural Tangent Kernels.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the Inductive Bias of Neural Tangent Kernels

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:23.919649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:23.919649Z digest=sha256:efb5d59fbd35d35cf1c133eb786b744f880bad9865aeb227d9cdc846940caa99

Observation 77b6dbda-ac72-4049-bbe7-f8caf6c48683 · outbound

This paper cites Flows for simultaneous manifold learning and density estimation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Flows for simultaneous manifold learning and density estimation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.586125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:23.989185Z digest=sha256:b21800cc0c082c8e64b5577a9bd2790ca965bba5ec2502b7e86293a60e5259e1

Observation 36370ff4-c1f4-4016-bb26-f32426d32543 · outbound

This paper cites OpenAI Gym.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood OpenAI Gym

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.040414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.040414Z digest=sha256:290c5be182ef44c468547f18ad3c96b876ab846b7c955612f5511429edcc56a0

Observation edfe8984-2ac4-41d4-864d-aa29aaf70d34 · outbound

This paper cites Bail: Best-action imitation learning for batch deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Bail: Best-action imitation learning for batch deep reinforcement learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.097958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.097958Z digest=sha256:c3acf359a93352d008490c6346972352e7c8cd012088e396796a88827c943d3e

Observation c13f2ac5-941f-4bbc-ae58-d7ee4c981374 · outbound

This paper cites Diffusion maps.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Diffusion maps

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.561407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.181646Z digest=sha256:fdfe23ee86735bd36c17747222022ec74f3a5a832f2a2905f62e4f0ca63b8e2d

Observation 46b1f972-6798-496f-a2c9-70fe477b9e2a · outbound

This paper cites Pink noise is all you need: Colored noise exploration in deep reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Pink noise is all you need: Colored noise exploration in deep reinforcement learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.547034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.256430Z digest=sha256:515f62db2b5bb14b13670904a197fc2a91afdf17d1d3b700d5a1e0737f5a5c0a

Observation e57d3fa9-29c9-4044-89e8-69d03c6d1aef · outbound

This paper cites D4rl: Datasets for deep data-driven reinforcement learning, 2021.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood D4rl: Datasets for deep data-driven reinforcement learning, 2021

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.311145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.311145Z digest=sha256:9c1e16bc0e477741e749c9e5a1b8dd140e22523f863e91f316e4f191e6b2ea34

Observation 7ae79ce4-115e-4677-9086-b1f6e9c8af6d · outbound

This paper cites A minimalist approach to offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A minimalist approach to offline reinforcement learning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.522372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.363317Z digest=sha256:154e2ce2e7aad5fb9802eac891f96b48247d4a961a88d5bf839ffe7f8c317ba8

Observation 423edc72-3bbc-4e63-a4cb-c285062e1a18 · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Off-policy deep reinforcement learning without exploration

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.507852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.417323Z digest=sha256:1f0fa9128444fd183f9d6d806362f5bef83539a4c1a8c72b677f4e8ddb655550

Observation 01b87718-6770-47eb-be7a-7a426b0c7700 · outbound

This paper cites Learning rankings via convex hull separation.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning rankings via convex hull separation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.493119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.490677Z digest=sha256:496faaee37d6248aecdce5d00c1b23070c77e4a6b1d858d37557b325e5f486c9

Observation 14dd046c-95d9-4c78-9362-9a3b7e400315 · outbound

This paper cites Extreme Q-Learning: MaxEnt RL without Entropy.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Extreme Q-Learning: MaxEnt RL without Entropy

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.560434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.560434Z digest=sha256:094dee5e9061c69d58aacfda56ddd5e0ffae830ce30676d6969b2e2c426a2016

Observation 0762145d-4da6-4aea-8d80-90ebb6106a89 · outbound

This paper cites Improving Offline RL by Blending Heuristics.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving Offline RL by Blending Heuristics

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:28.060732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.609399Z digest=sha256:142fb8123a35a4bc21a3ad6b53e0d014a784004da26c74e22a2bf4db553b44e5

Observation 5c018bd8-872b-4025-8e05-9254dd4f0ef2 · outbound

This paper cites Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why so pessimistic? estimating uncertainties for offline rl through ensembles, and why their independence matters

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.478149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.668798Z digest=sha256:042edb2e4026ad43728fb26a5c8d6ce39b3edefdd6b6e9e78ce32ad3f6ffbe1e

Observation 4e505ac3-d422-4752-b19e-2ae4a178802b · outbound

This paper cites Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:24.748337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:24.748337Z digest=sha256:8eff622ce5bc629177ec81eabd97aa678b37d537dd93b840e5c6cc0d11bc2c7f

Observation 087f99af-5649-49de-bd47-2c2e0f950ff8 · outbound

This paper cites Random projections for manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Random projections for manifold learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.451972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.810006Z digest=sha256:223c9eb84974b4abeb48a0990516ab3c00db986c20f896ef4daee71e9eac308f

Observation f174c80b-b129-4129-8507-7b5ac451a2dd · outbound

This paper cites Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Beyond uniform sampling: Offline reinforcement learning with imbalanced datasets

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.436324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.868522Z digest=sha256:fb8c2a8965f3d13738d00b2c0c41aafe9f81d3948fc2ba47f5c2a984f559a4cf

Observation 7483e25c-d457-4a7d-85ad-9f0c26e589c7 · outbound

This paper cites Mild policy evaluation for offline actor--critic.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mild policy evaluation for offline actor--critic

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.421916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:24.941158Z digest=sha256:2925d627cd72d360e4abb5a8746dab796c166d56be2933d05b64578bc915c01b

Observation bb277ef5-c440-4243-8801-ae2ca4e47a57 · outbound

This paper cites Neural Tangent Kernel: Convergence and Generalization in Neural Networks.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Neural Tangent Kernel: Convergence and Generalization in Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.087828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.087828Z digest=sha256:20ff568fdfd27527fb372f4447e166512229a75975cb7e9a1dbcbb1eb9f0b45a

Observation 0971bad8-7bec-440e-91f0-f6a9c4586f28 · outbound

This paper cites A convex hull-based data selection method for data driven models.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A convex hull-based data selection method for data driven models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.407761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.188248Z digest=sha256:70de59076a4f817242269d1d7927115120171ae43a27a94fa3f2e2e1b5ab9d21

Observation b435e2c8-7c51-4615-a19b-45bfc93b67ce · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adam: A Method for Stochastic Optimization

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.300203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.300203Z digest=sha256:5a8b599be3470a3ad058aa98f6b7b68b1ea880ad385f4d3a8ce6a7732cf515d1

Observation b4ac79f4-d8d5-4a13-8266-8afc8aabe59c · outbound

This paper cites Offline reinforcement learning with fisher divergence critic regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline reinforcement learning with fisher divergence critic regularization

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.392898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.361489Z digest=sha256:efdb322f0631866939a73e9ee15a1f6a88e4655185ba5f39b4386e9912117e95

Observation 2796467c-3a3c-4708-a2c5-91d74746154e · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline Reinforcement Learning with Implicit Q-Learning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.460008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.460008Z digest=sha256:8372743802c4e563926c8fde19f55a370a58b978fb6a385a6db762afc408a308

Observation 950a4f7b-b3a5-4673-bf86-e4a380591bf4 · outbound

This paper cites Stabilizing off-policy q-learning via bootstrapping error reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Stabilizing off-policy q-learning via bootstrapping error reduction

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.377984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.549357Z digest=sha256:fe8e5b0824d70a9994ecb957c6ba27f2aa050699079a1b8ad54f1741dfda5888

Observation 7acef59d-6aa9-4981-9cec-af33d8400982 · outbound

This paper cites Conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Conservative q-learning for offline reinforcement learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:25.669895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:25.669895Z digest=sha256:81f41fba551a8ecb5734fcd84ac404456a354251c84c594918270f45130663de

Observation 8b551c13-6061-4f88-90c0-b6a1620cb95d · outbound

This paper cites Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Kernel Metric Learning for In-Sample Off-Policy Evaluation of Deterministic RL Policies

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.992371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.764096Z digest=sha256:6489176027b6a13c87417a2b52163a7c8b9bbb643061c3c7035f28ca0fd8a293

Observation 7a30ff04-8ad2-4ad5-a2df-3330feda9fd5 · outbound

This paper cites When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood When Data Geometry Meets Deep Function: Generalizing Offline Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.971229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.844016Z digest=sha256:3b8c0a70fec94059c75441e20f0a924928d8221c7b14158d65a18c3c0e2a0a31

Observation 258495b4-b0ff-4cb4-99e9-f73d6b6800dc · outbound

This paper cites Mildly conservative q-learning for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mildly conservative q-learning for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.352784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:25.916094Z digest=sha256:97b7f7af66e9604eadfcb12f7fea6941abef4df4cd9505e943ab42b719641210

Observation b067009a-5214-4801-85fc-419e7b097027 · outbound

This paper cites SEABO: A Simple Search-Based Method for Offline Imitation Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood SEABO: A Simple Search-Based Method for Offline Imitation Learning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.045813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.045813Z digest=sha256:ad7835391ac8a3dfcc3e9af635b770965e7cadeadf8e2eaab97481be3a31f11a

Observation 1196c970-65f4-4ac2-8829-4bdcd0bce33a · outbound

This paper cites On the role of general function approximation in offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood On the role of general function approximation in offline reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.337075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.115252Z digest=sha256:28430f4dda351d392d5ca77c8cc79e4b24a3d26984e0fd21f7bac01535e29510

Observation 3d6c7838-3c2c-4f35-874e-6992398433af · outbound

This paper cites Machine learning algorithm based on convex hull analysis.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Machine learning algorithm based on convex hull analysis

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.322758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.203111Z digest=sha256:eed5e328859ea14c8b25295a3477cc8ec3594eb10855f029f5c42ce0329fb0cb

Observation 44f94f89-08f2-41ca-be92-788441ea4c08 · outbound

This paper cites Why is Posterior Sampling Better than Optimism for Reinforcement Learning?.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Why is Posterior Sampling Better than Optimism for Reinforcement Learning?

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.932706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.273819Z digest=sha256:85e04c541832402af49352da49fcacffb60df2b666d3259f2bd1452fc5db34da

Observation e3c0c889-bdd8-436e-ba64-193c5f90571d · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.320316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.320316Z digest=sha256:5155c20a20b30a261abb8edb2cda2398345ee72c78e66d82fda48d8f1d2730e4

Observation 4fbc39b4-afb8-4167-9a6f-cd1fdce17504 · outbound

This paper cites Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.393335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.393335Z digest=sha256:7aeb45524da811eb514be5ac7ba3a82f9a838c7e38af8667e2aa5610dd79af96

Observation 87e71b41-e966-446e-9f88-910d9bf76560 · outbound

This paper cites Policy regularization with dataset constraint for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Policy regularization with dataset constraint for offline reinforcement learning

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.308226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.477113Z digest=sha256:d8cc068e5736522e5f59253cc379a3b6e7930dc1a589361d6ad56c6eff365485

Observation 1e868df6-a8bb-44bf-9b57-e2a19a62e438 · outbound

This paper cites Nonlinear dimensionality reduction by locally linear embedding.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Nonlinear dimensionality reduction by locally linear embedding

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.293800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.530029Z digest=sha256:6135746ae6963c90b51bf4c312516d66bc8c5e0dd6f70003710e95c1eb79df9a

Observation e2562a35-92bd-4490-8447-166d9e165625 · outbound

This paper cites A dataset perspective on offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A dataset perspective on offline reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.279695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.572723Z digest=sha256:eb0904227d6b0cc0c147547ab8471115a21d13c35d6e46bb5d7fc3d17c4ed180

Observation 5c6daea8-bc9b-4774-9770-09b969caaafe · outbound

This paper cites Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Accountability in Offline Reinforcement Learning: Explaining Decisions with a Corpus of Examples

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.632916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.632916Z digest=sha256:742253b3211204ad2ee46d1d73f4db175d672ad9e282442cbad3739a9487604c

Observation be1e0c20-24d3-4c46-9125-6725309df06d · outbound

This paper cites Sutton and Andrew G.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Sutton and Andrew G

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.731486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.731486Z digest=sha256:b9ff5bd0c391302535398f09cd581ce6c7feb88cdadd474cf9b7863ebe6b5e8a

Observation bdf784a3-85fc-4a18-a1b8-10cce6db4427 · outbound

This paper cites A global geometric framework for nonlinear dimensionality reduction.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood A global geometric framework for nonlinear dimensionality reduction

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.254341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.777449Z digest=sha256:a68c156edeea5b3daa57cefe569c0cdaf96b524d2a0e8a73a0517693b7c68681

Observation c3550023-30a5-4f75-860e-b65b43c43a7e · outbound

This paper cites Mujoco: A physics engine for model-based control.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Mujoco: A physics engine for model-based control

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:26.818469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:26.818469Z digest=sha256:34d76b1c9a900816a9f343d6ec5c5c65c0ed3af88dd01a0245675d5533a29771

Observation 97666142-1f76-4ba5-8747-d477965e14a5 · outbound

This paper cites Adaptive manifold learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Adaptive manifold learning

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.237822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.909702Z digest=sha256:674462ec9ecaaad7ea3a5aa79ff05714f3f8ad202b2ec897669c35884ad467f6

Observation c32c712c-f7b1-4f43-826c-ec7805ae0790 · outbound

This paper cites Improving generalization in reinforcement learning with mixture regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Improving generalization in reinforcement learning with mixture regularization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.223654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.975185Z digest=sha256:c325544a904ddabab5420b50572c92cd076e79348f0209a2e44ed098fee9ed58

Observation 6b0d68aa-5d91-45a0-872b-69df93598980 · outbound

This paper cites Exponentially weighted imitation learning for batched historical data.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Exponentially weighted imitation learning for batched historical data

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.209242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:26.993358Z digest=sha256:86bb775b5f04c8d8e4d31fdc3b143785f67c48c6d7979588ebb2e6282883e6bd

Observation 77c722db-48ec-49f6-b92f-dfe28046357c · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Behavior Regularized Offline Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.055708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.055708Z digest=sha256:4edc2b2415acd3c3e6d7cf858c45e546bcee1b9998349e844e0644e97d80eac6

Observation fbb913a5-e59d-418c-a15d-31470aad2381 · outbound

This paper cites Zeta hull pursuits: Learning nonconvex data hulls.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Zeta hull pursuits: Learning nonconvex data hulls

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.194510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.159787Z digest=sha256:cbcbf7e8da0cae974050dfb58f690bd70cf6038e90788229adeb17204d1b09e8

Observation 45f17633-9104-4ffd-9565-f2fa0d64e7c0 · outbound

This paper cites Uncertainty svm active learning algorithm based on convex hull and sample distance.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Uncertainty svm active learning algorithm based on convex hull and sample distance

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.179977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.217496Z digest=sha256:4aaf2125617611bb0c0155435bce839d463bcfe283bc0c6217132f88f360ff2d

Observation a73e2a5f-1b28-4238-8d50-5f4e3eb823c2 · outbound

This paper cites Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Offline RL with No OOD Actions: In-Sample Learning via Implicit Value Regularization

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.299725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.299725Z digest=sha256:0fb8ce99ad94c5173a29f53adac435410a66c86c40fe54936d4c294abdca9bb2

Observation 29ef1edb-2f09-4290-8ace-a5f42f6e7ef8 · outbound

This paper cites Rorl: Robust offline reinforcement learning via conservative smoothing.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Rorl: Robust offline reinforcement learning via conservative smoothing

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.402534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.402534Z digest=sha256:1d88caf2d738c9a76640ea4598b5b0e2954de6c86a17785a42ed61e4d1b72360

Observation a42cc906-eef6-49be-9a4f-edf5155a047b · outbound

This paper cites Towards Robust Offline Reinforcement Learning under Diverse Data Corruption.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Towards Robust Offline Reinforcement Learning under Diverse Data Corruption

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:22:27.698977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.451131Z digest=sha256:41969fdea9050d33d87b00a6c17c193973b1da34ce297e2682dc6aa7bea67043

Observation a0d3b574-9bee-46ee-8dc8-2945e3a3ccb4 · outbound

This paper cites In-sample actor critic for offline reinforcement learning.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood In-sample actor critic for offline reinforcement learning

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:22:28.154600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:22:27.547421Z digest=sha256:d2588497b563d580dcff8c2360c53a768f80e75bf12865a6ea57cb48e5c6b9a7

Observation 2572a492-724e-4fa6-94af-d2a55b0b8b92 · outbound

This paper cites @esa (Ref.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood @esa (Ref

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.644839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.644839Z digest=sha256:d25484c4b5f73c54ce61e3ecd9112723614ba472348b1584a0a70cdb5d886741

Observation 772415cc-6c2a-46be-9dac-6ee889b10779 · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.649510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.649510Z digest=sha256:eb336c7890613d76740b005f5bc55fd4f0487e5d1868091da6c6b692941245d7

Observation 9b3481cc-3210-418a-a671-83b889d1099b · outbound

This paper cites an unresolved cited work.

Offline RL with Smooth OOD Generalization in Convex Hull and its Neighborhood Unresolved cited work

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T05:22:27.653919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:22:27.653919Z digest=sha256:fb606967d7b4e8c503bd0ab29bbc7cd54388014dd9e216a8a108331496e7b47e

Pith citing papers

No inbound Pith citation observations are available.