Pith. sign in

Paper Citation Record · LEDGER

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

As of 5 August 2026, this Paper Citation Record lists 100 of 284 outbound references and 100 inbound Pith citation observations for arXiv:2005.01643.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2005.01643 v3

Coverage vector

measured 100 of 284 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-11T11:33:20.892688Z

measured 200 of 200 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 100 of 229 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T11:39:57.880351Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 284 outbound references displayed

  • verified exact2
  • verified fuzzy77
  • unresolved15
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch4

External citation measurements

19
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 017d7b1b-83b8-473f-b2f4-b88dc64ca4fa · outbound

This paper cites and Friedman, N.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Friedman, N

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.783805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:f39864e3df679001d05ed3e2611fb5fe093113bfc4c3aa31dd378f519322647e

Observation 99c1b101-f61b-4b1d-80b4-89d8e5307477 · outbound

This paper cites 2019 International Conference on Robotics and Automation (ICRA) , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2019 International Conference on Robotics and Automation (ICRA) , pages=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.804895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:31ebe7bed366642392d1e03895e83178388252f975373f0e130d1c18fc05b10b

Observation b0f3a3c4-b6c3-4543-b2fa-4a06c9e36c6c · outbound

This paper cites International journal of computer vision , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International journal of computer vision , volume=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.794668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:2cbf6071622aac0621e554079e92a36d9a20b56c15cbc4a0d9bb3638a36bb13a

Observation f460aca3-0c65-4de5-8e18-5348041e0755 · outbound

This paper cites Sim-to-Real: Learning Agile Locomotion For Quadruped Robots.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Sim-to-Real: Learning Agile Locomotion For Quadruped Robots

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:21.431688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:7a2773479fb60c894f5a34fc008a103fafdda119cc40aaf533b731e293d5da02

Observation 5a9fe4aa-5d75-4d04-8414-5000f5573be8 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.826478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4dd17d1d85acd55d8a253da4f50ec307fbca2079684d3115872d1714fe835d51

Observation 97e98e5b-5812-48dc-854f-1e1dc3d0f2e6 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.914240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:10568922f884d9273bf46ee7019c506beb795ef3a49b8f05052860bed88ed464

Observation 6b4362ee-8572-4cad-9520-f98d8095fa38 · outbound

This paper cites 2019 , howpublished =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2019 , howpublished =

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.796281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:fa8232753fe9904a16b37a2621c56c655b831f252aed7ddae3d06f752a2803c9

Observation 9c5f9cad-cbb0-4bbd-ab2a-6ea4a9015b54 · outbound

This paper cites 2020 , howpublished =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2020 , howpublished =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.878462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:50027edfe49bf323aadbc68aeec9b7dae2cd467038010951543d477ec4ebb044

Observation a58eabfc-aa55-4c4c-b1f2-d17bc63f5660 · outbound

This paper cites The Journal of Machine Learning Research , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems The Journal of Machine Learning Research , volume=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.815645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:8f777b911a1bd4d3208a5f0b369c4048f9909f5e501327d40cdccdb82b2b7253

Observation a567dcbf-dcca-4ac1-b9a6-f694138c7c0a · outbound

This paper cites International Conference on Machine Learning , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning , pages=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.782204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:6d1c963d52f6b03ce30c888fe085bc566c13d1b1cd15eecae97f076e9306cb2c

Observation 2ada1639-130f-4b50-bf16-62f12eb21ac2 · outbound

This paper cites Nature , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Nature , volume=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.843852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:0bc2c9e41309722edf87055ba31873a7cb32a3000bfcade41814242a154b2108

Observation 13869476-25b7-4a9c-ae51-a66f7ae0e377 · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.792592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:2e305c49ee77f80a34fe708f3caee38b19903332fb0960f5e55e05e77c24fb19

Observation 3bf082b5-83d7-494a-8c1a-68e90153012a · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Playing Atari with Deep Reinforcement Learning

Reference 13

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T11:33:21.379465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:01e66af2e9088e293135dd7f4cfeead150da9791b52a003752f5c27490452269

Observation 47bfc1ba-181f-46e4-ac47-2e6a80d0d836 · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.880341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:061e6b433955d1e8bac363f0af444c12de6a6b9d036a072d576c6c8effc0fda3

Observation 26ad66e4-51fb-45dd-9a7e-bf843c8ce18b · outbound

This paper cites Journal of Machine Learning Research , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Journal of Machine Learning Research , volume=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.893011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:8cc0a82d05da1044d811ba1f3a32b2905846706b4abcebc2653e2f81540abbbf

Observation d0e7f085-0968-4e03-81c8-02ce96a8a80a · outbound

This paper cites Machine learning , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Machine learning , volume=

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.858305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:8edcea364d80202838916922204047ec7280eaf597aa66a79879c67fc9b996fd

Observation 587f61a2-a360-4d4e-a784-36deaec4be14 · outbound

This paper cites Neural fitted.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Neural fitted

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.785474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:7969f86bd80fa3a9132067b192beb9d88dcc4686ca45153e5d9b75c34b15f1e7

Observation 5ca0840a-1034-4231-a050-8580bf321120 · outbound

This paper cites Reinforcement learning , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Reinforcement learning , pages=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.882117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:159ef91307e5b3b6032e11a4fcff6b55231897d14f8caf3bd1ee3b1a9986705a

Observation ccb0f885-4128-4a95-80e2-c7e9716d2ee9 · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , year =

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.806566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:2d7c355c7936d884dd98c5ba382396294858904a2f93c9bbffb12a2f0bdc5781

Observation 219e78f2-6afc-4d95-84ee-9aa357c4dfa8 · outbound

This paper cites and Storkey, A.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Storkey, A

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.801348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:afd7dd24fc354258efd838b6b0b971e35a508a5465617de1731bceba5840494d

Observation a078c781-cee4-41f6-a2c9-901dd6f112e5 · outbound

This paper cites Attias , title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Attias , title =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.900481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:0b6eecd3ecd43586a2343b5967f17d8859aa2374a4ebde7ae9b3fbd82d3d7b26

Observation 6a747586-201e-4716-9640-7c49e488d92d · outbound

This paper cites 1994 , publisher=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 1994 , publisher=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.871837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a8ce05346128ea153356c07b217518a09d9925bddb099885a606783c5919e5fb

Observation 4e9c36e9-1fbb-4ad2-b23d-5b8fdac0c9be · outbound

This paper cites European Workshop on Reinforcement Learning (EWRL) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems European Workshop on Reinforcement Learning (EWRL) , year =

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.823151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:de42e065c18228c0fc6129585a82a7b2167d5ad1183d7cd08700cfde8b15d74e

Observation 8d1bbd75-fe54-4f35-a35a-4c4ce14afd9b · outbound

This paper cites International Conference on Artificial Intelligence and Statistics (AISTATS) , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Artificial Intelligence and Statistics (AISTATS) , pages=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.824882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a032dd5bb50c55aa1d7eb2fb001b3420b9a7632cfef9ee029dba89278d038b53

Observation 2f84cdf3-522d-4602-8c11-9d5228efe85d · outbound

This paper cites International Conference on Artificial Intelligence and Statistics (AISTATS) , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Artificial Intelligence and Statistics (AISTATS) , pages=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.788901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:8d6273d5055bad88df48c09288a05333ea4ffc6be187799853745c5787b07daf

Observation b6f6b8d3-b0fb-4924-85a5-1ce670763bc9 · outbound

This paper cites International Conference on Machine Learning (ICML) , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , volume=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.862908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:ad66cf379a9724dfc8ee00d58b1fbe2003e1ab3cd3dfb8e03599a5f6989d9dbf

Observation e069cc26-6dea-498b-b704-a2adb79a9d29 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.837165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:324b51d2f8ef2e9077de0ddba0ab460bd06509def959de8f0d1caf3bea633500

Observation fb2f69b3-6662-4dc3-b153-def1f8c19a91 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.840519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a6708d595aae4cec286a67c82ff4d59a84d43562056f6149f6447570411d5381

Observation ec825235-2bfb-4f2a-b7f2-e6a52a83d7bf · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.902271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9f168f252e8ebd13c911373ee65deba015ea71eed84d12f8bd7a291761186f78

Observation 603ea544-8099-4e37-8d22-0eb23adea9cf · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.864577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:1af2c36c5fea0d83999f5a38ee874420ece5b9162ef862a05ec7bd8f9abf7a6c

Observation 1e79c4ba-f65b-4b64-855f-de687b5cfd38 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.861387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:8a0d3c02ad0a64e39b8c5789cc1a1f57a1c1c50c982e839bc43a0868b6ded686

Observation 993532ce-941c-4a66-978b-8f7cbe89a558 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.920968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:6f6581aeda3983a3b99922fb086b695ecfaf05ff6963cf76f1eab4126383ccbe

Observation d6893da3-d4c0-4944-a2c6-836592955515 · outbound

This paper cites and Munos, R.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Munos, R

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.922500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:949d8f5fc349f7b3de11daa9eda907fb444379e1390699bd4b88fddba0a90458

Observation 912ed871-092e-47ce-8d43-7ecfd04754f5 · outbound

This paper cites and Hinton, G.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Hinton, G

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.848813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:957cd805544045b02b923c126625da96ac255c933bda39a420d4fe9e4275d3bd

Observation 1dc4f67e-6162-426b-8a9d-3d484bef25bb · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.856784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:5e3e2a05910580d731aeed864a24c9caeecdad5e33d5be0714eda82442c785ea

Observation 5fb240d4-6aa6-4b92-ac65-ea8a613361af · outbound

This paper cites Todorov , booktitle=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Todorov , booktitle=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.876791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:331716bdc1a6f5ae7b2484e3e3b638603bc33415656ad58bd9e83c58ed285427

Observation 78f55a18-d894-451f-acc4-ded24097d2a7 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.905866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:11959a1a3b0c3607691d57715e49ad6804dec4d60231c376189eb3ec9c49d54c

Observation c660bc69-1242-48f4-8ba8-fa80df1f8e7d · outbound

This paper cites u lling, K. and Alt \.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems u lling, K. and Alt \

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.850435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:ba96f8d19f83af319f6ea94084ccfce824ce3d90bc8ac73909f1f0ad833f87b6

Observation e70eabc4-7b65-4fe4-a2d1-ca0504ef657a · outbound

This paper cites Neural Information Processing Systems (NIPS) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Neural Information Processing Systems (NIPS) , year =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.831704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:71ae1f4708af45d9167a5a9be1770692604358b7184aed7d2419657c8e2a85bc

Observation f40398de-445f-4ccb-943b-c37eb5f6c90f · outbound

This paper cites Todorov , title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Todorov , title =

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.894916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9c33316553495a521777445613ccef6ae2b659edb94764835073ce27d6e23087

Observation 99bb6567-2151-4a9d-bb40-6474e727013c · outbound

This paper cites and Todorov, E.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Todorov, E

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.838993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4c68b519b5debe83dc7666d802f8f260a8d7da0ef1c096a7ee100a6ec5a5d16d

Observation 6ca4a406-5c61-427b-badd-3e8d1b9da512 · outbound

This paper cites Neural Information Processing Systems (NeurIPS) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Neural Information Processing Systems (NeurIPS) , year =

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.904264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:84bf371076963b9e3aa7a618f50932fbb78cd81fdf368d1aa94ad84f44a6d26f

Observation 93c6d035-6c80-4c42-82f3-543732fe82fb · outbound

This paper cites Learning Physical Intuition of Block Towers by Example.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Learning Physical Intuition of Block Towers by Example

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:52:29.539815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:670c88b5603acf322119274edc37d2af123cea2d1828ebd7f6e111c733c4234e

Observation 040187be-39b5-4cc2-aae5-a12ec00f5afa · outbound

This paper cites 2017 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2017 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.799695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:125c7ad7848c985b87f9020ad5419b9b44b73f8cc8d6b3408090ee83dc382cb1

Observation 5952dc03-2148-474c-95c3-69e436c1c24c · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.910808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:206a581b33a1fca9903dc88ef398d9bf8b41f08d4b4b07bbaa167f1ec0210f52

Observation bafe8fe3-fcfe-4e50-91a5-f1239b506e7c · outbound

This paper cites International Conference on Learning Representations , year=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Learning Representations , year=

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.912657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9d36ae3a1f179c1240109ac69d9d7c6f202d164dd6cbf88390c4e81f0b040ce9

Observation 2dd2e29e-31f7-4bdf-90b2-5979171ffdb5 · outbound

This paper cites Certifying Some Distributional Robustness with Principled Adversarial Training.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Certifying Some Distributional Robustness with Principled Adversarial Training

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:21.220075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:efd2e24a7a87541f1254ebca13a1f40de838caf2550b27803201a8ffa86373da

Observation 90c3a46e-8b5f-4284-96ef-cc0390ce4741 · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.884467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:82acf4717c6c32af645a3d567751c621ff53f5845e8b4670014d4b05e7483818

Observation 141421d8-0334-4944-8146-bc7a4b444597 · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.790566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:6c2843b63f790b2877c4bf7f41f462591e0da012ff8a9ab0a4928524f780e0a0

Observation cee9bcc0-cf40-4469-993e-1ec4336af1ee · outbound

This paper cites international conference on machine learning , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems international conference on machine learning , pages=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.817411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:7f184f358b5649c0013e125bb16208dfd70940357f80300b143d0eeeab83d19c

Observation ce7360c0-08ce-4081-8619-f35f2d8f7271 · outbound

This paper cites Causality for Machine Learning.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Causality for Machine Learning

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:21.171791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:b524fb70fb3c8ccad8d64c757e6f0e5b8351e86bc8088be8c658fe80b91d0796

Observation 9900ef4d-f96f-4705-b290-cc233452b690 · outbound

This paper cites nature , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems nature , volume=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.798031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4be986bb79776ad2caeb8cfdb0bd3e2f97711136693168705cdef89f98e95599

Observation caac2d98-8701-49eb-a680-2df16cd46c67 · outbound

This paper cites Deep Reinforcement Learning and the Deadly Triad.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Deep Reinforcement Learning and the Deadly Triad

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:21.360067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:85d760ca950be4447cdc174decc3d7cdb4edce0ed58b89225f16e505cb11daf5

Observation eb374b03-bd38-4bc8-98cd-ab00a9f24897 · outbound

This paper cites International Conference on Artificial Intelligence and Statistics (AISTATS 2010) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Artificial Intelligence and Statistics (AISTATS 2010) , year =

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.819191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:2a1dabaefaad2a64ba88184bdfe45b5e788f25dd53adf042036039738b71fc42

Observation 4ccbe37c-eee5-4df8-a45f-1c7119456e51 · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , year =

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.873578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:e6db1029555279944af4ec60b1922c94a4bbf325f7350b9e64c4f27631cffec9

Observation 66a5395e-34b2-4c9e-8343-3ffda9795d55 · outbound

This paper cites title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems title =

Reference 58

Resolution
parse uncertain
raw_fallback, observed 2026-05-14T10:08:28.842175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9a0e84128e8e976b7d050f14453a4217630417428cb5db4ca9827ef0bcfa8cc2

Observation 13cc71b9-6675-4a2b-9dcf-c40b654231b1 · outbound

This paper cites title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems title =

Reference 59

Resolution
parse uncertain
raw_fallback, observed 2026-05-14T10:08:28.809896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:e02b0173adcd3a428a1549efe132b6a334121a92b9a1655130dfaf33d625e88c

Observation d09b6329-442d-4aef-8bf6-34e071f6409a · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.891283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:801bcb60e6bdee42533bb3a9e99a889aab4391d3ccfbb433b7b26d4bce1f24df

Observation 1424b93a-f4e8-4d5f-b973-932edeba5784 · outbound

This paper cites Advances in Neural Information Processing Systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in Neural Information Processing Systems , pages=

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.803033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:39bbb62aec34f3413c3d71c76c3b3b838730679877958aab5c4e3cbcccff8c4c

Observation 16608f98-dac3-43af-90cc-87a0ff85e566 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning-Volume 70 , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Proceedings of the 34th International Conference on Machine Learning-Volume 70 , pages=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.835574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:c7823b85944cb2a1cb58a00f64148564090c83663c8f109df4d677a5448d51b5

Observation a399f551-3c4a-491c-a470-be7b8f3901d1 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-05-14T10:08:28.833591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4bb480ba62a2f5853292f37fc1e426a511e35e578cbe9a5c19b126ab05767509

Observation 57daec8e-9498-4e2c-a935-6648389f29c9 · outbound

This paper cites Rawlik and M.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Rawlik and M

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.813476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9a0ab87970f0da3ebb417ca566e1f4d277131ae01597b3a9477409c6840ad955

Observation 46df2233-e771-4154-bfb7-14b8d7460186 · outbound

This paper cites Toussaint , title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Toussaint , title =

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.889668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:4a13f663c0f726dbd396295812c290387e4ae09b1798e2775989852b96ceaf8e

Observation c6db0ca3-0bb1-4885-8ac6-d38fc7522d0a · outbound

This paper cites Uncertainty in Artificial Intelligence (UAI) , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Uncertainty in Artificial Intelligence (UAI) , volume=

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.898612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:991c09416be1153427f300ffea268ab3ba3a8bf45c3f44e8e682c6408df7488b

Observation fbd8530c-2ab2-4807-96dc-a0eabad54227 · outbound

This paper cites Kalman , title=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Kalman , title=

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.923972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:6ec41a782bee454106d5eb32de3174bc1770b9fb81ff4cb3d86e0b2174dfee57

Observation a9e0edc7-8f05-48a3-9248-b6d61cd039ea · outbound

This paper cites Advances in neural information processing systems , pages=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Advances in neural information processing systems , pages=

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.828059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:83217d61e9fd3558b65db0260ef51ad536760ab14d05916df90caf1f2dec0874

Observation eb548699-f795-4bc5-94e1-f14d8f8a1fd3 · outbound

This paper cites Machine learning , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Machine learning , volume=

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.870212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:5cf96b1009f515040a512dc21f1ca3ffd47057221fe76e20acfbb6b85c1ac882

Observation 8a8d5695-d1ef-4573-b444-af10c596fc82 · outbound

This paper cites Machine learning , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Machine learning , volume=

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.886287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a5dcc5c376f6300afcb791fcb69c9dc8a9bb4dab65e41da69b47b3364cfba003

Observation 9704249c-e542-4a9d-a7cb-a8c861cfc365 · outbound

This paper cites Todorov , title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Todorov , title =

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.887954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:ebb479df6a14019b7c12ff1735e58fd255b4e6b8a1bee75fec7f25eebe39d485

Observation b010ca1b-f766-4b39-8d6f-169ca18fe369 · outbound

This paper cites and Schaal, S.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Schaal, S

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.866386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:ead5efdcdfb8bea4a51a790d10ebe22d10a7b68cbfe37cbe92802bb6f336f169

Observation efc5bd70-edb7-474d-99b8-22f7cb184ff9 · outbound

This paper cites Neumann , title =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Neumann , title =

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.919341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:e5875e00c52c4eb37f78124b8453764ab2f1523a2fc20c4e2e36c5a01e01b458

Observation f81744cf-1c30-4691-a18a-917fbbef84ca · outbound

This paper cites Levine and V.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Levine and V

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.853691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:356f565f9e895848a1c7ba23211d49079f6dacb8a1e267d74619dc3535e515c1

Observation a598e534-ab2a-4f97-9e34-90854a5948f4 · outbound

This paper cites European Conference on Machine Learning (ECML) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems European Conference on Machine Learning (ECML) , year =

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.859868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a2c7787bde1d3d6b12a201aad27b959b7e7f0867eed198ded5e50c30942028a1

Observation 78de8f95-197e-4cd3-a315-360f03990d38 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Journal of Machine Learning Research , volume=

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.845494Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a3e251f95fcb39bae36cd54ac4d57cb84a4c98cabf9a787f734f038aeaa6cbca

Observation 6fe590dc-06d4-43ac-8a85-81f43203b9e2 · outbound

This paper cites International Conference on Machine Learning (ICML) , year=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , year=

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.868499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:f9c8b943bd3a79d260247a2652d6e78848d691f595459845b9ed6ca5efd2adab

Observation ed0b712c-afe7-4c84-8985-147cb7a1b35e · outbound

This paper cites 2018 , booktitle =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2018 , booktitle =

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.811747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:778c1cbbc02c92e1b6459a3847e439e70b1fae4ff297042084ca90e33ef24cc4

Observation d6871d2d-813c-4643-89b0-193b84aef1dc · outbound

This paper cites 2020 , booktitle =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2020 , booktitle =

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.852095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:60982222b39c73db61255a354de758417101b547aeb72f883c37551561660d74

Observation 729ba718-0073-4eec-9431-996330d01b0c · outbound

This paper cites 2017 , booktitle =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2017 , booktitle =

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.915831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:dbc5eb4bc7bc586aec447fe8fd10a042be165686949d7a4fc7dcb487b6849488

Observation 26624040-4489-4354-a09e-a28f75c6d8f1 · outbound

This paper cites and Koltun, V.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Koltun, V

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T10:08:28.917753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:118207383a5a5ec9e567f478b88b4aab566f3de51f20c0760c2b7f8aa6af5097

Observation cfe2ece7-d32a-492a-860e-8ea4d23a6c88 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-05-11T11:33:21.714139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:157e1fb64808a7d14d3f23ce91927f550c9d0d277d1cf5ef578a3799cb62db81

Observation aaca1e55-a15f-4f7b-b26b-f8f5161cdc67 · outbound

This paper cites and Pong, V.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Pong, V

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.718120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:51a7c2c885484b5514d846d73bc0453f779673eb9a4ad334abb6ee6fb5b7aab9

Observation 21d6ff20-bd47-40c7-8b7e-8616ebe8846b · outbound

This paper cites Nachum and M.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Nachum and M

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.722900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:91eb5c9cd5029f0ff08f4f97567113491440866b0b0eae6aeec85cb2f25c0857

Observation 6eeb3b54-aa3d-467b-8cf5-c70f27717686 · outbound

This paper cites and Popovi\'.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Popovi\'

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.727855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:15fd3e15fd99c8516ef09624ffe5a9f4f5504d65ca8c5c42b1fd57f26d6cd98a

Observation 8032c13b-c8d5-4c19-84bb-24042d8689ab · outbound

This paper cites Levine and V.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Levine and V

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.733754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:44d1731e154023602c485f4be17fd0f5cdb583f6778731d951756b5b113d4c40

Observation 3c51fd4e-d0ac-4f50-9cc8-290fe3fd2209 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-05-11T11:33:21.739083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:38cc88be2bf4c238a84b201ae5944bd3d2bc553453f84427bc0a6e30cfaa2809

Observation 524ca5d7-3f8b-4b6d-bff8-fa0e0b30e02b · outbound

This paper cites Wulfmeier and P.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Wulfmeier and P

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.745334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:1b38cf70beaa032f3bc728066d40a23a027f53c6a6c2f73c38ce2a2fb249dd88

Observation 1165a998-d2ed-4e30-a0e0-9ac7b49a63f0 · outbound

This paper cites an unresolved cited work.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-05-11T11:33:21.750548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:c27a167fb002fab092ed0c06fad1afcba75681189fd5c240148027e9c47b7e61

Observation 9c2226fe-349c-4ed3-8655-f2ca6791d08e · outbound

This paper cites Huang and K.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Huang and K

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.755430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:3a5508979fff9faca067b77b9e3d7bd2be493c00402cad8ef02c317e4c42d448

Observation 300e8f57-5b77-440a-93a0-eae67aafc52a · outbound

This paper cites Huang and A.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Huang and A

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.760387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:a9e75979af9cb8d5a743f6db6d35e138d9339049928023630b3a0659b3b160db

Observation fa80f8e7-95c2-40f6-b46a-2f589c87367c · outbound

This paper cites CoRR , volume =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems CoRR , volume =

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.769769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:d40ce06620c33db8c58ef4773a282b9b090ccc7a62c8bee001cbe59d26b54c0f

Observation 22563612-0b83-411a-a13c-0893c0dd71c3 · outbound

This paper cites Javdani and S.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Javdani and S

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.777702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:7c82698a16c1ec82fb1530406725cff3c943c14042929d4c74eee9cbaebe24cd

Observation d45f29a8-2cdf-479b-8a03-a289d9d8f81f · outbound

This paper cites 2018 , author=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2018 , author=

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.783811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:875265841c5f742505aabae06ac3a2a6d19f4d01f16237f9a629725a83a2d0b0

Observation 7c23d3c6-1c6a-4cfb-a909-5702a7ac3643 · outbound

This paper cites International Conference on Learning Representations (ICLR) , year=.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Learning Representations (ICLR) , year=

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.790421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:9e73e20738f97936d113565b902d7cf61bd166f3edcd6e3f86c058c60c229df9

Observation 10602982-8f0f-4abb-8482-4f7918b9b704 · outbound

This paper cites Gupta and R.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Gupta and R

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.797488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:d0192cf3bab70e2c603652d30f35de6a9e9c9830ebf9a7a3dae90b75f7777613

Observation b6c630a0-7ab4-4352-86cf-a609963390e6 · outbound

This paper cites and Finn, C.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems and Finn, C

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.802664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:d3a19b8eaf300e95ddade8b2e497a425c9e4c5c6b543c7a261b13c97eb17c520

Observation 5a31aef9-3e84-471c-87be-caa11d9529fd · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , year =

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.808893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:1652b8e87b2b70e9f8e6fc0536b90ece1d59c152906b7dad2ff005045b0b4248

Observation ddf69ec6-8a02-4549-a84a-37bcb3a224f1 · outbound

This paper cites 2017 , booktitle =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems 2017 , booktitle =

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.812618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:d9496a3ab5fc4d8075bd23f11ea4ff940aebde90d3c8dc0115416106dc2b0921

Observation cc95fd18-d503-4564-ae63-7222ba6dfc59 · outbound

This paper cites Mnih and K.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Mnih and K

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.816384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:0673bee177306700fc9264205a98abd2a37839629f9c9871ebd33bc2d361d12a

Observation 8d001b6d-35cf-4d9d-ac16-97999581fc78 · outbound

This paper cites International Conference on Machine Learning (ICML) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems International Conference on Machine Learning (ICML) , year =

Reference 101

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.825058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:39e7a0b4bfd2deafcbf74585cc4b0ab14398a521aa88cc58b0f95f41b3025a43

Observation ebbba56a-785d-496a-8e8e-4f57c08e9cab · outbound

This paper cites Neural Information Processing Systems (NIPS) , year =.

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems Neural Information Processing Systems (NIPS) , year =

Reference 102

Resolution
verified fuzzy
raw_fallback, observed 2026-05-11T11:33:21.831420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T11:33:20.892688Z digest=sha256:0f2d9607cc77b0112c0dedc1175bf3630ebe5f1bf72154eca16902dfa8d5445e

Pith citing papers

Observation b1bc3922-d0e6-40bc-89f5-08672e92a0da · inbound

D4RL: Datasets for Deep Data-Driven Reinforcement Learning cites this paper.

D4RL: Datasets for Deep Data-Driven Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-12T23:19:17.449574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T23:19:17.322890Z digest=sha256:f16d028596cab1514bea68c730a239ee3f5f9534009e52ded4b5d3dfaf6317f4

Observation 67458dd5-8f57-4ddb-bd19-04980927a21b · inbound

Decision Transformer: Reinforcement Learning via Sequence Modeling cites this paper.

Decision Transformer: Reinforcement Learning via Sequence Modeling Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-18T15:11:11.344917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T15:11:11.056013Z digest=sha256:c99f20253302bdb57f60e3f68041daf2ac507ac1a6abdc9ca4bb962a3cff0444

Observation ccee5352-07a0-494c-b973-e2f5fc11e219 · inbound

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation cites this paper.

What Matters in Learning from Offline Human Demonstrations for Robot Manipulation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-13T08:51:55.877106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T08:51:55.826747Z digest=sha256:52890e6375039eee114bf2ba5f4635c161377104ee8eb44a5f0cacc49447c081

Observation 03bee1b7-29d0-421c-a60e-cd7edc4a9e25 · inbound

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training cites this paper.

VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-15T04:42:52.738375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T04:42:52.627166Z digest=sha256:5373b47478ef4057f09ee53435d3ad9a52b879d2b9a9e8300b2faacaa4b62686

Observation 94e12715-8977-40b8-aea8-de7e1433cefa · inbound

IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies cites this paper.

IDQL: Implicit Q-Learning as an Actor-Critic Method with Diffusion Policies Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-13T13:48:36.515834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T13:48:36.369334Z digest=sha256:5a1abe78310879d5a6eff5e10959772b504a00b006829e20d0298c62303569c5

Observation 230da3f9-9d98-4a05-adfa-58a638b57531 · inbound

MiniLLM: On-Policy Distillation of Large Language Models cites this paper.

MiniLLM: On-Policy Distillation of Large Language Models Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-12T17:40:27.920802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T17:40:27.827493Z digest=sha256:0128e33f2fc5ed61e2f5ba05b02f0edecae57d1614fc36986d1aa6120b3fe356

Observation 15b3fab1-a780-4cf7-982b-c213faa7ef53 · inbound

CROP: Conservative Reward for Model-based Offline Policy Optimization cites this paper.

CROP: Conservative Reward for Model-based Offline Policy Optimization Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-24T06:39:01.288888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-24T06:37:57.104233Z digest=sha256:d892b2e45d2ae685567709fe1f1905ac88944167abf771a25571aac3907e76bb

Observation 00bd38f2-6fad-4a31-85ac-22977eb76bed · inbound

MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations cites this paper.

MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:47:55.083411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T09:47:54.977716Z digest=sha256:0a12d2f2f8e08a720f8817fc0692849aedd66b08401e6619346657a58c6c3a89

Observation a0d5459e-2f70-4cd1-ad06-11e95c3c0290 · inbound

Remember what you did so you know what to do next cites this paper.

Remember what you did so you know what to do next Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 15

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T08:25:33.818549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-25T08:20:54.753641Z digest=sha256:d18467331c3a7277eb390a9a5db741beac62469bda8dd61dc0289dd7b6e8ac44

Observation bcddf5fb-13ae-4edb-898a-605610b07764 · inbound

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots cites this paper.

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-12T23:46:30.288548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T23:46:30.198761Z digest=sha256:aa1234d2f194a614327fb212fd52916f15379660cc01e6a82e7fc13a2cd1238e

Observation 04116204-3098-4c78-a6b0-3049b13ff9de · inbound

Multi-Objective-Optimization Assisted Data Collection Framework for IoUT Based on Offline Reinforcement cites this paper.

Multi-Objective-Optimization Assisted Data Collection Framework for IoUT Based on Offline Reinforcement Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-23T19:13:21.908396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T19:09:56.559531Z digest=sha256:ba6f034cc6ad1ed18e18b73f1aae1aa76e0950a3b5f1c62550a8659cf767bee8

Observation f19f4684-9397-4f86-867c-52b14d28a2b1 · inbound

A Tale of Two Cities: Pessimism and Opportunism in Offline Dynamic Pricing cites this paper.

A Tale of Two Cities: Pessimism and Opportunism in Offline Dynamic Pricing Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-23T17:43:17.889714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T17:38:17.004171Z digest=sha256:0304ed96bb515a31f9d6164d45e98596a90dab4601d4f5faa01061e735c554ef

Observation 97fa3e47-0c49-4e9f-b1ee-8f7439296c38 · inbound

Meta-Offline and Distributional Multi-Agent RL for Risk-Aware Decision-Making cites this paper.

Meta-Offline and Distributional Multi-Agent RL for Risk-Aware Decision-Making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-23T05:12:34.919542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T05:08:26.606825Z digest=sha256:ed5e1e4646b8abae8bc5b0b6958d31eb9565c83aaf4f8608101d824286d75e80

Observation 37b2b7a1-cb1f-4ef4-9fc1-0ed1aa40400c · inbound

A Review of Causal Decision Making cites this paper.

A Review of Causal Decision Making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-05-23T02:47:26.265952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-23T02:47:21.582667Z digest=sha256:5beb8e18ce2329f033a0c1e7793dc2af920338c385b86d7b5907ae92bd85a5b0

Observation 0aa23041-a176-4a0e-8f20-38976287eb82 · inbound

VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning cites this paper.

VIPO: Value Function Inconsistency Penalized Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-22T19:35:03.977055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T19:34:33.263674Z digest=sha256:e32f298a38cf61687efa88e307d7145071641809bf909e3df3aa8ffcc7277337

Observation b9da10df-a74c-4cec-83fd-3d14da4f5134 · inbound

Offline Constrained Reinforcement Learning under Partial Data Coverage cites this paper.

Offline Constrained Reinforcement Learning under Partial Data Coverage Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-19T14:22:24.068643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T14:18:48.007454Z digest=sha256:ecb0d1c78d0292b913567bca20e6521b120ed6e62f2b7a711b75819f6d734a2d

Observation b21890e2-3a96-40fe-b354-5b868aa601e0 · inbound

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning cites this paper.

BiTrajDiff: Bidirectional Trajectory Generation with Diffusion Models for Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-19T10:52:15.246580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T10:48:28.980868Z digest=sha256:f599cfb9de60852cedba932adb9da5bdb4342a387d907b3ce83a769857b15069

Observation 2fbac1f2-4a0e-47b7-a636-c258239421ce · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-19T09:17:14.217433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:6cfc24dec828bd554c2ce038c38c07d03c2a6c18e1ffc5b1dc0a9f09b7889c2f

Observation 1a90a0e5-82de-450c-8ab6-40cc06427cbe · inbound

Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling cites this paper.

Blending Supervised and Reinforcement Fine-Tuning with Prefix Sampling Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-21T23:40:46.464894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-21T23:39:39.018498Z digest=sha256:8044f06b18bf20f642770718ddba9fc53f6a158a86e07bf63353883bf31b68f8

Observation dfd9504e-4c08-4393-b1a9-432863d59bce · inbound

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers cites this paper.

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:57.880351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:57.880351Z digest=sha256:8366e508fd8cde77df9286a1ecaaadb38ca5b094c98d6650ad5835f81cb56ed2

Observation 722d2c77-9226-4913-90dc-3c73ee52b394 · inbound

Reinforcement learning meets bioprocess control through behaviour cloning: Real-world deployment in an industrial photobioreactor cites this paper.

Reinforcement learning meets bioprocess control through behaviour cloning: Real-world deployment in an industrial photobioreactor Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T23:04:16.944975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:04:16.944975Z digest=sha256:916cd86891cff6bd2bbc1c0d9354a5da7561da26a7a038de4654793b4952d60a

Observation 55c5924b-a1cb-4929-8e47-abb71b374845 · inbound

Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning cites this paper.

Wavelet Fourier Diffuser: Frequency-Aware Diffusion Model for Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T10:30:59.055699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:30:59.055699Z digest=sha256:160c03a4cbd21941991952928c65c7f7042d947dc8cf4b9607528b6b98e3a54a

Observation 54cacf35-c3c2-465d-a4e0-e4b79061e6bc · inbound

RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations cites this paper.

RobustVLA: On Robustness of Vision-Language-Action Model against Multi-Modal Perturbations Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T14:58:36.786424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:58:36.786424Z digest=sha256:d34eea38eb4874a48f9e54b894acf79190bb5976528521116d5fe8b8735b976c

Observation 5fa3a69c-1c3a-49d6-abe8-1b2a4bd32e81 · inbound

The Three Regimes of Offline-to-Online Reinforcement Learning cites this paper.

The Three Regimes of Offline-to-Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T13:00:07.505888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:00:07.505888Z digest=sha256:4860bb44355b9e5107e52969ceee344de4b41c5ce5e2d62f3e128e86c0bf6898

Observation aab9c65b-3ee7-438f-8ac7-dcbd634937b6 · inbound

Comparative Field Deployment of Reinforcement Learning and Model Predictive Control for Residential HVAC cites this paper.

Comparative Field Deployment of Reinforcement Learning and Model Predictive Control for Residential HVAC Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T12:56:15.017647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T12:56:15.017647Z digest=sha256:9207036a8ae22426b22a4355a9335e8035d994f03f6a1a25e53bae0b452984a9

Observation f6adef32-a557-4f37-a420-a78dad14b28b · inbound

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets cites this paper.

Density-Ratio Weighted Behavioral Cloning: Learning Control Policies from Corrupted Datasets Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T20:54:21.637716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T20:52:08.062450Z digest=sha256:4680a17430f9b4b9dc42b666baf351b10d519f6946a65628dd4789121bb80820

Observation c5fa8801-c96a-4919-94b8-24471e516dac · inbound

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care cites this paper.

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T10:51:10.090896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:51:10.090896Z digest=sha256:5b465da47b52692abafc090e50a50c301f92028f18fd2520a6b00fda40de121f

Observation 99e667d1-68b0-491c-a0f1-a3dc85b2110e · inbound

Mixed-Density Diffuser: Efficient Planning with Non-Uniform Temporal Resolution cites this paper.

Mixed-Density Diffuser: Efficient Planning with Non-Uniform Temporal Resolution Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-18T05:00:54.923049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T04:56:47.744306Z digest=sha256:ec094cf16691433b58576879defa287b480f561ca09c86b016313b4c5b9bb058

Observation 4e107a36-59c2-43d5-bbd7-e9b6dba8b038 · inbound

Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach cites this paper.

Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T08:03:19.816134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:03:19.816134Z digest=sha256:85f02c46678d121669b52a056fd9c8c97f56c8ac29d273a9c72e961467b99dd5

Observation fc706218-b02e-4a30-8c44-20b76a4c55d4 · inbound

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer cites this paper.

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:31.568689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:31.568689Z digest=sha256:d48ee87eb6b951400d217040180c9a97dda9f393ffa3bd65bdd8c5552912153c

Observation 9bd99190-6207-4cd0-b160-27a68062b4ac · inbound

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters cites this paper.

Stability of Certainty-Equivalent Adaptive LQR for Linear Systems with Unknown Time-Varying Parameters Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-18T00:00:31.574994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T23:57:37.299127Z digest=sha256:f11ed6d7982c933c36396a3b903b762704b76cbf91ad989868ad4e89d7c70505

Observation a8beca63-e9f0-4739-8ca2-675a1853e085 · inbound

$\pi^{*}_{0.6}$: a VLA That Learns From Experience cites this paper.

$\pi^{*}_{0.6}$: a VLA That Learns From Experience Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-12T10:34:59.242880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T10:34:59.134604Z digest=sha256:f24d2f0cad8483230b323b3487fec0d553b655aba94a8aef2206769036556c11

Observation 381624a3-fd54-42a8-8cf0-9789245ee286 · inbound

Outcome-Aware Spectral Feature Learning for Instrumental Variable Regression cites this paper.

Outcome-Aware Spectral Feature Learning for Instrumental Variable Regression Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T19:26:22.911048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:26:22.911048Z digest=sha256:296b7f460807c5e5c0f58e38777b277ea44514ad613d94a460698ed48266b58e

Observation 1355548b-b4e1-4464-aea1-bf78080d6e00 · inbound

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism cites this paper.

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-17T01:48:51.134138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T01:46:16.923948Z digest=sha256:ee45a7029a586fb3a8c11e52c48abd9dd341aca837de2df1f1f2fadca5fdd977

Observation ebe15c9c-7834-4c68-9801-01c5c0da807e · inbound

Pseudo-Expert Regularized Offline RL for End-to-End Autonomous Driving in Photorealistic Closed-Loop Environments cites this paper.

Pseudo-Expert Regularized Offline RL for End-to-End Autonomous Driving in Photorealistic Closed-Loop Environments Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:48:32.649055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:45:30.382235Z digest=sha256:6172e100944f6254971ebb32c47eb625cc84e4d635acb32ee01822a771b9d617

Observation a3c0d4c6-a65e-4d8d-8583-62085bf3c4cf · inbound

Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning cites this paper.

Generalisation in Multitask Fitted Q-Iteration and Offline Q-learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-16T20:21:13.582594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T20:19:28.727598Z digest=sha256:c8ad44924df942850a50f53405233656fa7a6a315f2c9780715f257bda0e1189

Observation 7238425d-08c9-4eeb-9a6b-7d30bc83d945 · inbound

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management cites this paper.

MetaboNet: The Largest Publicly Available Consolidated Dataset for Type 1 Diabetes Management Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T13:20:57.304198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T13:20:02.622343Z digest=sha256:40a24b7b4d47f61e49ae45a116fbb60aa4f82b9848460a65c5bdfa975ff2d558

Observation c40be30f-8a3a-46cb-a4b7-fc4eb8cfabb1 · inbound

Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data cites this paper.

Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T09:02:27.089561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:02:27.089561Z digest=sha256:8a3e62a5b42d7292e244a90b7a20714f2668bb49ba58d64b5b90f4f92899d22c

Observation 9312abe6-38bd-44c6-844d-7373ed5e7670 · inbound

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care cites this paper.

On Safer Reinforcement Learning for Sedation and Analgesia in Intensive Care Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:40:12.613570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T13:36:12.920714Z digest=sha256:a2406496ee56429e5d21efece8fdb646c7589a436de49ace671e5f9ae1635bcf

Observation 92aec2ef-d032-4ea3-ab13-04374749dbd1 · inbound

Adaptive Control in Autonomous Driving via Real-Time Recurrent RL cites this paper.

Adaptive Control in Autonomous Driving via Real-Time Recurrent RL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-21T13:50:12.734085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T13:47:12.875888Z digest=sha256:18cf3ffdbd7ab7b5c15d7b075a660907aca2b5d6ac4fa6787ddd85597dd45db8

Observation 91a39f13-16cc-4343-b581-d8828b6e1951 · inbound

Beyond Success Rates: Trainability and Extractability for Offline GCRL cites this paper.

Beyond Success Rates: Trainability and Extractability for Offline GCRL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T04:18:40.519980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:18:40.519980Z digest=sha256:110aee831ab6eee6b3b3dc9a566679f461a8f4fb5b5598495c2e953a3f8c548e

Observation ee16f694-368e-4cf0-927c-23755537ff15 · inbound

Can Vision Language Models Learn Intuitive Physics from Interaction? cites this paper.

Can Vision Language Models Learn Intuitive Physics from Interaction? Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-03T04:05:23.006057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:05:23.006057Z digest=sha256:4e457dad096092838f1db2c37cbc995ab05053b68e11660df47cd61f8a31cce4

Observation f1abecfe-8b5f-49e6-8acd-95049ea199a3 · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:34:12.906138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T14:34:01.332088Z digest=sha256:f26cd85312d1ae147b90e5a632b1396516672a5a5a1fe36b9c98fa154c673b3a

Observation 1b07d353-ac51-4c15-8b1d-c3ca0c98df48 · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T04:00:18.374520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:00:18.374520Z digest=sha256:8d9f9e94920607f8e4244ed84353e24a3d1f24b6a46bd3202ab2807216ec8cf5

Observation 9ec55555-4e6d-4a27-af2b-968df06fa147 · inbound

The hidden risks of temporal resampling in clinical reinforcement learning cites this paper.

The hidden risks of temporal resampling in clinical reinforcement learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-16T07:07:29.666521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T07:05:16.379996Z digest=sha256:328cbfe213ebc3b2bb2f5fe8a49a7201d7d4aca3ef367bffb786250bb5980d7f

Observation 6e13b17d-89c0-4158-8e60-67fe2b30ed21 · inbound

VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation cites this paper.

VGAS: Value-Guided Action-Chunk Selection for Few-Shot Vision-Language-Action Adaptation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-25T07:00:26.181856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T07:00:01.741166Z digest=sha256:db4c8ab32c4410cc8027c8d98400cda9c96c5e74e05184931477d09ec3c5d5c6

Observation 7ea1ead5-d6a2-43ba-8552-a7b346e8cf69 · inbound

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows cites this paper.

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-16T03:37:13.947853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T03:36:09.272019Z digest=sha256:2f0c91e6074b48546265f3c96f5b4a77aae5a4cc0727db8aa96a3298724ef524

Observation dfddc979-3e71-4b82-8edf-f7c9a5014c74 · inbound

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows cites this paper.

SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T02:53:16.595689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T02:53:16.595689Z digest=sha256:e30e0346c44af54657333459c96dc02c9a5d9eae9bdbf473f82c527fa04319ba

Observation 65a9af51-c76e-4870-9f5f-d7ca59236381 · inbound

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates cites this paper.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.075775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.075775Z digest=sha256:b1568b483aa203e7b80298aac622eafe7fe79b5d1d9c02daa60dbc69e2f21b9a

Observation 3007d3db-bf77-4310-acdb-f6862a5bc763 · inbound

RISE: Self-Improving Robot Policy with Compositional World Model cites this paper.

RISE: Self-Improving Robot Policy with Compositional World Model Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-05-16T02:30:31.963711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T02:28:37.997148Z digest=sha256:a6fa86f4fe7dee2dc0391a2742d35c0cb86b4f7c42ccb6a26965154a58fcf954

Observation 2755953c-b859-48c9-a2d3-eb29f534a7b7 · inbound

What Matters for Simulation to Online Reinforcement Learning on Real Robots cites this paper.

What Matters for Simulation to Online Reinforcement Learning on Real Robots Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T21:36:19.144684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:36:19.144684Z digest=sha256:7fb0bedc1292053604fa3054e556fce18c21140ce1e3bb967cb30fe4e2fc750b

Observation 63fd0a51-1e7f-485e-8265-ec4ca40cff1c · inbound

Mollified Value Learning cites this paper.

Mollified Value Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T20:29:19.038948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:29:19.038948Z digest=sha256:415f0a4af053da5d047d3907a22bbbb3948210860d10993c4ace002eb7a452bd

Observation 5aa4410a-6f85-4c8f-87c6-82e227245baf · inbound

Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning cites this paper.

Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T20:00:58.981863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:00:58.981863Z digest=sha256:e2c8f016026ca35fe8eb6f75dc3466b62619ce4722b9fb2285d60408dcb9e589

Observation 1205f5df-a97c-4723-8a8c-c8958a5c26a0 · inbound

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning cites this paper.

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-15T17:40:11.795471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T17:37:03.704563Z digest=sha256:584e0cb1e7b2fbc558dff1c06c3fb6fea1f3f798bc0440f7b47d1d7ec7b07d62

Observation e4dc9d65-5c2b-43e0-83cf-55e578101f71 · inbound

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning cites this paper.

The Informational Cost of Agency: A Bounded Measure of Interaction Efficiency for Deployed Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-21T11:45:03.184188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T11:44:30.880602Z digest=sha256:15189a3755365472e378223ad70ba2563463ae2abdee5176dfd3659ab5d48876

Observation 927b2e5c-82e4-448e-bd24-903a4f5f2053 · inbound

A Survey of Reinforcement Learning For Economics cites this paper.

A Survey of Reinforcement Learning For Economics Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T18:34:53.020591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:34:53.020591Z digest=sha256:365ac63897b97e67675912f64d215fadc513248f647f5553b5fb7635f57cfc76

Observation 484d1eaa-9c7f-4e61-a1d9-35cb57109e32 · inbound

Contextual Intelligence The Next Leap for Reinforcement Learning cites this paper.

Contextual Intelligence The Next Leap for Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-15T21:56:40.742119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T21:54:41.748596Z digest=sha256:d2ebf5055fe5edcffa014947f4536d6fbc09cee0f66b4da2a8605bf7a3007352

Observation eff3b642-bd4a-4c0b-bf8b-0876ca3c502b · inbound

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models cites this paper.

Mitigating Data Scarcity in Spaceflight Applications for Offline Reinforcement Learning Using Physics-Informed Deep Generative Models Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 40

Resolution
malformed identifier
local_arxiv, observed 2026-05-13T21:38:18.422160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T21:35:52.012244Z digest=sha256:85f1a77684c95204be4eb1b6e1001dd20f15610c1a9d282d9aaac72047f19439

Observation 43b1278d-07dc-41e8-b9c4-e289af13e39d · inbound

Offline RL for Adaptive Policy Retrieval in Prior Authorization cites this paper.

Offline RL for Adaptive Policy Retrieval in Prior Authorization Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:13:12.077414Z digest=sha256:60f3f9d255d125468a7e3f7974e79672f9ee6ab5fe8eea799aa16fa5813a8474

Observation e3d1c6f6-50ae-417f-bde9-799970ac52ac · inbound

Cross-fitted Proximal Learning for Model-Based Reinforcement Learning cites this paper.

Cross-fitted Proximal Learning for Model-Based Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T19:26:15.890536Z digest=sha256:4ac0c9ffc0084a9c48bce7629adf67508bd7e006c84153ca60a371af1f87e987

Observation f555d2f8-c58f-405e-90f9-2fe7ec31994a · inbound

JD-BP: A Joint-Decision Generative Framework for Auto-Bidding and Pricing cites this paper.

JD-BP: A Joint-Decision Generative Framework for Auto-Bidding and Pricing Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:40:01.560145Z digest=sha256:6ab7fc0d7e728a69be538195d9d9c27fc5cc67eb54fdb8fca7683d467094f91a

Observation 496a6180-c6f7-47f8-a477-02b877343bfe · inbound

JD-BP: A Joint-Decision Generative Framework for Auto-Bidding and Pricing cites this paper.

JD-BP: A Joint-Decision Generative Framework for Auto-Bidding and Pricing Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T16:46:55.922489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:46:55.922489Z digest=sha256:ce67a470dc4fa494e01838e8f8e9ca299a4c02caf0bc975e9901ba19bdca05ff

Observation ece9362a-9746-4cc6-8f1f-e9be583b0311 · inbound

The Cartesian Cut in Agentic AI cites this paper.

The Cartesian Cut in Agentic AI Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:54:45.333038Z digest=sha256:8023872fecd82a9e9b69385e97712b037c30e2d1dcbcba68ecffe151cc6950ea

Observation fa1ae9dd-14e2-414f-9337-874cb46f0f1a · inbound

Automotive Engineering-Centric Agentic AI Workflow Framework cites this paper.

Automotive Engineering-Centric Agentic AI Workflow Framework Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:32:09.896596Z digest=sha256:6a9b2471f6a1178de6386fbac06fde9a321d70e5620f67306da9b5ffc95f76fd

Observation 0af77945-106d-4639-ac7c-86d64126b930 · inbound

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning cites this paper.

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:28:58.515666Z digest=sha256:8b973d50cf49260f24ce630cff0614d28241cd78e507cf7d436247de8978d391

Observation 3d9a76c0-e56b-4300-bf6d-9a5535a2e345 · inbound

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning cites this paper.

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T17:12:08.970164Z digest=sha256:192e1469e3a89529abd5b64cc6c5c3b08ab5495ec01c4b39d4ad14175c71d6bd

Observation 0a373dd7-b962-4f43-ad12-48e207303235 · inbound

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence cites this paper.

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T00:03:53.609175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:03:53.609175Z digest=sha256:06df7b4e43eb5c1fa8699031ae832d5362fe326eaf8325fa0d8f3e15d1cb2302

Observation bd36e067-e400-4557-8b59-772d7fa4143f · inbound

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning cites this paper.

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T18:12:32.839681Z digest=sha256:a6a75c421354a276289c2cf7390d5c5bd35c61da75e55168e167f495ea7a3f07

Observation 2398f91e-7a7c-475e-b2aa-dddabd2ad32a · inbound

BankerToolBench: Evaluating AI Agents in End-to-End Investment Banking Workflows cites this paper.

BankerToolBench: Evaluating AI Agents in End-to-End Investment Banking Workflows Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:55:49.453068Z digest=sha256:f87a92dffb099e4991487b790533bdee75989fcaa2b0ba4128a0daf69a796b28

Observation b9ac097e-0099-41dd-a033-d643eff7cd8a · inbound

BayMOTH: Bayesian optiMizatiOn with meTa-lookahead -- a simple approacH cites this paper.

BayMOTH: Bayesian optiMizatiOn with meTa-lookahead -- a simple approacH Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:49:27.948704Z digest=sha256:d0f939d4479c994b6ae2f9dd8f01d61028bcaf9d415ea36f50de405e4cb5f586

Observation 2845ed5e-633b-4347-9ec5-9ea60ea8840f · inbound

Whole-Body Mobile Manipulation using Offline Reinforcement Learning on Sub-optimal Controllers cites this paper.

Whole-Body Mobile Manipulation using Offline Reinforcement Learning on Sub-optimal Controllers Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:06:38.569059Z digest=sha256:81173660f51110b695f7ab9da6c51e5bb88f06f9387fb7ad505f89bd88a1c971

Observation bd406ca3-6f37-4cb6-b64e-3956c9fd2229 · inbound

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation cites this paper.

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:12:06.985609Z digest=sha256:cdb667c0aeb323e90256ce29c25e730cfcb46862c1191f472f75e85c43e78044

Observation 7bcef570-679b-42fc-8b85-f4a8fb4c964a · inbound

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation cites this paper.

Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-11T01:04:12.454268Z digest=sha256:3a9f4ea8a44920cb09e5556110acafe2bcd2a77cbf7509fa70a6596f60541945

Observation 50fcafeb-21df-4501-bd37-5801302defcd · inbound

Adaptive Memory Crystallization for Autonomous AI Agent Learning in Dynamic Environments cites this paper.

Adaptive Memory Crystallization for Autonomous AI Agent Learning in Dynamic Environments Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-13T20:58:15.877238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T20:56:17.879349Z digest=sha256:aed2d0c4cfe6d299233340e70a9ca14ebcf35e3433fe7a4edcaf6d14aee04866

Observation 9375ceac-391d-4cb0-ba94-d0dd12ee133c · inbound

Fisher Decorator: Refining Flow Policy via a Local Transport Map cites this paper.

Fisher Decorator: Refining Flow Policy via a Local Transport Map Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T05:28:12.298066Z digest=sha256:8e9819cf5d768606052af12be4be1cf9de21eb8108843dec1e82573b8ca634dc

Observation c20328f3-1783-4651-a5d7-a6900bf00502 · inbound

Distributional Off-Policy Evaluation with Deep Quantile Process Regression cites this paper.

Distributional Off-Policy Evaluation with Deep Quantile Process Regression Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 139

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T12:06:03.864434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T04:10:14.476158Z digest=sha256:d1160181248d2c6a5ac77cc04d69030607caf60fcbb1a6786fce2cdc20dfe14a

Observation ee3a9c02-7553-43f4-a2f1-5c417f1240b8 · inbound

Locality, Not Spectral Mixing, Governs Direct Propagation in Distributed Offline Dynamic Programming cites this paper.

Locality, Not Spectral Mixing, Governs Direct Propagation in Distributed Offline Dynamic Programming Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T10:15:19.300542Z digest=sha256:6719d8801c5877ddab3571b103a648eeee14191d2130862d635f79646903af32

Observation 66be2b28-c73f-4209-96a6-1b8f621754c9 · inbound

Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation cites this paper.

Beyond Bellman: High-Order Generator Regression for Continuous-Time Policy Evaluation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T12:56:20.524933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T02:32:58.434291Z digest=sha256:87b5d543f9cbe967c096cd447f9a47366452a184c8fb7dcbdbc8a5d1e2068484

Observation 622e546c-bee0-4bec-abf2-27c776b0011e · inbound

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation cites this paper.

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-11T12:31:06.596604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T03:22:30.012610Z digest=sha256:b56e38952562bb4fde8f05d66c75283a45041a4cfe37d85640e1803a07324eec

Observation 6b03fbc8-db1d-47de-b3e7-a55a697ffad1 · inbound

Align Generative Artificial Intelligence with Human Preferences: A Novel Large Language Model Fine-Tuning Method for Online Review Management cites this paper.

Align Generative Artificial Intelligence with Human Preferences: A Novel Large Language Model Fine-Tuning Method for Online Review Management Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T22:40:31.084594Z digest=sha256:718e0084260d6b2c5471f2bdcefff94c0acb3cf5a9b6df03011956caaa4a4f38

Observation 677ab89e-82ad-4965-934c-1321fc8129f6 · inbound

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning cites this paper.

SOLAR-RL: Semi-Online Long-horizon Assignment Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:21:08.858283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T12:09:24.371878Z digest=sha256:11d5273e90f8c9ea4c3cc7e39186a12b6d9f6ef913662d5cdd12b3efa8564b45

Observation 1f4eecc5-fb15-4e2d-8f85-64cc30650c2a · inbound

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems cites this paper.

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-11T20:11:09.842470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T10:08:10.495654Z digest=sha256:d3c4f4d9e25ec7a3adebbf0c3c121770d6e94a2a808678b5b4fa573b6bee9f98

Observation ba714bad-9612-4ec0-827e-fdccd0466a5e · inbound

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning cites this paper.

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T20:36:10.109937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T08:30:00.096232Z digest=sha256:2fe82953d726df95883c4662fd9a708b0ea5b575ae80f1f13517822c5515baaf

Observation bbee5401-b5b1-464d-8f16-28149a27c026 · inbound

TSN-Affinity: Similarity-Driven Parameter Reuse for Continual Offline Reinforcement Learning cites this paper.

TSN-Affinity: Similarity-Driven Parameter Reuse for Continual Offline Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-11T23:36:25.858078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-07T16:41:13.286470Z digest=sha256:b96ae4d8ff739c3a1d934441cbd66951d1dded3da5ff36d0d6f36ef79c2cb889

Observation f21a88ff-207b-4f21-8da6-46166ad9f763 · inbound

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift cites this paper.

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-12T09:41:25.829369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T10:04:49.055655Z digest=sha256:383da02731b4d79ffc2679225c4fc56b5b73d01851c690eea14e517cf5fe31d5

Observation e2f0a9b4-5f25-4933-881b-531dfc2106d7 · inbound

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data cites this paper.

Model-Based Proactive Cost Generation for Learning Safe Policies Offline with Limited Violation Data Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-05-11T16:56:07.792990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T14:24:58.919333Z digest=sha256:62dca99c86f91977da6dc9504e3df67c8af693d72eea9fa4aedfa61d00d91424

Observation f4588942-99a4-4f20-8cbc-39a855788f11 · inbound

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making cites this paper.

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T16:56:07.294277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T14:29:18.385955Z digest=sha256:7153e2f8b226a1ba88e0752de66b47f6b64f4dae339c20b80ded0b3f9ca08ade

Observation 2341d5e7-2a7a-472f-b9dc-acb2830fea68 · inbound

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making cites this paper.

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-13T07:27:29.609753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T07:26:08.192587Z digest=sha256:93a9f148433a2fc6122b0729df2ef1ce387c38e5354dc6692d3a1961124fe3a7

Observation f4330bf2-3444-4119-9f6e-e87a1b28f2a1 · inbound

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making cites this paper.

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:29:28.814472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T21:28:28.236225Z digest=sha256:070ebab83e72430a16599b072dcdc0f67ffaeabe2d2cb0e95107539d6b166e1f

Observation 77ff837d-8b11-4acd-a0e9-0c51d7ed9254 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:87bd98b969b7f6ba5d0c4f87c5182c2d254c16fcd49e7de8747d0a37a7619feb

Observation 8870bb17-d89e-4d6d-8ab0-8a1db1903b7d · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-01T00:45:11.403899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:16ddc7e6bf887732d9cf210af57b3dbcfa8808d1caa5401bc0ae67109c087e09

Observation 3158d570-c503-4529-9b78-924a554035dc · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:975be88401f7bec90a10b1e38f1388731b7476ef54c102c12a67799cfdafe39a

Observation e9dfcd8b-277f-4cab-a545-739245b88960 · inbound

On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization cites this paper.

On the Optimal Sample Complexity of Offline Multi-Armed Bandits with KL Regularization Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T19:35:59.832868Z digest=sha256:43d6cffd06f002fe3ab8e044a63041ff9f2d354928af74eff6182f1b27c5584d

Observation 21819124-22b2-4451-acd5-dbfc9adc2307 · inbound

Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability cites this paper.

Recurrent Deep Reinforcement Learning for Chemotherapy Control under Partial Observability Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:33:22.706378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T18:51:16.753971Z digest=sha256:7de1b697ccea76e917a7ee4a341d05dd72affa63f0bdde190ce9a9ae5f289b16

Observation 80227e65-b65d-4b0e-b026-5b6025e4d9d6 · inbound

Adaptive Estimation and Optimal Control in Offline Contextual MDPs without Stationarity cites this paper.

Adaptive Estimation and Optimal Control in Offline Contextual MDPs without Stationarity Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 256

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T08:56:26.425783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-07T13:27:05.081050Z digest=sha256:908bd82572845c22a3ad91522787cfe5ebbaf521456cc5f1975e4100eb4797ff

Observation 02983fbb-34ad-48f1-9e28-c96323e7b9f0 · inbound

Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning cites this paper.

Adaptive Policy Selection and Fine-Tuning under Interaction Budgets for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-11T17:41:09.203353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T17:17:14.300877Z digest=sha256:229b5205902164d7ad3ddb3e967e35b0cfcbb366e2e91758f2b96355533f609b

Observation 3b597498-b9ec-4847-87a9-2019288c8b0e · inbound

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning cites this paper.

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 68

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T18:11:05.093888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T16:34:18.603134Z digest=sha256:446ee9c9823e297212c4e0715a53616ab48fa7c762f3dcaf47b9951f817a595a

Observation e7b46636-abdd-4eff-9a6b-205ad4755113 · inbound

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning cites this paper.

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:25:07.019630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-30T23:24:57.811919Z digest=sha256:755926f1ebda1dfb3fba9212db2f29f3f76fb44dbad9e62d13b53aca891851ee

Observation b4d3f5ce-9d80-4686-83d3-f6159a4ab8cf · inbound

Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning cites this paper.

Adaptive Q-Chunking for Offline-to-Online Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-11T18:36:07.666554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T15:09:00.758071Z digest=sha256:41956d9e372c17a77ec0d68a6cef9ec8cbfd6793d1158802c5b7c8fe0c0d6f5c

Observation b2edbc2e-35c9-48d9-a424-0f9c8cb5cdc5 · inbound

On the Role of Language Representations in Auto-Bidding: Findings and Implications cites this paper.

On the Role of Language Representations in Auto-Bidding: Findings and Implications Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-11T19:41:09.090854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T11:17:01.713604Z digest=sha256:cf2628fdb471b74c2960bd2e758d522b87e1d1732818e4c7f445049f4b8ca2a1