Pith. sign in

Paper Citation Record · LEDGER

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning

As of 18 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 1 inbound Pith citation observation for arXiv:2505.13144.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.13144 v1

Coverage vector

measured 92 of 92 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:25:09.689761Z

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T16:25:25.739019Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:56:00.573367Z

Reference resolution

92 of 92 outbound references displayed

  • verified exact0
  • verified fuzzy71
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e1cd538c-c209-40fe-be5a-228d999db8a0 · outbound

This paper cites Deep reinforcement learning at the edge of the statistical precipice.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Deep reinforcement learning at the edge of the statistical precipice

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.207105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.207105Z digest=sha256:cd68d7c22b027f1b480e9bb698b40be29b35e38afdbc90cb1c34176a4e40fb76

Observation b79eac3d-6c2b-4f5d-ba00-1a3c7534befc · outbound

This paper cites OPAL : Offline primitive discovery for accelerating offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OPAL : Offline primitive discovery for accelerating offline reinforcement learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.214646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.214646Z digest=sha256:dca12349fe762c328f1e6f18a12a59939fa887dc5733fcdbd6137dcbc734705b

Observation 3dad8488-e787-4b74-9043-d92174afae52 · outbound

This paper cites Learning M arkov state abstractions for deep reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning M arkov state abstractions for deep reinforcement learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.220642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.220642Z digest=sha256:f0e28e5ddda5fdbc0728b0305e34299d51949e032f1db77d8f94c435a4ff39e8

Observation 5304bbd2-c51b-440b-85ec-9c549d361fc5 · outbound

This paper cites Uncertainty-based offline reinforcement learning with diversified Q -ensemble.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified Q -ensemble

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.228809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.228809Z digest=sha256:96d29020c7e712b10bb0e99bf40b45a9b537d0bc7c3d1e67adeccc80ad496f03

Observation 08c05610-018a-4a1a-baab-68c7091a5829 · outbound

This paper cites Hindsight experience replay.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Hindsight experience replay

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.233815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.233815Z digest=sha256:9b1b12d400850d1331b983285d738a79a024d5bdd63268f3d1b2682c030c83fd

Observation 84f33a4b-dd7e-489e-848d-9c26c8ce5226 · outbound

This paper cites and Arnold, G.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Arnold, G

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.238908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.238908Z digest=sha256:0344234b34c4c877f3fdcd5ebeef67f6df35c7069f75299f69e58341ec5b8768

Observation 4328227d-61a6-456e-83dc-e3caa7bee374 · outbound

This paper cites Autoencoders.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Autoencoders

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.244693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.244693Z digest=sha256:35c3a1f94675e44584618f78410e2ac7944d4895ff8f7a828ede1c2a3cf362b1

Observation 08880532-99c2-43f8-abec-6a61ee1ae7a6 · outbound

This paper cites Successor features for transfer in reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Successor features for transfer in reinforcement learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.249490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.249490Z digest=sha256:c42f5f6c06202817aa211ac0a7bfe37a18838452e6ee3868f0f249c2d35e59e8

Observation 5a4ad987-9bec-41f6-be37-c0f4d2084723 · outbound

This paper cites OpenAI Gym.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OpenAI Gym

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.254310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.254310Z digest=sha256:d0ea9f68557018d0801204222e7cfb33d74595202a0d9f6143f25c4f7bfa68dc

Observation 10ecf7ad-651b-4486-9f70-b888bef0d872 · outbound

This paper cites Improving generalization for temporal difference learning: The successor representation.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Improving generalization for temporal difference learning: The successor representation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.261198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.261198Z digest=sha256:4032163aae2b1701c7c63c51c0d07596baac6865602361dc879dd83f239842cd

Observation f155b088-cc16-470e-947c-26f32ca6bb14 · outbound

This paper cites and Hazan, E.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Hazan, E

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.266476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.266476Z digest=sha256:76a6d563a8ba7a86fcc0dba15aedfae7ccab30249365bc0f5f9f1fadbf3705bb

Observation 29ce66af-dc7e-4063-899b-2b840df8c1cd · outbound

This paper cites an unresolved cited work.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.273044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.273044Z digest=sha256:ed57349f374047274deb9fdf70d1ce5fd27c2c77fe27182e1e765f7fd1bbd4b1

Observation 64b310e2-6840-47f6-b5b7-5828a548c520 · outbound

This paper cites The impact of dataset on offline reinforcement learning performance in uav-based emergency network recovery tasks.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning The impact of dataset on offline reinforcement learning performance in uav-based emergency network recovery tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.278593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.278593Z digest=sha256:9eb7bdfa7eb720b7de676d1e4ddd48cb0af517ddcc1bcdbd97cc88eba2f42891

Observation 0798deed-8bfe-4336-b17c-d3d141f73833 · outbound

This paper cites IMPALA : Scalable distributed deep- RL with importance weighted actor-learner architectures.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning IMPALA : Scalable distributed deep- RL with importance weighted actor-learner architectures

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.290043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.283840Z digest=sha256:0f271f9a6ceee929af458c8660188ffd6df6f3825097a353adf980a7de7207ad

Observation 7bd5089f-00f2-4d96-8777-847416a00f43 · outbound

This paper cites Bisimulation makes analogies in goal-conditioned reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Bisimulation makes analogies in goal-conditioned reinforcement learning

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.272971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.289412Z digest=sha256:93bae7bc9f73ccc699a05775f8f68f32ac627806b543c1ecb3938f5b4fb0d9f6

Observation f3286861-4c10-446b-8946-620e10a000e9 · outbound

This paper cites C-learning: Learning to achieve goals via recursive classification.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning C-learning: Learning to achieve goals via recursive classification

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.256448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.294951Z digest=sha256:cfa2f6d1f3f96c1b4578f55f485d26e93a2c364596520f891077c39636ae0ee7

Observation f6f3133c-2e10-4142-bae9-eea29404cf6a · outbound

This paper cites Contrastive learning as goal-conditioned reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive learning as goal-conditioned reinforcement learning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.239523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.300375Z digest=sha256:5937b7f825f924590a347d9df11ff5608348e65275b5481271eee9b9c2919ce0

Observation d05173ae-a940-4c91-896b-cc4dfb9172d3 · outbound

This paper cites D4RL: Datasets for Deep Data-Driven Reinforcement Learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.305951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.305951Z digest=sha256:7232fe2b99a1b0b36574fe26af6257e8b01f07aa38b3ed7a521047ca66df50e5

Observation 1c224caa-e7e8-4c81-9241-cd413759f639 · outbound

This paper cites and Gu, S.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Gu, S

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.223110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.311626Z digest=sha256:618eabeb4d0de1529f12638ba9e39bf7f372d5829ef9d73ed351a4f701468bb9

Observation 4ba86f66-6857-4c9f-8b92-c3f4f2617922 · outbound

This paper cites For SALE : State-action representation learning for deep reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning For SALE : State-action representation learning for deep reinforcement learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.205589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.316672Z digest=sha256:1e826a2f70e34a0bccde973810dd696d93e6c31461ae4831a5fd19d1c259ca23

Observation c1401b15-946f-4fa8-98d3-6d0b73b0a7e9 · outbound

This paper cites Learning to reach goals via iterated supervised learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning to reach goals via iterated supervised learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.188393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.321656Z digest=sha256:7c67f5349a08c61b3d1005edbd1434f9921ea1b0a96d5fe54e7c1f08ff2bcfc7

Observation ba634ecf-1b41-4b7a-96c2-0b67ced5571a · outbound

This paper cites Reinforcement learning from passive data via latent intentions.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Reinforcement learning from passive data via latent intentions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.170987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.326311Z digest=sha256:d37cec4abeb7a709cd758365f35ddb61105273c3b5296b79024d7ea19f13ce66

Observation 3d12a7d5-f62c-49cf-a560-d5661cc6dc39 · outbound

This paper cites Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.151011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.331080Z digest=sha256:1f0b961ac41dc28a65d35ed5a36a28a4d3618f58e4e480bb84ff5120b85770d2

Observation 8797393f-891c-4bca-96e3-34915fd994c4 · outbound

This paper cites Learning latent dynamics for planning from pixels.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning latent dynamics for planning from pixels

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.133992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.336160Z digest=sha256:a2e19db043812d66723dfa14d3d46b629b8784648ae0cc969206964876b71957

Observation 76f80694-f9f3-47b7-88de-4fb614c190c2 · outbound

This paper cites Distance weighted supervised learning for offline interaction data.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Distance weighted supervised learning for offline interaction data

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.117782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.340972Z digest=sha256:80043651482e242e153c052d2dffb090a02ec56b8a06248a89a3754e80674796

Observation 61f2304e-9073-463c-b304-9c429192e187 · outbound

This paper cites Efficient planning in a compact latent action space.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Efficient planning in a compact latent action space

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.101452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.346494Z digest=sha256:e1c68d4c73613e3d93477dbc8fb65623520f583a919ecbe1e26ee3ed42d691df

Observation 5669dc45-8d5f-4b47-9024-3a891a0d83ab · outbound

This paper cites Learning to achieve goals.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning to achieve goals

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.083605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.351267Z digest=sha256:b608c4e5a9acce2e8acc1d71821d6d280ea3a83447da2ad028e5783685611aa4

Observation f925b495-dafd-4217-80bc-71eb9d601531 · outbound

This paper cites MOReL : Model-based offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOReL : Model-based offline reinforcement learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.063662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.357567Z digest=sha256:0b55e163a302b8d9f39acca38e6e099ea6870a4556cc3088b49f3c3bc11b93de

Observation 63044ad9-abcf-453c-84df-cb066564d8f9 · outbound

This paper cites Auto-Encoding Variational Bayes.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Auto-Encoding Variational Bayes

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.362820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.362820Z digest=sha256:2ec7e5ead778a4bd717977c17a569a218e368adff9fa6337c5968870d4dfc5e2

Observation e3f3aacd-faf8-4a77-9d15-98b5c0490be4 · outbound

This paper cites Offline reinforcement learning with implicit q -learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline reinforcement learning with implicit q -learning

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.043843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.368787Z digest=sha256:c5eeaac9175e090641d4ee06499abd18af5dd8d1e4efdc2a31d4f45635765b5f

Observation b769f48f-2056-4a2d-bc0a-89ba07505f00 · outbound

This paper cites Stabilizing off-policy Q -learning via bootstrapping error reduction.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Stabilizing off-policy Q -learning via bootstrapping error reduction

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.026947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.374051Z digest=sha256:66b16d9031e2d232954323842fc35c3860216b2365147532f173d6b06ac4bd47

Observation 2badb7e8-344e-4a30-913e-f3decc7c16e6 · outbound

This paper cites Conservative Q -learning for offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Conservative Q -learning for offline reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:11.007839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.379740Z digest=sha256:0eeaf93fc2095284b071930aba884fec27c7cf6b7f1005ba68667796785003ef

Observation 9775d383-aeb1-4236-b82c-cfc2e62bf589 · outbound

This paper cites CURL : Contrastive unsupervised representations for reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.991386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.384613Z digest=sha256:3b92f2a0cc86ff8c85f34da1b906cf70d1e0093b4acb996ebc55e83e7d80d328

Observation 1b5aba79-32c6-42c8-b796-f408218bfe0d · outbound

This paper cites Lipschitz lifelong reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Lipschitz lifelong reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.973917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.389920Z digest=sha256:add9064a4ea6a9d295efbd9d948d58ba3c5304b0b495ad73bde936143e9a6e9f

Observation 3f77e42c-9241-4dff-8790-5e94b02f47ed · outbound

This paper cites Representation balancing offline model-based reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Representation balancing offline model-based reinforcement learning

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.953244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.395094Z digest=sha256:fd1fb83581661641a76bb34660316e403a6022d44e50da19e6c1e1f3e98c04a2

Observation 166f8952-892a-4551-8df9-769fb8291620 · outbound

This paper cites and Kwon, M.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Kwon, M

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.934775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.400321Z digest=sha256:9adf7d93fb4f7be51ae00072bbdc7bf8ae40c369ccd24316bc9203d7fa271407

Observation 85585392-8a0c-4e56-af50-4b7fe8d034f8 · outbound

This paper cites and Kwon, M.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Kwon, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.916390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.405947Z digest=sha256:5e313002f85885291ec2595fcda2381c3929c172d8680d2a45923e0ce021e03d

Observation 49797fa3-f996-4b15-a343-271f9087ddba · outbound

This paper cites AD4RL : Autonomous driving benchmarks for offline reinforcement learning with value-based dataset.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning AD4RL : Autonomous driving benchmarks for offline reinforcement learning with value-based dataset

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.899118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.411955Z digest=sha256:d7e08fcb8262ca0bfcbc31b13a7be842b593905dfb994829d13e3f2deabac182

Observation 8ffa6316-3475-42b9-8c05-c52d5089d772 · outbound

This paper cites K., Choi, W., and Woo, H.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning K., Choi, W., and Woo, H

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.878724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.416818Z digest=sha256:df47b996ca0d8d3408f8658522cf54c2cdef46bf250a27cc8e562f37e4bd631a

Observation c67fac9f-3df7-42c5-8d0f-c0215ce0a01a · outbound

This paper cites GTA : Generative trajectory augmentation with guidance for offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning GTA : Generative trajectory augmentation with guidance for offline reinforcement learning

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.857609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.421852Z digest=sha256:815d7304b57d2f7aebc7d3eb940c00b735f315d895b2370555b37940c8022d59

Observation 1987829c-3d00-416e-9349-b6144803d731 · outbound

This paper cites Metric residual network for sample efficient goal-conditioned reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Metric residual network for sample efficient goal-conditioned reinforcement learning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.839673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.427089Z digest=sha256:b1e29ef5e57d9dcd47420f45581ba31cdba7e9e1832399c999dc5001a5e2a028

Observation b52a406f-a2b8-4d15-bc66-eadd68548f75 · outbound

This paper cites Synthetic experience replay.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Synthetic experience replay

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.816339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.431842Z digest=sha256:87254b1bb3363d8922ee626c237c186a36fc7076e688fd0a51aa0ab4b69551d6

Observation 48346471-2cce-4c8e-8b53-aa72fb5b9f71 · outbound

This paper cites Conservative offline distributional reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Conservative offline distributional reinforcement learning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.800647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.436373Z digest=sha256:7bead9d32ec320c5e21c31241f3757f07ba091a724746c4d8eb8df58ac34350c

Observation 93fbf303-4e4d-4c3f-b341-64f181ff34db · outbound

This paper cites VIP : Towards universal visual reward and representation via value-implicit pre-training.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning VIP : Towards universal visual reward and representation via value-implicit pre-training

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.783509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.441535Z digest=sha256:9459cab0af43e157292ce0344674ae617f7b9d6bbbe7010232986b21c4a8c550

Observation 1b143d75-7add-48a6-bbc7-c0c9bf54e49d · outbound

This paper cites Contrastive value learning: Implicit models for simple offline RL.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive value learning: Implicit models for simple offline RL

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.767068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.446491Z digest=sha256:c903f9b90392f12241ce34ed471461859cb5da681416b03fb54bb2b78c869e0a

Observation e3843d57-d0c0-492a-a723-ac6b2d368148 · outbound

This paper cites CALVIN : A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CALVIN : A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.750532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.451310Z digest=sha256:4958897f50c577b24ead4368ae5754286d7c948012710158e0be4739ad622d46

Observation 98b2a25e-e1a5-4d2d-b190-cfab05ad6540 · outbound

This paper cites Discovering and achieving goals via world models.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Discovering and achieving goals via world models

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.733769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.456421Z digest=sha256:2e81ad48ed0c0ce029fd13734dbd537a32261a87af980223a9692ed9e1d3a5b4

Observation b0db7f68-7d60-4093-9226-bdb23bfc887a · outbound

This paper cites Offline meta-reinforcement learning with advantage weighting.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline meta-reinforcement learning with advantage weighting

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.717358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.461662Z digest=sha256:74ad9c95fed6b199ab78b355fcebacaf09f4053d55f43ed051b33037dd8ef8cd

Observation 3dd4bf2b-2620-45bb-945c-bd4216d3347f · outbound

This paper cites Human-level control through deep reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Human-level control through deep reinforcement learning

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.700490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.467043Z digest=sha256:2a0c5088575a0e68b927424245e7491b9bf0fc26decb27e943a93187105684c0

Observation 65c45e28-d60c-4698-92dc-4c740200a872 · outbound

This paper cites Learning temporal distances: Contrastive successor features can provide a metric structure for decision-making.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning temporal distances: Contrastive successor features can provide a metric structure for decision-making

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.682967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.472047Z digest=sha256:56baefbc992b4723e5d7e9a42fcc1f9671ee58b7aea7cbbf81fbc52f85b75d44

Observation 545ca6e0-e0cf-456e-8d4c-1a5982a732a7 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.477942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.477942Z digest=sha256:1f0fd89cd07f5b3edc9a350d3cd234787207c352105fc640082c6d83d8a62b35

Observation 379f56dd-9a02-46b9-948b-838d6f41623a · outbound

This paper cites Planning with goal-conditioned policies.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Planning with goal-conditioned policies

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.666063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.483589Z digest=sha256:deafdc50b883645ad4dee5d8fbabcb891235f51fb92aff6f1e76a0370c277823

Observation 3fee0249-592f-401e-aaa1-967076e9736c · outbound

This paper cites Geometric autoencoders--what you see is what you decode.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Geometric autoencoders--what you see is what you decode

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.649145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.488674Z digest=sha256:82722ae09f19d79a51dbc9a7eb19328c6f71dc731338ba44c0dd5df32fca2381

Observation d61df6c0-38a8-46db-a6b2-7af3428c78a3 · outbound

This paper cites and Powell, J.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Powell, J

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.632279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.493642Z digest=sha256:bc33210c17d98d4904feea102eeeec9356981416929379ce962215afbce84062

Observation eaf3475c-ff04-4794-9805-c9aa42404d85 · outbound

This paper cites HIQL : Offline goal-conditioned RL with latent states as actions.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning HIQL : Offline goal-conditioned RL with latent states as actions

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.616096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.498842Z digest=sha256:4ef69659a307a6d037e0b1b2ce08f2193a41879d117fc4667a6eb45d6a901001

Observation 8ed43098-98aa-4ce3-8e5f-fb4538020964 · outbound

This paper cites Foundation policies with H ilbert representations.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Foundation policies with H ilbert representations

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.599071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.504073Z digest=sha256:eafe984c879946a276c0d806c8cc7cf72b6b5bc1c1d75156ec611625e48fd126

Observation 07489cba-ed33-49e7-819e-d26c40eff11b · outbound

This paper cites Long-horizon visual planning with goal-conditioned hierarchical predictors.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Long-horizon visual planning with goal-conditioned hierarchical predictors

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.581846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.509257Z digest=sha256:fdba3a85a059176668d6634160ad2578cae0130c9b5da15c462e515cecc0049d

Observation e257e055-6e27-4205-9fde-48a1b24847a9 · outbound

This paper cites and Juditsky, A.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Juditsky, A

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.563193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.514130Z digest=sha256:d9b612c5640efa0fcfb0062c478c2b90d6a7abcad64045b012cb117e0571fcdf

Observation 09968429-d669-4a51-a08e-57f550de8483 · outbound

This paper cites Temporal difference models: Model-free deep RL for model-based control.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Temporal difference models: Model-free deep RL for model-based control

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.545500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.518915Z digest=sha256:54a97fcf34e3dde96563af8dbce21c78db472714ebc488677bc3f2de006dc658

Observation 501f3a5a-301f-42b2-b265-bb6bd2ce85aa · outbound

This paper cites MOTO : Offline pre-training to online fine-tuning for model-based robot learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOTO : Offline pre-training to online fine-tuning for model-based robot learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.528766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.523633Z digest=sha256:a2d27ea119a36a59bf2f5f740f9cb68f2557c192ceb609c9c60a59c06804e2e7

Observation 75e608e3-de1b-40d5-8a98-3f3ef04757d2 · outbound

This paper cites Goal-conditioned offline reinforcement learning via metric learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Goal-conditioned offline reinforcement learning via metric learning

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.528322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.528322Z digest=sha256:0913420daec76ea02c2093eb7053ab4a55c0923cbfbfd68e229915b5a1a07976

Observation 11a93f55-1063-4b51-b018-0147d8c7307f · outbound

This paper cites RAMBO-RL : Robust adversarial model-based offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning RAMBO-RL : Robust adversarial model-based offline reinforcement learning

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.512174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.533157Z digest=sha256:8b35b0c8b35e069d49f97676741ddefd4f85bc80dda3df6c2654220fb62f1a62

Observation f6e61a0a-aea5-4b8e-ac32-5eb4289fc6ad · outbound

This paper cites An overview of gradient descent optimization algorithms.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning An overview of gradient descent optimization algorithms

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.537802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.537802Z digest=sha256:39901e14dfdf6603caaac4e8f26b7d0805a9c9259c8f58064af64d264b142c66

Observation 7695a798-a148-4099-930f-95a27f05bffd · outbound

This paper cites Universal value function approximators.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Universal value function approximators

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.495276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.542684Z digest=sha256:7cf0c4299a459e500516d799b09afc845b9afd4a845ffb751fca77f5b79a00e4

Observation 4ebb8f12-af3b-4ec8-b530-eb1872c3b133 · outbound

This paper cites Reinforcement learning with action-free pre-training from videos.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Reinforcement learning with action-free pre-training from videos

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.476581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.547429Z digest=sha256:5ffaea36b9145294c162c4ddd194c09e38ee9b0eef5b5f8c0d9afb561e25c792

Observation d3ea083e-aabd-4351-806e-c57a9085a037 · outbound

This paper cites Skill-based model-based reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Skill-based model-based reinforcement learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.458026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.552038Z digest=sha256:9b2a7eca82b8bee7111717e2995ff9ab4866381d915b6e08924823bd876f43b5

Observation d1bfc6ce-35c8-4136-8805-abbc65109a81 · outbound

This paper cites K., and Woo, H.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning K., and Woo, H

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.440697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.557425Z digest=sha256:4a594d8a627c5e9d3e51726d51a4552cc8604ca23ab9ce202548f7e6f453513c

Observation 9bf83f89-1d6d-4ae1-8d8f-e57049288fa8 · outbound

This paper cites S4RL : Surprisingly simple self-supervision for offline reinforcement learning in robotics.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning S4RL : Surprisingly simple self-supervision for offline reinforcement learning in robotics

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.421547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.562414Z digest=sha256:75ff70191801237e40ef592a61c1f28cb91a86326325a184d930a2ad46ebf56a

Observation 04bf1025-d401-4cdc-bdea-3af053d7825b · outbound

This paper cites Offline RL for natural language generation with implicit language Q learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline RL for natural language generation with implicit language Q learning

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.400577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.568039Z digest=sha256:9e14d3e28429a875880902d5fa4abd38104047add0148945345d0c8382af8ce3

Observation c065a14c-3941-4103-893d-eecd266b00f6 · outbound

This paper cites Intrinsic motivation and automatic curricula via asymmetric self-play.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Intrinsic motivation and automatic curricula via asymmetric self-play

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.381987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.573333Z digest=sha256:f27570dbea939957b979e79385bc0250e44de3ef1fe9500a6aeb487b07162d15

Observation e6a33efc-f50e-4827-bd30-771b9b32f575 · outbound

This paper cites Model- B ellman inconsistency for model-based offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Model- B ellman inconsistency for model-based offline reinforcement learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.359352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.578993Z digest=sha256:02322a3a11538c426ed112b3e2256300006739d416c5a2948834b1ee4099de93

Observation 60834f7b-27a5-489c-8e87-fe12b6448648 · outbound

This paper cites Leveraging factored action spaces for efficient offline reinforcement learning in healthcare.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Leveraging factored action spaces for efficient offline reinforcement learning in healthcare

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.341074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.584308Z digest=sha256:970fe2a63d3f644fa9dedcbd13e1e3fdb00c0ae05d46a05fa196453d03622409

Observation 9c3bf189-2b1e-498e-a47a-a193da26113c · outbound

This paper cites Revisiting the minimalist approach to offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Revisiting the minimalist approach to offline reinforcement learning

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.323486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.590352Z digest=sha256:416d9c8417303c5d1ae4bdda4be9e6ec145b8982129f7db06715718579721013

Observation e4dd58b2-71ad-426d-a282-a339a3f5cc27 · outbound

This paper cites CORL : Research-oriented deep offline reinforcement learning library.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CORL : Research-oriented deep offline reinforcement learning library

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.302213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.595661Z digest=sha256:7007c00c3546e615e0a2d3038f2904e5a6853455063a75dece16f8263b1e4903

Observation e34ed147-2a30-4712-9fd6-4f6723d47e53 · outbound

This paper cites and Mannor, S.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Mannor, S

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.284608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.600941Z digest=sha256:a75a31e1f4a8bbc6044d660d8f40bc65a6990e1e79f3dca58d346631a8d84b4e

Observation 064133d3-37be-42b8-ba84-4e50f3938e4e · outbound

This paper cites Offline reinforcement learning with reverse model-based imagination.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline reinforcement learning with reverse model-based imagination

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.263319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.606971Z digest=sha256:e9f772fe9125a1c90f1fd936b2346e6570c55605ecac6ec297d107d8a1a72200

Observation 389f57b2-b4ce-43d4-9a2f-ae5149ce1488 · outbound

This paper cites and Isola, P.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Isola, P

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.240000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.612981Z digest=sha256:771fe225f529402802f9de60de37a874577c1dca875d3067e527ab93b3577dee

Observation f3a0d81e-7b9c-4061-a12a-e24211017145 · outbound

This paper cites Optimal goal-reaching reinforcement learning via quasimetric learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Optimal goal-reaching reinforcement learning via quasimetric learning

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.222515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.618021Z digest=sha256:23946d308d2956b0d43f77cc2586c945d81884bc659f687984c3db711fa88d3e

Observation 498aac16-3c47-4786-8817-530c5f367e61 · outbound

This paper cites Critic regularized regression.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Critic regularized regression

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.203046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.622779Z digest=sha256:b4c76eda972e01c5b56c04ebad332e25c84a1b33b87f8b008c9a9dc3540ff494

Observation a9e55d8a-34d5-476f-9bd4-ef7926ae8098 · outbound

This paper cites OCEAN-MBRL : Offline conservative exploration for model-based offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OCEAN-MBRL : Offline conservative exploration for model-based offline reinforcement learning

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.184568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.627535Z digest=sha256:b58ca755b9301e5f303fcec7d53bffa6c2a99324f02b70d203652214263cf6a3

Observation ad7d5074-944c-424e-bc95-982aa4363e5c · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Behavior Regularized Offline Reinforcement Learning

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.632253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.632253Z digest=sha256:622282f5add7cc457c96330d592aa23c330dc7bfde92ccbb7e81850fb01dc5a2

Observation a62dea00-cbf1-45b7-874b-be5bf8aef162 · outbound

This paper cites A policy-guided imitation approach for offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning A policy-guided imitation approach for offline reinforcement learning

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.166948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.637365Z digest=sha256:f9ea88439f448ec4c8de8545526abaf2afb29a000f6819c21663ff6925381d73

Observation 900ecb0c-a6c7-43db-8919-70a9f11f0e77 · outbound

This paper cites an unresolved cited work.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:25:10.148776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.642253Z digest=sha256:2751e4b6b808763a2a308f590010c4b76a0180046d50079a45f4a89d1e494b70

Observation 478de198-5e9d-49f3-8fb8-2f882a3fc592 · outbound

This paper cites Image augmentation is all you need: Regularizing deep reinforcement learning from pixels.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Image augmentation is all you need: Regularizing deep reinforcement learning from pixels

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.129625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.647705Z digest=sha256:61fbeb051f72bdc42ea5302e821c9e4c29eddde4f675ef663d9ec8506f9bdd35

Observation 00f9f802-f9ba-4e02-8452-8db4eec90fee · outbound

This paper cites MOPO : Model-based offline policy optimization.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOPO : Model-based offline policy optimization

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.105493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.652509Z digest=sha256:2bcc272c270de32e4673d651a1d13ace9948282075921dfa1a4045e7acd32c69

Observation a6832146-84fb-412d-9529-9dca4c43c8dd · outbound

This paper cites COMBO : Conservative offline model-based policy optimization.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning COMBO : Conservative offline model-based policy optimization

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.083276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.658143Z digest=sha256:b0473de4eaac1ba2b1e65c96754655b78624a897a920aa463f3216bcacbcb7d1

Observation 374e67d8-7dd6-457a-af84-fb5095ac550d · outbound

This paper cites BRAC+ : Improved behavior regularized actor critic for offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning BRAC+ : Improved behavior regularized actor critic for offline reinforcement learning

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.052484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.663112Z digest=sha256:869e76bacc51e4722b8d405acf2bf9d2385e69740af1b19afecdfa88a7532f66

Observation b31c4ee2-01e6-49d9-8348-a85cf8f357ef · outbound

This paper cites Discriminator-guided model-based offline imitation learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Discriminator-guided model-based offline imitation learning

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.035057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.667972Z digest=sha256:650fd50ed53b60002a8fb46b3bb772a17169bd7b54aa4825fb243aa136b3271a

Observation b4b97e3c-5988-4b6f-8794-185d72226a91 · outbound

This paper cites Contrastive difference predictive coding.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive difference predictive coding

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:10.017605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.673237Z digest=sha256:d294ed6b1478d83c9b6d400f9cd469558d1ad4be99f3190c6524ff6934762cfa

Observation 35bf770d-1907-4a23-9dc6-6dba20c99c4a · outbound

This paper cites TACO : Temporal latent action-driven contrastive loss for visual reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning TACO : Temporal latent action-driven contrastive loss for visual reinforcement learning

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:09.999200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.678721Z digest=sha256:b37e84d835a8b4022932004b890eead1e3902eaa7c4af02d7d368fbd2280da0b

Observation b5471c6f-52fa-432a-b3b6-64483d3709a4 · outbound

This paper cites PLAS : Latent action space for offline reinforcement learning.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning PLAS : Latent action space for offline reinforcement learning

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:25:09.978520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-15T20:25:09.683966Z digest=sha256:b947c2e4f6dc483876b89cffee6ec94b9531a324f84db86099e588f5a8b94e81

Observation 28f766fc-8c4f-40f2-907a-9b22376fc385 · outbound

This paper cites write newline.

Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning write newline

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-15T20:25:09.689761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:25:09.689761Z digest=sha256:742962211f1f0062dcbc213b1f8473031912759c6cdb73f805068e1ca0d4fed2

Pith citing papers

Observation 1fec2f2a-66aa-4ca9-bcf3-9f64390f4523 · inbound

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning cites this paper.

Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:56:00.575502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T16:25:25.739019Z digest=sha256:b71b28de45a2f71771b8eb62b0f2b139ed91796d325fc2546264f53d9776b1ac