Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:25:09.689761Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 1 inbound Pith citation observation for arXiv:2505.13144.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:25:09.689761Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:25:25.739019Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T08:56:00.573367Z
92 of 92 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation e1cd538c-c209-40fe-be5a-228d999db8a0 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Deep reinforcement learning at the edge of the statistical precipice
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b79eac3d-6c2b-4f5d-ba00-1a3c7534befc · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OPAL : Offline primitive discovery for accelerating offline reinforcement learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dad8488-e787-4b74-9043-d92174afae52 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning M arkov state abstractions for deep reinforcement learning
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5304bbd2-c51b-440b-85ec-9c549d361fc5 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Uncertainty-based offline reinforcement learning with diversified Q -ensemble
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08c05610-018a-4a1a-baab-68c7091a5829 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Hindsight experience replay
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f33a4b-dd7e-489e-848d-9c26c8ce5226 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Arnold, G
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4328227d-61a6-456e-83dc-e3caa7bee374 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Autoencoders
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08880532-99c2-43f8-abec-6a61ee1ae7a6 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Successor features for transfer in reinforcement learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a4ad987-9bec-41f6-be37-c0f4d2084723 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OpenAI Gym
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10ecf7ad-651b-4486-9f70-b888bef0d872 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Improving generalization for temporal difference learning: The successor representation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f155b088-cc16-470e-947c-26f32ca6bb14 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Hazan, E
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ce66af-dc7e-4063-899b-2b840df8c1cd · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64b310e2-6840-47f6-b5b7-5828a548c520 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning The impact of dataset on offline reinforcement learning performance in uav-based emergency network recovery tasks
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0798deed-8bfe-4336-b17c-d3d141f73833 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning IMPALA : Scalable distributed deep- RL with importance weighted actor-learner architectures
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7bd5089f-00f2-4d96-8777-847416a00f43 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Bisimulation makes analogies in goal-conditioned reinforcement learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f3286861-4c10-446b-8946-620e10a000e9 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning C-learning: Learning to achieve goals via recursive classification
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f6f3133c-2e10-4142-bae9-eea29404cf6a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive learning as goal-conditioned reinforcement learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d05173ae-a940-4c91-896b-cc4dfb9172d3 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning D4RL: Datasets for Deep Data-Driven Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c224caa-e7e8-4c81-9241-cd413759f639 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Gu, S
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4ba86f66-6857-4c9f-8b92-c3f4f2617922 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning For SALE : State-action representation learning for deep reinforcement learning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c1401b15-946f-4fa8-98d3-6d0b73b0a7e9 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning to reach goals via iterated supervised learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ba634ecf-1b41-4b7a-96c2-0b67ced5571a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Reinforcement learning from passive data via latent intentions
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3d12a7d5-f62c-49cf-a560-d5661cc6dc39 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Relay policy learning: Solving long-horizon tasks via imitation and reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8797393f-891c-4bca-96e3-34915fd994c4 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning latent dynamics for planning from pixels
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 76f80694-f9f3-47b7-88de-4fb614c190c2 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Distance weighted supervised learning for offline interaction data
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 61f2304e-9073-463c-b304-9c429192e187 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Efficient planning in a compact latent action space
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5669dc45-8d5f-4b47-9024-3a891a0d83ab · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning to achieve goals
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f925b495-dafd-4217-80bc-71eb9d601531 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOReL : Model-based offline reinforcement learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 63044ad9-abcf-453c-84df-cb066564d8f9 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Auto-Encoding Variational Bayes
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3f3aacd-faf8-4a77-9d15-98b5c0490be4 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline reinforcement learning with implicit q -learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b769f48f-2056-4a2d-bc0a-89ba07505f00 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Stabilizing off-policy Q -learning via bootstrapping error reduction
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2badb7e8-344e-4a30-913e-f3decc7c16e6 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Conservative Q -learning for offline reinforcement learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9775d383-aeb1-4236-b82c-cfc2e62bf589 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CURL : Contrastive unsupervised representations for reinforcement learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1b5aba79-32c6-42c8-b796-f408218bfe0d · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Lipschitz lifelong reinforcement learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3f77e42c-9241-4dff-8790-5e94b02f47ed · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Representation balancing offline model-based reinforcement learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 166f8952-892a-4551-8df9-769fb8291620 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Kwon, M
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 85585392-8a0c-4e56-af50-4b7fe8d034f8 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Kwon, M
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 49797fa3-f996-4b15-a343-271f9087ddba · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning AD4RL : Autonomous driving benchmarks for offline reinforcement learning with value-based dataset
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8ffa6316-3475-42b9-8c05-c52d5089d772 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning K., Choi, W., and Woo, H
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c67fac9f-3df7-42c5-8d0f-c0215ce0a01a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning GTA : Generative trajectory augmentation with guidance for offline reinforcement learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1987829c-3d00-416e-9349-b6144803d731 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Metric residual network for sample efficient goal-conditioned reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b52a406f-a2b8-4d15-bc66-eadd68548f75 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Synthetic experience replay
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 48346471-2cce-4c8e-8b53-aa72fb5b9f71 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Conservative offline distributional reinforcement learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 93fbf303-4e4d-4c3f-b341-64f181ff34db · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning VIP : Towards universal visual reward and representation via value-implicit pre-training
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1b143d75-7add-48a6-bbc7-c0c9bf54e49d · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive value learning: Implicit models for simple offline RL
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e3843d57-d0c0-492a-a723-ac6b2d368148 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CALVIN : A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 98b2a25e-e1a5-4d2d-b190-cfab05ad6540 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Discovering and achieving goals via world models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b0db7f68-7d60-4093-9226-bdb23bfc887a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline meta-reinforcement learning with advantage weighting
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3dd4bf2b-2620-45bb-945c-bd4216d3347f · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Human-level control through deep reinforcement learning
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 65c45e28-d60c-4698-92dc-4c740200a872 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Learning temporal distances: Contrastive successor features can provide a metric structure for decision-making
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 545ca6e0-e0cf-456e-8d4c-1a5982a732a7 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning AWAC: Accelerating Online Reinforcement Learning with Offline Datasets
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 379f56dd-9a02-46b9-948b-838d6f41623a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Planning with goal-conditioned policies
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3fee0249-592f-401e-aaa1-967076e9736c · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Geometric autoencoders--what you see is what you decode
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d61df6c0-38a8-46db-a6b2-7af3428c78a3 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Powell, J
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation eaf3475c-ff04-4794-9805-c9aa42404d85 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning HIQL : Offline goal-conditioned RL with latent states as actions
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8ed43098-98aa-4ce3-8e5f-fb4538020964 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Foundation policies with H ilbert representations
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 07489cba-ed33-49e7-819e-d26c40eff11b · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Long-horizon visual planning with goal-conditioned hierarchical predictors
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e257e055-6e27-4205-9fde-48a1b24847a9 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Juditsky, A
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 09968429-d669-4a51-a08e-57f550de8483 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Temporal difference models: Model-free deep RL for model-based control
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 501f3a5a-301f-42b2-b265-bb6bd2ce85aa · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOTO : Offline pre-training to online fine-tuning for model-based robot learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 75e608e3-de1b-40d5-8a98-3f3ef04757d2 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Goal-conditioned offline reinforcement learning via metric learning
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11a93f55-1063-4b51-b018-0147d8c7307f · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning RAMBO-RL : Robust adversarial model-based offline reinforcement learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f6e61a0a-aea5-4b8e-ac32-5eb4289fc6ad · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning An overview of gradient descent optimization algorithms
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7695a798-a148-4099-930f-95a27f05bffd · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Universal value function approximators
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4ebb8f12-af3b-4ec8-b530-eb1872c3b133 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Reinforcement learning with action-free pre-training from videos
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d3ea083e-aabd-4351-806e-c57a9085a037 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Skill-based model-based reinforcement learning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d1bfc6ce-35c8-4136-8805-abbc65109a81 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning K., and Woo, H
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9bf83f89-1d6d-4ae1-8d8f-e57049288fa8 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning S4RL : Surprisingly simple self-supervision for offline reinforcement learning in robotics
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 04bf1025-d401-4cdc-bdea-3af053d7825b · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline RL for natural language generation with implicit language Q learning
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c065a14c-3941-4103-893d-eecd266b00f6 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Intrinsic motivation and automatic curricula via asymmetric self-play
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e6a33efc-f50e-4827-bd30-771b9b32f575 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Model- B ellman inconsistency for model-based offline reinforcement learning
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 60834f7b-27a5-489c-8e87-fe12b6448648 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Leveraging factored action spaces for efficient offline reinforcement learning in healthcare
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9c3bf189-2b1e-498e-a47a-a193da26113c · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Revisiting the minimalist approach to offline reinforcement learning
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e4dd58b2-71ad-426d-a282-a339a3f5cc27 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning CORL : Research-oriented deep offline reinforcement learning library
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e34ed147-2a30-4712-9fd6-4f6723d47e53 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Mannor, S
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 064133d3-37be-42b8-ba84-4e50f3938e4e · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Offline reinforcement learning with reverse model-based imagination
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 389f57b2-b4ce-43d4-9a2f-ae5149ce1488 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning and Isola, P
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f3a0d81e-7b9c-4061-a12a-e24211017145 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Optimal goal-reaching reinforcement learning via quasimetric learning
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 498aac16-3c47-4786-8817-530c5f367e61 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Critic regularized regression
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a9e55d8a-34d5-476f-9bd4-ef7926ae8098 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning OCEAN-MBRL : Offline conservative exploration for model-based offline reinforcement learning
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ad7d5074-944c-424e-bc95-982aa4363e5c · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Behavior Regularized Offline Reinforcement Learning
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a62dea00-cbf1-45b7-874b-be5bf8aef162 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning A policy-guided imitation approach for offline reinforcement learning
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 900ecb0c-a6c7-43db-8919-70a9f11f0e77 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Unresolved cited work
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 478de198-5e9d-49f3-8fb8-2f882a3fc592 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Image augmentation is all you need: Regularizing deep reinforcement learning from pixels
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 00f9f802-f9ba-4e02-8452-8db4eec90fee · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning MOPO : Model-based offline policy optimization
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a6832146-84fb-412d-9529-9dca4c43c8dd · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning COMBO : Conservative offline model-based policy optimization
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 374e67d8-7dd6-457a-af84-fb5095ac550d · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning BRAC+ : Improved behavior regularized actor critic for offline reinforcement learning
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b31c4ee2-01e6-49d9-8348-a85cf8f357ef · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Discriminator-guided model-based offline imitation learning
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b4b97e3c-5988-4b6f-8794-185d72226a91 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning Contrastive difference predictive coding
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 35bf770d-1907-4a23-9dc6-6dba20c99c4a · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning TACO : Temporal latent action-driven contrastive loss for visual reinforcement learning
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b5471c6f-52fa-432a-b3b6-64483d3709a4 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning PLAS : Latent action space for offline reinforcement learning
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 28f766fc-8c4f-40f2-907a-9b22376fc385 · outbound
Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning write newline
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fec2f2a-66aa-4ca9-bcf3-9f64390f4523 · inbound
Towards Efficient and Expressive Offline RL via Flow-Anchored Noise-conditioned Q-Learning Temporal Distance-aware Transition Augmentation for Offline Model-based Reinforcement Learning
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.