Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-22T07:46:20.289421Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2605.22711.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-22T07:46:20.289421Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
63 of 63 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fcbf8917-a44c-45f6-a2fa-418a20312a24 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning to achieve goals
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a4a4f3d2-acd3-45ca-927b-4e69af28d3b4 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Universal value function approxi- mators
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aa4726ca-7294-4998-b9b7-664d7204e41a · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5177c369-e874-4405-b84c-90252e5d2498 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Understanding the World Through Action
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 68eaf269-05be-4142-b8dd-52ae9d8b9432 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning OGBench: Benchmarking Offline Goal-Conditioned RL
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d764c1be-3a1b-4cc9-a6ad-b995f7f8964c · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning doi:10.1109/TNNLS.2023.3250269
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 92e13170-33b8-4903-b4e4-56cf37ebfb05 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation adc95831-6a56-4005-b662-987fcfe72c55 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 961bcf7d-6779-45d9-80f6-d5ab4545ef13 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Revisiting the Minimalist Approach to Offline Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a1598660-f353-440c-975e-4a371a192175 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Challenges of Real-World Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 36c580df-c18e-4b42-a49a-79faaefd2165 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Is Value Learning Really the Main Bottleneck in Offline RL?
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5a8387a0-abaf-46c9-8338-7fe458418401 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Horizon reduction makes rl scalable
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 98361091-8d67-41e2-a9c2-67b8a67c4df9 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Sutton, Doina Precup, and Satinder Singh
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d2122a0b-5a16-4da6-8e38-eb86db5b5ed2 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Feudal networks for hierarchical reinforcement learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2f096f01-0c67-4c69-9e63-8bbb2a7c4034 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cb262baf-ff93-46a9-8c5d-7a522ca053fb · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Real-Time Execution of Action Chunking Flow Policies
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3b241c59-7c4b-4896-a468-132d688f6fb5 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Scalable Offline Model- Based RL with Action Chunks, December 2025
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 58ccb305-5dd0-4103-b2af-7882204efcbc · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Geometric Deep Learning: Grids, Groups, Graphs, Geodesics, and Gauges
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1c76bcec-6df0-4551-8b3e-3f1ad8010346 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Equivariant Goal Conditioned Contrastive Reinforcement Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 06568a53-1fe3-4df0-a7d9-a45d2c0e6d72 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Riedmiller
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 845905e7-99e0-45df-a245-2889189cf6f5 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Sutton and Andrew G
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 03c1c226-f725-477f-96d6-b44808b6c064 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Scaling Life-long Off-policy Learning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bd4ef190-0bd4-47b6-820c-7d7f670a2440 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Hindsight Experience Replay
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 967ca42f-56e3-4e18-866d-c82355edd8e0 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning HIQL: Offline Goal-Conditioned RL with Latent States as Actions
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 91af34ea-1812-483d-9e44-00cd4875c82b · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Conservative Q-Learning for Offline Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 30b9fe42-f41f-49ba-94d0-ba300624c751 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Uncertainty-Based Offline Reinforcement Learning with Diversified Q-Ensemble
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e7cd1b31-8a6e-4d93-994c-cc1191d1722c · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Contrastive Learning as Goal-Conditioned Reinforcement Learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9a684a99-7f3b-44ab-81da-afe85ad7232c · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning A Minimalist Approach to Offline Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 42acfa00-10cb-44a1-9353-437dd73ac8d8 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning EMaQ: Expected-Max Q-Learning Operator for Simple Yet Effective Offline and Online RL
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dae6d8e1-bc45-4c00-b156-9a16bb69a7a8 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning The Option-Critic Architecture, December
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a29fed8b-62ad-44f0-b7f8-c8526f5a8ec2 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning The Option-Critic Architecture
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e7505b51-cb54-4ea0-9da9-3ca2d8cd13c8 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Graph-Assisted Stitching for Offline Hierarchical Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f0a75844-4e58-44a9-a9f1-d1047259d396 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Intra-Option Learning about Temporally Abstract Actions
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5b922b75-9a58-4877-92d7-3317e80815f1 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 97af16e3-57f0-4e95-b45e-d1a0b03a4dcc · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Data-Efficient Hierarchical Reinforcement Learning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c2946719-8979-4fb7-b88f-b90275c49b94 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Near-Optimal Representation Learning for Hierarchical Reinforcement Learning
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 11ae49c6-7149-4f27-829e-509333980c88 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9c96ac77-f8ad-44e1-8d75-725914f54521 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 51774ece-c80d-4819-9210-3f6f07557c5c · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning options in reinforcement learning
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7f730393-593e-4afa-84d0-ebf6a7ffc9c2 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Hierarchical planning through goal-conditioned offline reinforcement learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation a9657b27-731f-4ea8-b007-c828a52ad564 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Towards a Unified Theory of State Abstraction for MDPs
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f24f5adc-2de7-46bd-a733-af71cf354a62 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Metrics for Finite Markov Decision Processes
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f9243d5a-e4ad-4e8d-b411-6e37128e3770 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning Representations via a Robust Behavioral Metric for Deep Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f8e3d58-0158-49b3-bc5a-a4fbcba1d4c5 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning A Survey of State Representation Learning for Deep Reinforcement Learning
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 257eea23-8077-40f4-9a73-88e3b81553c4 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Phd thesis, University College London
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2d77abf6-907e-4a7e-b205-929f0174bb9d · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Finite-Time Bounds for Fitted Value Iteration
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4f99cb7e-8d9f-4f19-8ad9-99138bde01c6 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning PAC Bounds for Discounted MDPs
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3bbee3e0-9172-4e6c-a174-a617bef7b9c7 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Sample Complexity of Goal-Conditioned Hierarchical Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 81e8a703-9183-4090-97e2-609cbb8c5686 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Transitive RL: Value Learn- ing via Divide and Conquer, February 2026
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 349592fc-9b42-48e1-93a9-cd9c859f3614 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Reinforcement Learning from Passive Data via Latent Intentions
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 586fda6b-8f06-4f7b-b297-2bb70c52b169 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning A Policy-Guided Imitation Approach for Offline Reinforcement Learning
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96923287-9fb2-4900-baf3-5f02670420b7 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Deep reinforcement learning at the edge of the statistical precipice.Advances in Neural Information Processing Systems
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 37a01df8-4a66-4638-84ea-57b387be753f · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning A Clean Slate for Offline Reinforcement Learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4f47fa5e-dbfa-4c14-ae36-95c0f7336288 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Scaling Laws for Neural Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c62e827a-694d-4048-bd2e-62e5940ab89a · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning 1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities, February 2026
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 50ac6e59-8c51-4a53-8b29-de1f7c13a18d · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Flow Matching Guide and Code
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d5be0bc9-8e4f-4ef9-947d-d6fb154911c1 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Dual Goal Representations, February
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 12253bbf-6d0b-4a49-972d-a67d5e4ecb7f · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning arXiv:2510.06714 [cs]
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 12027911-1874-436c-9969-0b495e0b58b6 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Adam: A Method for Stochastic Optimization
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7ff9557a-0121-4b94-9db2-f317ec17ab1a · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Layer Normalization
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1cb285bd-6a3c-45c0-8fbc-1bfede5f0140 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Gaussian Error Linear Units (GELUs)
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8bfb6577-6d7b-416b-8078-3e66a43ebf3e · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Addressing Optimism Bias in Sequence Modeling for Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 51449718-5ca5-4438-9cbe-3cc21615e971 · outbound
Abstraction for Offline Goal-Conditioned Reinforcement Learning Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.