Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T14:49:29.225582Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 27 of 27 outbound references and 0 inbound Pith citation observations for arXiv:2607.23726.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-30T14:49:29.225582Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
27 of 27 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation de6624ef-83b6-4868-b8e7-69758832922e · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning doi: 10.3389/frobt.2025.1567211
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ee8f3fd-2c78-4d56-9063-4e3ead66ba83 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning ISBN 978-981-97-5035-1
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e540d67b-2188-418b-b546-5bc8c3045a16 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e82ab8b3-acef-4e7a-a1df-96add91289d3 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f2ec826-c1a8-4a9a-a0ef-73a03fbc7537 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Chao Lv, Ming Zhu, Xiao Guo, Jiajun Ou, and Wenjie Lou
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1498600a-6106-42e9-a1c1-082cc01dc09d · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning doi: https://doi.org/10.1016/j
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af259677-0dc2-4111-a467-f297a97dc073 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Andrew Y Ng, Daishi Harada, and Stuart Russell
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b60851a7-cd1b-4d9c-ba1a-f7ce0bdcb7a2 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning RescuedBy
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 74e21419-114d-420c-a697-2603ef0fcf6c · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Version 1, 1980 images, CC BY 4.0, Accessed: February 12,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5632b7b4-5eb9-45d5-8fb9-2c118d360f47 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b806cbc4-3d3e-48fd-8e0c-7d7bd88721ff · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Richard S Sutton, Doina Precup, and Satinder Singh
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 492dc1b7-eeac-4f43-8489-29515e4f5cef · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning SHIRO: Soft Hierarchical Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b060fd3a-194b-4d3a-b24f-fe91dfb76067 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Jiarui Yang, Bin Zhu, Jingjing Chen, and Yu-Gang Jiang
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f990ca5e-e7de-4e09-a792-0958609e7e83 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning 48550/arXiv.2508.11143
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0df5e0c-38d5-42ca-a02e-0e8bdbc2495f · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f024449-3b0e-4c6c-a2d0-9c05ccb418a1 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning URLhttps://doi.org/10.1016/S0004-3702(99) 00052-1
Reference 1999
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f1e540b0-6f2c-4286-9aea-4d72b2420d3d · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Botao Dong, Longyang Huang, Ning Pang, Hongtian Chen, and Weidong Zhang
Reference 2000
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a5e4796-f0dc-4f98-ba3f-bd888bf8ed04 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Mohammad Al Homsi, Maja Trumi´ c, Adriano Fagiolini, and Giansalvo Cirrincione
Reference 2015
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation baac3ab8-4536-4ff0-a891-14085507671b · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Michael Laskin, Aravind Srinivas, and Pieter Abbeel
Reference 2016
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1b7c7fb-ee14-4065-b7e3-98477741ef57 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Large-Scale Study of Curiosity-Driven Learning
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6abb979-3707-465b-b990-17cf6a5d0515 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Large-Scale Study of Curiosity-Driven Learning
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a40db98-bb95-44ea-b329-24d0d2ff932c · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Off-policy Maximum Entropy Reinforcement Learning : Soft Actor-Critic with Advantage Weighted Mixture Policy(SAC-AWMP)
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71793d86-e21e-4908-943f-7c122a9fed3b · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning 24 Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Pierre-Luc Bacon, Jean Harb, and Doina Precup
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ee3656f-18f9-4e4b-aa52-c9ed85da9786 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning doi: 10.3390/make4010009
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 982eef5a-619f-43ad-bdaa-0262b4c8fe08 · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Jiangyue Yan, Biao Luo, and Xiaodong Xu
Reference 2023
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9dbd8e67-502b-4cb4-845f-4da8c81c8a7a · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Nikolas Gegenava
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fee421ed-faa5-4ca6-9f17-76f81655ad7f · outbound
Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Mostafa Al-Emran
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
No inbound Pith citation observations are available.