Pith. sign in

Paper Citation Record · LEDGER

Learning Multi-Level Hierarchies with Hindsight

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:1712.00948.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1712.00948 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T00:28:21.497424Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T02:07:34.156817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 81199ecb-c522-41f6-b07a-18363b25401f · inbound

Learning World Graphs to Accelerate Hierarchical Reinforcement Learning cites this paper.

Learning World Graphs to Accelerate Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-25T12:35:49.133772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T12:31:38.848720Z digest=sha256:d2076fcd4f99b3e7af30943f433464fb726f777b55e04f1802cd41d16313c33f

Observation 257f28c5-d6ed-4b06-bfeb-5d0b6e4cb1e6 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 217

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T12:04:10.833730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:a869b76093e12aaa71d858c4669b0b7540ff03d87523d8172a5af19b5f03a0d6

Observation 55873d67-a843-40b5-a060-7d878673b2f0 · inbound

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning cites this paper.

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:54:31.206433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T00:53:46.002945Z digest=sha256:92ba9b12772db52613eca7a2115b47a165f194c55fea72f219c07b93b9f0976f

Observation c4a50341-8965-468d-98c0-0d740f3f11d1 · inbound

Scalable Option Learning in High-Throughput Environments cites this paper.

Scalable Option Learning in High-Throughput Environments Learning Multi-Level Hierarchies with Hindsight

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T20:06:49.738978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-18T20:04:58.064472Z digest=sha256:58f64e55bbb6799d6696f0f9967ef145df3f0be245d7e036f5d28465917bdcb2

Observation 5e35741b-8585-4856-96b7-a9ebb544b128 · inbound

Combined Constrained Sampling and Reinforcement Learning for Robotic Manipulation cites this paper.

Combined Constrained Sampling and Reinforcement Learning for Robotic Manipulation Learning Multi-Level Hierarchies with Hindsight

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T03:18:43.091575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:18:43.091575Z digest=sha256:9fa99346a06c2bd88fedc9ea2618978b94c5bb1770727f582a2a87115215effa

Observation a0c15070-9519-4b6b-864c-67ed56e00bd4 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Learning Multi-Level Hierarchies with Hindsight

Reference 101

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:30:58.065014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:90129fcc3ec253e7f1bbf2b52cf748d07562e7abb2823e472f3c78ff0820866d

Observation 10a79b59-f989-4d6f-9a5f-ff99fc4abf32 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:47:21.233473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:36a7c261cea225c137ab0ba9c124da11b1802b9940ef2b9b9040778399de7483

Observation 11ae49c6-7149-4f27-829e-509333980c88 · inbound

Abstraction for Offline Goal-Conditioned Reinforcement Learning cites this paper.

Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:16.522569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T07:46:20.289421Z digest=sha256:98a7b27f44ea75e02e690a73dbec0b7b137e78d822fbc073af37bd158e4b51ee

Observation 4bd6a974-0da4-402a-b89c-7116527be498 · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Learning Multi-Level Hierarchies with Hindsight

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.158320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:ae23aaf25ed90d40847953581248347d26889ea880645c33db5e76e560ab547c

Observation 99c8ec5e-840c-437a-9ce2-1e0ccfab4028 · inbound

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning cites this paper.

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T03:42:47.680209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:42:47.680209Z digest=sha256:e2d2b151efe3c44f30cee10eca8cdb3bccc5e91470ca08e749f7933996e217c6

Observation fcd44bf2-ed5b-4e57-bc47-5d60f92935d3 · inbound

S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning cites this paper.

S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:39.713561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:39.713561Z digest=sha256:a5938266abbf3f1ca14448eb547c9ee493f80074469d31a0140fadca1afce118

Observation e82ab8b3-acef-4e7a-a1df-96add91289d3 · inbound

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning cites this paper.

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-30T14:49:29.163862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T14:49:29.163862Z digest=sha256:7100b5ac5afd599d0ce5bcf5485dc561d104f36bc911e761f4304c36e0eba16b

Observation 460194cc-9184-4098-9d02-d495822f57c1 · inbound

Hierarchical Residual Policy Optimization for Generative Recommendations cites this paper.

Hierarchical Residual Policy Optimization for Generative Recommendations Learning Multi-Level Hierarchies with Hindsight

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T00:28:21.497424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:28:21.497424Z digest=sha256:3350163370799c8379838fe17f3176cf09f2920b1bdea8ad23d35c1531752c15

Observation 72a1f002-9694-4f28-b2f1-e02318a63af0 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Learning Multi-Level Hierarchies with Hindsight

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:34.875368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:34.875368Z digest=sha256:3a9da62412af4003449616d099cae1dc8888ee8f3881fc7c91d14ba4d05246ca