Pith. sign in

Paper Citation Record · LEDGER

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2509.05545.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.05545 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T16:27:33.377046Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 05bd76c5-0d52-49fd-a501-51c0239a0ed9 · outbound

This paper cites Hindsight experience replay.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Hindsight experience replay

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.920035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.232940Z digest=sha256:bd6a7289dd687b6a0783c64da16611ae758505320e4b3c62803a51e56688f3d1

Observation e8ad6641-f987-4468-96d9-d0b70a184f94 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:27:33.900806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.238619Z digest=sha256:a6d70f67cbfa586a7f32cd83d28de82e8ddde44c1cc2abc17ec67c2cdbd31338

Observation 56f1df9e-3bb9-45c8-bc3d-6505cf4d91ed · outbound

This paper cites Barto and Sridhar Mahadevan.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Barto and Sridhar Mahadevan

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.882425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.243883Z digest=sha256:21f298272eff78626940734556ba41d90ff39bebf96c4b2d5a1147c61d83ce87

Observation 6fc8c6bd-0a98-4c7b-aa32-a0a4be840ef4 · outbound

This paper cites Bertsekas and John N.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Bertsekas and John N

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T16:27:33.249281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:27:33.249281Z digest=sha256:28e75b65208ff05ba99cc929627cdae8b4d76ff7de0d3418255ab848372ee337

Observation 96003cee-06a7-4f06-b1a3-47336397cc95 · outbound

This paper cites Bertsekas and John N.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Bertsekas and John N

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T16:27:33.255041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:27:33.255041Z digest=sha256:b14129b04d64dd60c8010f7228e811680af5ea6b78920c1b4820d834889d4647

Observation a6181469-524e-429d-9779-8f176f16ae0f · outbound

This paper cites Near-optimal regret bounds for stochastic shortest path.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Near-optimal regret bounds for stochastic shortest path

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.836185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.260527Z digest=sha256:84639013bb8fa4f8096bec77c8b46f0e11599bbb043af4614bee96f034da3cd0

Observation 5aed9829-5eba-4702-88a9-9e5474ffe8c5 · outbound

This paper cites Improving generalization for temporal difference learning: The successor representation.Neural Computation, 5(4):613–624, 1993.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Improving generalization for temporal difference learning: The successor representation.Neural Computation, 5(4):613–624, 1993

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.815472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.266474Z digest=sha256:0a740e1094cf3925b36ff4f968f0880ddee879ea181f855db763c177196a4275

Observation 6aa8b705-3a1e-4ab9-b947-46ae8de869db · outbound

This paper cites Dietterich.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Dietterich

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.796513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.272000Z digest=sha256:fa0dbfa608d2c9fab2b5cfba7848454e95e06fae3bb3e05f228f084652d0affc

Observation 4a8f384b-3ef5-4401-a2b0-90770e7f07aa · outbound

This paper cites Diversity is all you need: Learning skills without a reward function.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Diversity is all you need: Learning skills without a reward function

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.778644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.277163Z digest=sha256:7cf232a741bae2c64ddee3440a7c80b2b1db5a564fe1dccaf948f09377f11907

Observation 29ad330f-ba64-4167-bcc6-4e73e3819cbb · outbound

This paper cites Contrastive learning as goal- conditioned reinforcement learning.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Contrastive learning as goal- conditioned reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.761464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.282625Z digest=sha256:e2562ed4c7449faf3c9d7bcb46cecfbf52772b23db1758c874f6d92220eac43c

Observation fc62e767-aa32-4749-87ae-6bf39a49999b · outbound

This paper cites Gershman.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Gershman

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.743477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.287910Z digest=sha256:9f17a06a78c8775d8708af96da7ee7225f5f2ba3a9229a93b559988a2ef4a9d4

Observation 9577541d-712f-4db9-ae4d-f9326ebcec4f · outbound

This paper cites Learning to Reach Goals via Iterated Supervised Learning.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Learning to Reach Goals via Iterated Supervised Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T16:27:33.292977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:27:33.292977Z digest=sha256:d35fa50343d73efa071ad8a6e00b8720d020a7c3642c527b3c625b4c801ea227

Observation 872f54b0-5188-4491-a590-eaaf70df424e · outbound

This paper cites Dynamical distance learning for semi-parametric control.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Dynamical distance learning for semi-parametric control

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.724725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.298878Z digest=sha256:ab3763b918ad669462d3468562170367f33abdac3b0137a4fc9f6d5962072249

Observation 587b32ff-c44f-424e-a71a-9ab1179dc598 · outbound

This paper cites Finite-sample convergence rates for Q-learning.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Finite-sample convergence rates for Q-learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.706269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.304257Z digest=sha256:c2de36e75396a0a0cbfbc4c1a1a241648b468da45f66340506c61e3ac5820ceb

Observation dd3af30b-7977-4a94-bc32-fbb71eeace47 · outbound

This paper cites Learning multi-level hierarchies with hindsight.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Learning multi-level hierarchies with hindsight

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.686190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.309289Z digest=sha256:5ff7f363bb7e1a358ca1d96a4f82c3818b20c28d441e92972d438b6eca305563

Observation 03c69c57-e2ca-4da5-bd5d-8cd2f845dd46 · outbound

This paper cites Lillicrap, Jonathan J.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Lillicrap, Jonathan J

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.668970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.314318Z digest=sha256:e5340af647655a80197d942cc6377a819c9a1a7013550f7c7f5c09a5b020d15b

Observation 932a3765-cc86-4ec1-80b2-af34429bbea7 · outbound

This paper cites Learning stochastic shortest path with linear function approximation.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Learning stochastic shortest path with linear function approximation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.650598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.319086Z digest=sha256:e532d33d14901294d6d63ca20117885660478f40c970fd62dadcf54d0ba739c5

Observation 84603847-4463-482e-92cf-c1a60ee9a232 · outbound

This paper cites Data-efficient hierarchical reinforcement learning.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Data-efficient hierarchical reinforcement learning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.629928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.324260Z digest=sha256:bb5f7e6d50876fdc210c7fe6f709c92b069c656dd47b2c148d814097e89b3155

Observation dab6480a-b62f-455c-835f-06d3add4426f · outbound

This paper cites Hierarchical reinforcement learning: A comprehensive survey.ACM Computing Surveys, 54(5):1–35, 2021.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Hierarchical reinforcement learning: A comprehensive survey.ACM Computing Surveys, 54(5):1–35, 2021

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.609131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.330109Z digest=sha256:0ea46d527ca3969784c2e94fa14b0b3fd3f8ce4d883b98485d4d56c7dc80cf2d

Observation 935be90d-192e-467b-b257-59597388b7cf · outbound

This paper cites Efros, and Trevor Darrell.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Efros, and Trevor Darrell

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.589629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.335666Z digest=sha256:6709cff1408a1de7909c8d16a7d96f3384fc9787db8d8fbdfa18ab562632a871

Observation 382b9d60-c1fb-4091-b7d9-79728ea1233e · outbound

This paper cites Universal value function approximators.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Universal value function approximators

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.570166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.340887Z digest=sha256:b2ecb80768136be3cb91a9ddbd1d1ebd54cc969f71ffc6a1c72f9a36bb4d57f3

Observation 66b03e62-c692-42f4-83f3-c9e313735709 · outbound

This paper cites Toshev, Sergey Levine, and Brian Ichter.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Toshev, Sergey Levine, and Brian Ichter

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.552596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.346188Z digest=sha256:c4b12615be4b977c3b638083a6f30fdc8dec5ff69ba37148407ab8da2a0bdbd2

Observation 07b67f88-073e-4a8d-8b9e-607a63254adb · outbound

This paper cites an unresolved cited work.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T16:27:33.351379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:27:33.351379Z digest=sha256:1bd867efe10354151e7e842c4e83daec374f2e25ac3685c770c30956a7262518

Observation 54821549-38df-486b-8da5-b03b789a7aaa · outbound

This paper cites Sutton, Doina Precup, and Satinder Singh.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Sutton, Doina Precup, and Satinder Singh

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.515944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.356283Z digest=sha256:2616b8ace5db719f1b2fb04eaf1b62182275a5c2044cc702af89e709e5cfb569

Observation 9dc56b6d-a174-4c62-adc3-dca78457cbea · outbound

This paper cites Morgan & Claypool Publishers, 2010.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Morgan & Claypool Publishers, 2010

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.498179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.361600Z digest=sha256:c2ff5acdfaad996a97e0b2c98524f605d5f1706a69cf5626f90e36517fc0beca

Observation 42a9d886-b834-4087-9119-7104e373c288 · outbound

This paper cites Value iteration networks.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Value iteration networks

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.480646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.367104Z digest=sha256:4642de087170dc77e997dd427c42ef2f85db726d4c1bf0734cfda40c3d3e0420

Observation c21094b6-4904-447e-928e-071c048c4745 · outbound

This paper cites Sample complexity bounds for stochas- tic shortest path with a generative model.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Sample complexity bounds for stochas- tic shortest path with a generative model

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T16:27:33.462195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.372191Z digest=sha256:aea97dda6ae49aa027a69e3bb87a5f2580f4666a0d21e52a6d527cae2006c4b1

Observation 03cd3234-0a4b-4fbc-9567-7c5636a29d88 · outbound

This paper cites an unresolved cited work.

Reinforcement Learning with Anticipation: A Hierarchical Approach for Long-Horizon Tasks Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-15T16:27:33.439595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-15T16:27:33.377046Z digest=sha256:9f829f42b4b9b54bb2ff7c0b1cde5a8b6a00b230e1b380e3d18f128dfbbfbc72

Pith citing papers

No inbound Pith citation observations are available.