Pith. sign in

Paper Citation Record · LEDGER

Learning to Reach Goals via Iterated Supervised Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:1912.06088.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1912.06088 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:21:25.585866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T06:49:37.764826Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd565aa9-bc56-47dc-acfe-3f509a04729c · inbound

Decision Transformer: Reinforcement Learning via Sequence Modeling cites this paper.

Decision Transformer: Reinforcement Learning via Sequence Modeling Learning to Reach Goals via Iterated Supervised Learning

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T15:11:11.146419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T15:11:11.056013Z digest=sha256:08d873c9d67a2e8394a5b6cb1de187895ffef680bc8ecdb32e0572f69e6df13b

Observation a2365892-1fd6-4786-9f4b-ca7f9256155e · inbound

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies cites this paper.

Prompting Decision Transformers for Zero-Shot Reach-Avoid Policies Learning to Reach Goals via Iterated Supervised Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:21:25.585866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:21:25.585866Z digest=sha256:ed32671a085465c423e909ff31bad69064490a4778b1839e1abe74c6c407ab82

Observation 501446ab-2c25-4b60-8f7b-9466e8128aff · inbound

Normalizing Flows are Capable Models for Continuous Control cites this paper.

Normalizing Flows are Capable Models for Continuous Control Learning to Reach Goals via Iterated Supervised Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:49:57.280493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:49:57.280493Z digest=sha256:a82ca238e7fb87daaca1b90b8ff118a760074e119f8a1c2f05336870deaa89ad

Observation 77335ca6-e087-4292-8657-b2892f585840 · inbound

Efficient Skill Discovery via Regret-Aware Optimization cites this paper.

Efficient Skill Discovery via Regret-Aware Optimization Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:34.963486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T22:41:34.963486Z digest=sha256:973b67006e9b9c4a5165ef58b872fa80bcdec190522153fcca2d7d6f5b2efb8d

Observation 5cf61d41-82c4-49f2-a1c1-391bcedd557c · inbound

Behavioral Exploration: Learning to Explore via In-Context Adaptation cites this paper.

Behavioral Exploration: Learning to Explore via In-Context Adaptation Learning to Reach Goals via Iterated Supervised Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:15:44.449930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:15:44.449930Z digest=sha256:4a170973dc6f99630b47efb1ad4432a2a8582c55ae7195eac3d9eec7b4741811

Observation 43cca234-29f1-4921-8ad4-e0236c4c209a · inbound

Equivariant Goal Conditioned Contrastive Reinforcement Learning cites this paper.

Equivariant Goal Conditioned Contrastive Reinforcement Learning Learning to Reach Goals via Iterated Supervised Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:24.538792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:24.538792Z digest=sha256:652bdf653c35ca00440681a9ee19fb1c0d4165599caa016453ef4c1eb3b510c2

Observation 6ad19f8a-a269-4666-9acf-cc9855668f46 · inbound

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers cites this paper.

Generative Sequential Notification Optimization via Multi-Objective Decision Transformers Learning to Reach Goals via Iterated Supervised Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:57.847334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:57.847334Z digest=sha256:1481815645f597a8faf73f0f44c253c84f2714effd47a9110cadb961da146586

Observation f91be620-0707-400c-ad6c-ff4412ea2e29 · inbound

Compositional Diffusion with Guided Search for Long-Horizon Planning cites this paper.

Compositional Diffusion with Guided Search for Long-Horizon Planning Learning to Reach Goals via Iterated Supervised Learning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T13:13:43.891012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:13:43.891012Z digest=sha256:594a6fc607abcfd8b97c6c2e11b2237142f5566a03b8efc1afdf423d6b2f372d

Observation 61541cb6-5ace-4b62-b49a-b960b8966cbe · inbound

LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels cites this paper.

LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels Learning to Reach Goals via Iterated Supervised Learning

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:09:22.537167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T04:09:22.328844Z digest=sha256:be07507a46791eb26c60548ae4c417631d2d76a1e5503f4d76fe789980b078c0

Observation 6bf7c55e-445b-40e5-b3dd-c16b65b7fb51 · inbound

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning cites this paper.

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning Learning to Reach Goals via Iterated Supervised Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:20:58.412122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:12:32.839681Z digest=sha256:24a37ab53822532742718a4d60d290c994eeff983fe728e8af25c74dfda752a8

Observation 5aa919d6-8f34-4df2-b8f5-833489def7b7 · inbound

From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning cites this paper.

From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning Learning to Reach Goals via Iterated Supervised Learning

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:45:59.479633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:30:34.434984Z digest=sha256:bda2c6b14bcedd675e890018dcdc53c3a1887f909d1f797947afa2c0ea500c77

Observation ebc466fa-a539-42e0-8872-1dece75ebddf · inbound

GCImOpt: Learning efficient goal-conditioned policies by imitating optimal trajectories cites this paper.

GCImOpt: Learning efficient goal-conditioned policies by imitating optimal trajectories Learning to Reach Goals via Iterated Supervised Learning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:36:14.937117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T11:26:22.149056Z digest=sha256:d946eba2d145e478db48ae6c8bc78b506f2d64d5ec1a640c840e3ae0e6ccb937

Observation 9cd5e1be-e9e7-426e-89cb-ac5bc540f741 · inbound

Refining Compositional Diffusion for Reliable Long-Horizon Planning cites this paper.

Refining Compositional Diffusion for Reliable Long-Horizon Planning Learning to Reach Goals via Iterated Supervised Learning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:11:18.892607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T17:47:58.141126Z digest=sha256:2968f91b91f3e73affcc0b991444f87730a3c9401240490edfbab58b9ffd394a

Observation 1f251c31-f01f-44ec-bb2d-a33d443c71a1 · inbound

Predictive but Not Plannable: RC-aux for Latent World Models cites this paper.

Predictive but Not Plannable: RC-aux for Latent World Models Learning to Reach Goals via Iterated Supervised Learning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:50:57.291741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T02:11:26.526418Z digest=sha256:61ffd349f29525143e9a7429a27c8c545f651edb651013a5e6783df67571682d

Observation eedc35fc-db0c-491c-b15a-4d9bf6b32975 · inbound

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning cites this paper.

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning Learning to Reach Goals via Iterated Supervised Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:16:26.289713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:35:37.739085Z digest=sha256:aa850942b73e68aac3963406649ad7670d3b7386cd530eb6a8c567b3431dc378

Observation 0c3561a8-9ed5-4b13-b510-7f5d4cfafe5b · inbound

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation cites this paper.

stable-worldmodel: A Platform for Reproducible World Modeling Research and Evaluation Learning to Reach Goals via Iterated Supervised Learning

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:01:20.332386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T08:57:19.834179Z digest=sha256:769759d3179fef8baefabb447f550bf24e752e79a4826cdbafa1e76680940af5

Observation eaa4089d-ea58-428a-b5cd-a7c926920358 · inbound

Goal-Conditioned Agents that Learn Everything All at Once cites this paper.

Goal-Conditioned Agents that Learn Everything All at Once Learning to Reach Goals via Iterated Supervised Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:00:21.639083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-25T04:59:48.867927Z digest=sha256:a77aa7fbbc784162d745a8592d63336ea99ff9bfac63f6bc51fb7f529a9aebe9

Observation 44fae367-5737-45ed-a48a-14de6a58cc3e · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Learning to Reach Goals via Iterated Supervised Learning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.180587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:37678b35d2c32a89c83a9e88983adcd10b12c8be83f48fb02cdfde25e5f6d234

Observation 35711ff6-2e43-42c2-9206-285dde74bc2e · inbound

Energy-based Compositional Diffusion Planning cites this paper.

Energy-based Compositional Diffusion Planning Learning to Reach Goals via Iterated Supervised Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:49:37.767185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T14:12:33.307261Z digest=sha256:a96ff5d5e31b6b517f8ba0769fb52e5253b8d20479e5ed153afc75bc1cab0aef

Observation f9624c5c-ee09-4152-827d-13cbb574d786 · inbound

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations cites this paper.

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations Learning to Reach Goals via Iterated Supervised Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-11T20:45:51.403759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T20:45:51.403759Z digest=sha256:94855946fb323fc8152271459fd8c2fe1f104f5b663e8b7cc5339640d07258d3

Observation 9583241c-4031-4b02-b8df-99f34b72dffa · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models Learning to Reach Goals via Iterated Supervised Learning

Reference 163

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:13.164857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:13.164857Z digest=sha256:3635f5ceace1a432a6f93c75aa85dbb4c7863681456ee7fd9988f13b77138471

Observation d74aa734-1585-4a62-9ec3-50ea4030e825 · inbound

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning cites this paper.

VisualPatchWorld: Code World Models as Latent Structured Representations for Planning Learning to Reach Goals via Iterated Supervised Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T03:05:32.177521Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:05:32.177521Z digest=sha256:daeb2feb68b8fbc718a1b780c48821e1179d715c973ba05c426d2734054215de

Observation 21c83659-5f2d-4b0d-815c-58825c68a30d · inbound

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control cites this paper.

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control Learning to Reach Goals via Iterated Supervised Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T02:48:10.458568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:48:10.458568Z digest=sha256:ab3c1c1aa7b58448eeb627f9e5c87a9bc2c364a8730d0f269308d0e552c06a3c

Observation c44246e3-72aa-4150-9137-819f79b53e22 · inbound

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models cites this paper.

INTACT: Isomorphic Intent-to-Action Learning for Search-Free World Models Learning to Reach Goals via Iterated Supervised Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-01T00:49:00.033074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:49:00.033074Z digest=sha256:0aab2cd9b917574c6f7b1030aa0e2e4f9a350fd1893ec01630b2110cb54ac05a