Pith. sign in

Paper Citation Record · LEDGER

Goal-Conditioned Reinforcement Learning: Problems and Solutions

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 32 inbound Pith citation observations for arXiv:2201.08299.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2201.08299 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 32 of 32 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:25:49.129256Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T03:24:28.836511Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b9f768c3-f6c6-4e19-8ced-8946163901b7 · inbound

Vision-Language Foundation Models as Effective Robot Imitators cites this paper.

Vision-Language Foundation Models as Effective Robot Imitators Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:44:27.676683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T21:44:27.562453Z digest=sha256:8580bdf08e1200eae4295aa78c809d8b38f91e406ca218cdb3657b5076155771

Observation 106e6232-3df4-4967-9c6b-b83b751e6ca6 · inbound

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning cites this paper.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.129256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.129256Z digest=sha256:89e628d67ef2facffde911a344c2257c574aae7fdb69986404dca3b77a5f312c

Observation 72854cbe-c926-437f-b070-49ce31435517 · inbound

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion cites this paper.

DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-19T12:52:17.883282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T12:50:17.902979Z digest=sha256:62e7c637200ca334ac009ae6e65d9702229de2bc76aa5046572c67d8e9938123

Observation c5d1b977-6ba4-481b-a841-4a852e7a9bbc · inbound

Reachability Weighted Offline Goal-conditioned Resampling cites this paper.

Reachability Weighted Offline Goal-conditioned Resampling Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:24:59.397642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:24:59.397642Z digest=sha256:f443a0227d0c4728455fefbcc6b418dcc30e8c613ee6b5e476d71c034344624c

Observation 7974a871-e96a-47a5-9f66-4fe6ab15f670 · inbound

Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models cites this paper.

Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:03:30.107695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:03:30.107695Z digest=sha256:dd7335f65ed798f3852dd6a423b76a80a8bda7bae02083324260238b938161f2

Observation a2a05581-2fcd-41d8-8b9b-66230649f4ba · inbound

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning cites this paper.

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:54:31.197090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T00:53:46.002945Z digest=sha256:ae8dd42917208f4f20a05cd51bfd0167184503db37630ad5481b801655885f9f

Observation 4ea218e3-983c-4711-9946-e715fba6e942 · inbound

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving cites this paper.

Bourbaki: Self-Generated and Goal-Conditioned MDPs for Theorem Proving Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:27:36.349835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:27:36.349835Z digest=sha256:7be1159f68383208fff52d2fb34d91f84de6246fcb619080b7e12742452958da

Observation b1ecc38e-6980-4d53-9207-fa812b98ce3b · inbound

Equivariant Goal Conditioned Contrastive Reinforcement Learning cites this paper.

Equivariant Goal Conditioned Contrastive Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:26:24.504219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:26:24.504219Z digest=sha256:2eafbd35751da4b812653a4905fbecab37086a02f20b8eecb3301c35c34c5fb0

Observation f57e4368-b5c0-4e4f-b89a-b56b308c0b8a · inbound

Self-Curriculum Model-based Reinforcement Learning for Shape Control of Deformable Linear Objects cites this paper.

Self-Curriculum Model-based Reinforcement Learning for Shape Control of Deformable Linear Objects Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T20:56:58.581176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:56:58.581176Z digest=sha256:563e5536d62a31385bc1206d45002e86d3c8ece43bc4b2f2ebd0861fe01a935e

Observation 3ea200fb-773a-400b-8b47-c9146e1a8b61 · inbound

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning cites this paper.

Efficient Hierarchical Implicit Flow Q-learning for Offline Goal-conditioned Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:20:58.280249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T18:12:32.839681Z digest=sha256:f81d6ca8760aed6174209af0741a811c7ed165484673d9f50c62000270e99d12

Observation fe268555-00e9-4533-96a9-0d2d9d6035fc · inbound

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking cites this paper.

AdaTracker: Learning Adaptive In-Context Policy for Cross-Embodiment Active Visual Tracking Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:34:47.312157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:34:26.106672Z digest=sha256:f81ca645047e7b1fe76fd8f72fc86eb0b4395952ddff5593dd5d4cd195eefef3

Observation 56a29c79-bb3c-482c-925d-d536bda25bad · inbound

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning cites this paper.

Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:24:47.179747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T00:19:48.053466Z digest=sha256:89a3717992ed85ef4b3d7209b6246e84131260929a49136043fd21cae2b3afec

Observation 738fbe2b-4b8d-41aa-b108-1fac1b1be53e · inbound

GCImOpt: Learning efficient goal-conditioned policies by imitating optimal trajectories cites this paper.

GCImOpt: Learning efficient goal-conditioned policies by imitating optimal trajectories Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:36:14.892367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T11:26:22.149056Z digest=sha256:850200df567d33ccea662d405bb14f78f680d7eb84a77171ece2a776de9ce095

Observation 14498daa-0227-42f1-a0fc-433d5d8fe9a5 · inbound

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning cites this paper.

When Policies Cannot Be Retrained: A Unified Closed-Form View of Post-Training Steering in Offline Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:49:16.044470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T22:15:16.221059Z digest=sha256:e3de769820753dcf8572d74ddbbec3175aa549728205f5d2e238cd7f84d7f478

Observation 2d00f843-1424-42ac-8c40-fc1ba8dcdd41 · inbound

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning cites this paper.

SpecRLBench: A Benchmark for Generalization in Specification-Guided Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:30.363716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T04:01:31.928056Z digest=sha256:64afc97b21ab92c674f9abaa3e33698539af9d836a194789ff56559d0a1a48de

Observation df4f7c2a-1c27-4f4d-b391-6da788d9c30e · inbound

Improving Zero-Shot Offline RL via Behavioral Task Sampling cites this paper.

Improving Zero-Shot Offline RL via Behavioral Task Sampling Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T23:41:18.632617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T16:27:15.347522Z digest=sha256:aa922936355969c0b64fe6ada6154061ccd7f3e1b09e2408ca23a9b753c48ae4

Observation dd71071a-da58-4d56-8dc2-c48e975db8a1 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 108

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:30:58.762428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:6cc31b222f0f21b252c5c1b063d35601cb05cd27936b6d6e6f4ab4c03103cd98

Observation 6cb6eebc-8701-4752-a0d8-752b89664768 · inbound

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation cites this paper.

Plan in Sandbox, Navigate in Open Worlds: Learning Physics-Grounded Abstracted Experience for Embodied Navigation Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:11:27.293862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T03:36:24.941205Z digest=sha256:418ebb78adc13f69911e4d4e26625e3d95900a2d3ffe29497fe69a9711a54946

Observation 1b2e9950-94a1-48fc-9d8b-069ac64a58dd · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:32:30.552367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T07:28:24.455817Z digest=sha256:a4ac446090f5f8fc6fad1d281a181e148c8ec640e82344d7330c423c52b0bece

Observation f9442c02-19e0-4661-a17f-b2f8c42540f0 · inbound

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning cites this paper.

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:49:10.042051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:48:55.661356Z digest=sha256:0583b5afdbdf8f4b853004fbc117975f4994fbce94aadfb9743efd6e90e07e1c

Observation 9e615d03-41ff-43dc-8644-a2710baab1c4 · inbound

Goal-Conditioned Supervised Learning for LLM Fine-Tuning cites this paper.

Goal-Conditioned Supervised Learning for LLM Fine-Tuning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:39:10.101400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T22:37:46.345159Z digest=sha256:6a127c57bd7b848a2b77e494c3baa6ce37a36a03437d703a73080f230a9bce86

Observation 0cc70f80-1952-4bd7-b18d-166027b674c9 · inbound

CurveRL: Principled Distribution-Aware Context Reweighting for LLM Reasoning cites this paper.

CurveRL: Principled Distribution-Aware Context Reweighting for LLM Reasoning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-06-30T14:04:44.914592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T13:54:52.129738Z digest=sha256:3928f80168be5238fd516c19d5382609ff57f0c1b98d70ca982324eb5b1053c8

Observation 37a7e5ff-c5f6-486f-a464-71898833ec2f · inbound

Decoupled Behavioral Cloning for Scalable Inductive Generalization in RL from Specifications cites this paper.

Decoupled Behavioral Cloning for Scalable Inductive Generalization in RL from Specifications Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-28T20:42:37.703603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T18:25:37.397393Z digest=sha256:dcb621d46bd4dd78179a3cbae4b31fd1443e30ccb792815f846734f3fd26d047

Observation 9b1e5aaf-2f60-4689-916f-f6175e76c5b8 · inbound

Dual Advantage Fields cites this paper.

Dual Advantage Fields Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T02:46:27.870593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:43:13.615850Z digest=sha256:4ec1e8171205dfe1b4f6b80c2b6b4bd438e2abc030a4cc56a631c8f8a12cf65e

Observation 3de0bea2-bdc0-4011-88ed-02739649657b · inbound

GUIDE: Goal-Initialized Directional Understanding for End-to-End Visual Navigation cites this paper.

GUIDE: Goal-Initialized Directional Understanding for End-to-End Visual Navigation Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:47:42.059834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T13:01:33.109009Z digest=sha256:98649bbcaa605322d7379601b2444407f856cad1eebecaacab4b67a23bc644eb

Observation 2c96bc35-26b7-44d2-8ffd-c6054725e69c · inbound

Learning Object Manipulation from Scratch via Contrastive Interaction cites this paper.

Learning Object Manipulation from Scratch via Contrastive Interaction Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.436177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:10:21.427118Z digest=sha256:30b82a4708976ecd5c661d63c5e63ad57bd533d0e5625b7c0d0307536f9b709c

Observation aa3b42f0-f16f-47a4-96c5-f77931f3d7fa · inbound

Embodiment Shapes Rolling Behavior in a Multimodal Infant Model cites this paper.

Embodiment Shapes Rolling Behavior in a Multimodal Infant Model Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:56.261209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:08:24.273577Z digest=sha256:cd0ae83b33978f1516981f87576f172c454ddd940a0d0f5ca4584006d9ed6386

Observation 1faef4cd-863f-438d-9025-a7b4142d9312 · inbound

World Models in Pieces: Structural Certification for General Agents cites this paper.

World Models in Pieces: Structural Certification for General Agents Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:00:01.600010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-25T23:04:11.950176Z digest=sha256:e56f364d7a1ac34caa28928f77655a0f47c518f0469b59d50d6a7a1219dbc011

Observation 6edea12b-d8cc-420d-806a-21d0d2a678ee · inbound

Coachable agents for interactive gameplay cites this paper.

Coachable agents for interactive gameplay Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:56.521838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T12:52:05.010028Z digest=sha256:ca0de596c80605c5e479d80ac33083614bd9ed1940a34cc635975b2b0cbac693

Observation fee86fc1-268c-4449-afde-bcd9c2db1f0b · inbound

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games cites this paper.

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 14

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T03:24:28.837788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-08T03:17:09.608451Z digest=sha256:ad9d4782c40dfbc6c98db129e1b7360e00628e8d8194d74003d63b6650aaa9cf

Observation 502c102c-5cce-4c05-baf4-bccce6d4c1bb · inbound

A Single Diffusion-Policy Controller for Multi-Task Block Pushing with Zero-Shot Sim-to-Real Transfer cites this paper.

A Single Diffusion-Policy Controller for Multi-Task Block Pushing with Zero-Shot Sim-to-Real Transfer Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T08:30:10.584105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T08:30:10.584105Z digest=sha256:f7faf52272b30f284d179469b1f1e42a98fc251cb1cb0b9c98882d8d7db89bd5

Observation 8c3fc251-e5d9-4816-a034-a9522b7f4a3e · inbound

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention cites this paper.

DAGR: State-Conditioned Goal Representations via Difference-Aware Goal Cross-Attention Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T04:15:53.392271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T04:15:53.392271Z digest=sha256:a62eda448b8fd157c44fe0c629e9069fb21a9d007cb0e259d6fdfb3a82b5bae5