Pith. sign in

Paper Citation Record · LEDGER

Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2101.07123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2101.07123 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:20:39.307427Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T06:55:29.516438Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b7f302f2-cda0-4f36-abda-0e42be550535 · inbound

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following cites this paper.

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T19:20:39.307427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:20:39.307427Z digest=sha256:7ec4b96abc546b69ba9f0d5ca9c59e1e8271586c934e735f98fc600abd32922c

Observation 0eec7250-8f73-4841-8163-0584315efc57 · inbound

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization cites this paper.

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:25.733026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:05:25.733026Z digest=sha256:2c029cb4455758ce3f79ab8deedde346cae8c824354566d991c8a202f4f6ead9

Observation 477dd4b7-d9b0-455f-9344-1668513ca4ba · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.616008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:b045d5490609546def14534a65eaeda01f906d3af13740b6b49a8599f3dd4082

Observation f794af6c-6908-480e-9bcd-2e4b4551201b · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:17:14.107283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:7d38d653abd8a5005684cc2f5377c7de922378ee4637836eab8000beabaa7acb

Observation aa8d2a68-e130-47e1-8ba9-d8cd7914e5c6 · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:16.986525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:16.986525Z digest=sha256:d140cd2acc97997f6e968976b5e37c22400b070e19056ad5ce58f6dc3ebe8f79

Observation 0211dde7-c40f-4a47-b775-6b561654b0ed · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:32.924634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:32.924634Z digest=sha256:9c22063ac3a323c466834dcec14214baf91c786b7ee542b1bfd979d1dc0b6f56

Observation e061b358-255f-4ce7-9a09-2e8fc4a78851 · inbound

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction cites this paper.

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:19.258076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T08:32:21.705085Z digest=sha256:52a53a5d13f1ecb74e4cc955bba6092d59f927d477056fc8409c1a450c6a8ca2

Observation 0307ed90-5d82-46c7-a7dd-57725bd5f701 · inbound

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning cites this paper.

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:46:37.135018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:45:54.155848Z digest=sha256:05625e8655ee7d8787f9b4a9c60328d071811fb8445125082645e6caa98a9c15

Observation 6f39062a-6a80-4e0a-ac1a-6d90532133d5 · inbound

Understanding Human Actions through the Lens of Executable Models cites this paper.

Understanding Human Actions through the Lens of Executable Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:10:09.777561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:54:29.565609Z digest=sha256:11a3a8612ddff2cd0982ebceaca1921909cc076dc7f3f0b0aedfbc9a16f813b1

Observation 0f1edb4e-c796-4a8b-a7ac-8a37a60e40b7 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:57.816224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:8118bb527992f5645c4748ac9af26b399a6576bc69f6f9bc785593250c738dd5

Observation 9f6d1144-0655-4e83-928e-8e6bb24dd369 · inbound

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning cites this paper.

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:42:58.552888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T20:26:35.019753Z digest=sha256:b39156ca5d33d967bfe674fc560fcaeb57b4dd9b34cbf511c9108018c99254f8

Observation 44ae2fa9-89eb-434c-b699-60bf093fbad7 · inbound

Offline Reinforcement Learning with Universal Horizon Models cites this paper.

Offline Reinforcement Learning with Universal Horizon Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.499849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T19:45:15.458347Z digest=sha256:ae9e3c171d8dadc04de1cabff361fa3c4c398a4606576b4d05dc329d9d7b3929

Observation fb48cde2-afe1-4053-a90d-38dd781e6fc2 · inbound

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited cites this paper.

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:37:45.250688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T20:37:36.030165Z digest=sha256:ae4244f55cc4f96e302356b3338cc9583554ad60da9d93b09ce5397505ce5095

Observation 2329491f-01f8-4d0e-8440-266a531305c8 · inbound

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM cites this paper.

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:53:15.706839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:51:14.809528Z digest=sha256:e828a63f3a524c5684ae2ae9023c658b25beedaa9a4edad83814ec085347a82b

Observation 678e5fb1-3c4f-4430-b7b8-b626efcb2039 · inbound

Exploration and Online Transfer with Behavioral Foundation Models cites this paper.

Exploration and Online Transfer with Behavioral Foundation Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:55:29.518692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T06:50:59.743859Z digest=sha256:c46f5fccd85d9f689594b69543bd15834cfc79de35c202d4e4c2cf7a0549ce63