Pith. sign in

Paper Citation Record · LEDGER

Bridging State and History Representations: Understanding Self-Predictive RL

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2401.08898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.08898 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T00:48:43.773234Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T10:27:14.544741Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 772587d5-f837-45c5-b9a6-a8171c8d85e6 · inbound

Latent Action Learning Requires Supervision in the Presence of Distractors cites this paper.

Latent Action Learning Requires Supervision in the Presence of Distractors Bridging State and History Representations: Understanding Self-Predictive RL

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-09T19:18:45.052790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:18:45.052790Z digest=sha256:ef68d563daba297e5e6c61f85b0fed5d6fdb9d2d4b4cf6bba021fd3ac56262fd

Observation cfd0025f-e6c3-4433-adbf-41ade8207f19 · inbound

Improving Transformer World Models for Data-Efficient RL cites this paper.

Improving Transformer World Models for Data-Efficient RL Bridging State and History Representations: Understanding Self-Predictive RL

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T14:59:44.999274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:59:44.999274Z digest=sha256:8b331b8257a14030e09ee11b491c7f9b5fb7b49d687b370afc4b16359e4361b8

Observation 0decce05-3a09-4d87-85da-b092720781fe · inbound

Hadamax Encoding: Elevating Performance in Model-Free Atari cites this paper.

Hadamax Encoding: Elevating Performance in Model-Free Atari Bridging State and History Representations: Understanding Self-Predictive RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:18.249383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:23:18.249383Z digest=sha256:50be45b62cde6b61a3d9eee2a2ed96cfff48f27d9165c6e1549967fdbdd18b5c

Observation c7d1b2e4-9bf3-443a-8c6e-ece296354c84 · inbound

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments cites this paper.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bridging State and History Representations: Understanding Self-Predictive RL

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:04.252390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:04.252390Z digest=sha256:c1b048429534e152f96cacf23cf50a4d8018c051da08281d01edadb19d561b86

Observation 5f3ef2cc-45e1-413c-958c-aa0fbc723349 · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Bridging State and History Representations: Understanding Self-Predictive RL

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.546674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:97fe83439f2b614f4bb5ab43ff15d113930fd0e1b170dac6f145867468de3447

Observation a5e4e9e1-19f5-4833-874f-66ad4f364888 · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Bridging State and History Representations: Understanding Self-Predictive RL

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:14.174785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:4a6bd3818d9e18348fc8c14250810ba34bbf82a449ed791f066eec2e4361f3a2

Observation e961f44d-52cf-4de8-99ab-6c98522680bb · inbound

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access cites this paper.

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Bridging State and History Representations: Understanding Self-Predictive RL

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:30.088955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:43:30.088955Z digest=sha256:dd8dc34730cd43a9ac2fb070f12202709df5b014167c4c735ea7b6178198b827

Observation 5e01d52c-1e94-4f20-9bfd-5669015cfaf7 · inbound

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels cites this paper.

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels Bridging State and History Representations: Understanding Self-Predictive RL

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:27:32.259434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T07:23:03.673259Z digest=sha256:a434461fcf28dccfba336f80ef285069612ad9a6f8df6ee005908ec761900121

Observation e95d30f0-92fd-4d0e-a404-9e1d6634045b · inbound

Can We Really Learn One Representation to Optimize All Rewards? cites this paper.

Can We Really Learn One Representation to Optimize All Rewards? Bridging State and History Representations: Understanding Self-Predictive RL

Reference 2023

Resolution
malformed identifier
no resolver link, observed 2026-08-03T00:15:46.425044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:15:46.425044Z digest=sha256:e7323b673de8f16dfd2b264dc3c07f68300aff7fa6f8342db58dffa5c00e333f

Observation fc878abf-3964-4754-982b-2fc77cd7285a · inbound

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction cites this paper.

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction Bridging State and History Representations: Understanding Self-Predictive RL

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:50:05.098809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T14:46:18.645009Z digest=sha256:47f0eebc92c58d1fec0d5d58ec2fdf1472a5af821280721f4106e0eda265c1f9

Observation 089e678a-6793-4f26-94a9-74919ce31458 · inbound

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control cites this paper.

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control Bridging State and History Representations: Understanding Self-Predictive RL

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:33:10.250941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T19:30:23.447901Z digest=sha256:436bdced93cf8744d29ef922da08ca9e3dbb8a61ee384c9c516ad4cf855ab14f

Observation 8698c271-a127-42da-be19-1afc2fc38300 · inbound

The University AI Didn't Replace -- Rethinking Universities in the AI Era cites this paper.

The University AI Didn't Replace -- Rethinking Universities in the AI Era Bridging State and History Representations: Understanding Self-Predictive RL

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T17:22:42.591825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:22:42.591825Z digest=sha256:b6431622e9662cd11f7bc279b3cae804f8ba8c777e9c87f7f35c7ec3adb48e68

Observation 1c04a257-0296-4fbd-9fdb-2c13687a2507 · inbound

Integrating Causal DAGs in Deep RL: Activating Minimal Markovian States with Multi-Order Exposure cites this paper.

Integrating Causal DAGs in Deep RL: Activating Minimal Markovian States with Multi-Order Exposure Bridging State and History Representations: Understanding Self-Predictive RL

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:56.173195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T02:34:12.141437Z digest=sha256:6c74f4a7e18d37be3c33ce35b58f157b80aa24cdbb6e63b8e02419bd3fa7a4d2

Observation 0cb8191a-1a14-4841-9e13-0f12692b0a2c · inbound

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback cites this paper.

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Bridging State and History Representations: Understanding Self-Predictive RL

Reference 257

Resolution
unresolved
no resolver link, observed 2026-08-03T04:39:32.161140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T04:39:32.161140Z digest=sha256:80a9ccae735b186de785ff3e48a93be33d3794de63d67d8ed88cc3dd023d6d78

Observation fe6e40b3-25ba-434b-815b-4e402073984e · inbound

Hierarchical Latent Prediction for Language Models cites this paper.

Hierarchical Latent Prediction for Language Models Bridging State and History Representations: Understanding Self-Predictive RL

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T23:13:07.349666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:13:07.349666Z digest=sha256:68abf5ec34a95ba4a5ddb20c11a7e24ce35aac5a31d57ee1e71fac826ee581de

Observation 66e176ce-c77d-4edd-a955-5227dad8bba4 · inbound

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control cites this paper.

V-Simba: Unleashing the Architectural Potential of RL in Visual Continuous Control Bridging State and History Representations: Understanding Self-Predictive RL

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:43.773234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:43.773234Z digest=sha256:326380bb45d86ce77f6cb41d4d40904db402d7e4a16b533cf1f2d980cce3b00b