Pith. sign in

Paper Citation Record · LEDGER

Data-Efficient Reinforcement Learning with Self-Predictive Representations

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2007.05929.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2007.05929 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:12:58.088970Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:19:31.235420Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation db04718e-1378-4546-b164-cd9b2dd2f3b0 · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.543460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:46fc525e72a0d0aa9e838447107a45d5c4ab8ea06dcd60dc126fddebd6711cf9

Observation 1bc51aeb-a694-41a9-9a81-2578af3f264e · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:17:14.006364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:b30551377fafd15170ec207a2006929da2292545e9e51535445b0ef9d6b3e5d3

Observation d81c0600-606e-4fb7-a4e2-bbbdbaeac6d6 · inbound

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease cites this paper.

Sample-Efficient Reinforcement Learning Controller for Deep Brain Stimulation in Parkinson's Disease Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:12:58.088970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:12:58.088970Z digest=sha256:5cb306a99e57eba692e31ff143f47d3b1ee716eb627ebf83903a659158f056d2

Observation 3c24405b-13cf-45ee-9636-e5c7338eda5b · inbound

Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning cites this paper.

Balancing Expressivity and Robustness: Constrained Rational Activations for Reinforcement Learning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:55:36.073397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:55:36.073397Z digest=sha256:5f83a867e6f9132c6adcaf6f81a70367840a16ce66542d5ec46736ac05a5675c

Observation 44b43f42-3d22-423e-804b-418aea13d67b · inbound

Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach cites this paper.

Human-Like Goalkeeping in a Realistic Football Simulation: a Sample-Efficient Reinforcement Learning Approach Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T08:03:20.231886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:03:20.231886Z digest=sha256:ee92423b85445ce3f6d16025597a6dcb7726ce2714a583225caa300b7d9aadc8

Observation 1b91d14b-6869-4daa-ae7e-6b4283beb883 · inbound

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction cites this paper.

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T14:50:05.076527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T14:46:18.645009Z digest=sha256:b3307408d9e1c8178a64d70de32206d3e0386e88c1a46f1b782b422985faaf86

Observation 308ea5c0-99ae-4d98-96ca-95b6ac0ed558 · inbound

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control cites this paper.

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:33:10.238247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T19:30:23.447901Z digest=sha256:ab29aa9dcd17211be6c9b2a73f695b7a3f178bf3e649865df9feb3866a63e222

Observation 16cf2cd0-cf7f-4dda-bd87-35236b489e52 · inbound

Hierarchical Planning with Latent World Models cites this paper.

Hierarchical Planning with Latent World Models Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:18:13.756848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T20:13:58.991298Z digest=sha256:b4898665e5133fa42be476e18e08e5c2ab3cbae9846c3c887582ab4423e8d735

Observation 864bbf0c-aaa1-4dd7-b1cb-cb89dfaea35b · inbound

Abstract Sim2Real through Approximate Information States cites this paper.

Abstract Sim2Real through Approximate Information States Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:44:38.083839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T10:39:57.845600Z digest=sha256:53fd2abc3ddb68378c8999249441acb9ba8541f2946b04644448d3755c0c884d

Observation 79ac0ef4-a397-45fa-bbc5-90a7ef811000 · inbound

Self-Predictive Representation for Autonomous UAV Object-Goal Navigation cites this paper.

Self-Predictive Representation for Autonomous UAV Object-Goal Navigation Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:06:03.869664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T23:31:49.766698Z digest=sha256:b8238de64c27dab5b486addc9c3c821d9417fda84c81f644bd06c7f3704f69d4

Observation 2fd12924-e127-4769-bd31-e37d3ec5b271 · inbound

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning cites this paper.

Odysseus: Scaling VLMs to 100+ Turn Decision-Making in Games via Reinforcement Learning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 92

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:16:08.814352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-09T20:22:58.061772Z digest=sha256:dffd4ea2a01640d311b871d3b8c936f2dc332a8f33400fe041d14df198fcc039

Observation acb10449-413e-4bd2-a7be-4c65ccc9e078 · inbound

Predictive but Not Plannable: RC-aux for Latent World Models cites this paper.

Predictive but Not Plannable: RC-aux for Latent World Models Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T03:50:57.219964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T02:11:26.526418Z digest=sha256:4255b6f617058219d2b39f2f6718dde04837234d4579feec80e0e84debca2730

Observation ef9b80c9-0dc1-4021-925d-761bb39ee295 · inbound

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning cites this paper.

Multi-scale Predictive Representations for Goal-conditioned Reinforcement Learning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:16:26.310030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T03:35:37.739085Z digest=sha256:219a43d8ef0727f4f3c6a4f087cc4982c8471847a2dbe9f3757ac798741b7e9d

Observation e22b60f7-f737-45f1-bdf4-862ca67203dc · inbound

ECHO: Terminal Agents Learn World Models for Free cites this paper.

ECHO: Terminal Agents Learn World Models for Free Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T15:04:46.501267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T14:57:03.095107Z digest=sha256:6582cbf71dd45fbdf7d82f7bd4fb15806492e60f4b6d88eb525fb4c378954f70

Observation 10d7ab89-9dcf-4f5d-990d-8842368fc212 · inbound

ParkourFormer: Integrating Predictive Supervision and Sequence Modeling into Parkour Locomotion cites this paper.

ParkourFormer: Integrating Predictive Supervision and Sequence Modeling into Parkour Locomotion Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:43:59.611385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T21:37:13.833256Z digest=sha256:07c2315d1e97f518cb732529824e551f4c6710161067cc3b1940b5c8a6ca143b

Observation 5cd1a620-781d-48d1-84a2-0666f1610554 · inbound

SCALE-COMM: Shared, Contrastively-Aligned Latent Embeddings for MARL Communication cites this paper.

SCALE-COMM: Shared, Contrastively-Aligned Latent Embeddings for MARL Communication Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T17:13:44.511218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T17:12:21.761535Z digest=sha256:46be71bb31685fdbb806a21a74d902a9929be90314c39f6b4cf466523c517c2c

Observation 223507fe-f61c-4e66-b364-8134ecbf89a3 · inbound

Structured Representation Learning with Locally Linear Embeddings and Adaptive Feature Fusion cites this paper.

Structured Representation Learning with Locally Linear Embeddings and Adaptive Feature Fusion Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:57.470577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T01:24:04.887225Z digest=sha256:70f1112a59550678944443ec67b8aa9b4ba42e8c4ff28b857310be1617b8045a

Observation 015906af-65cb-4a51-b26d-6ec2cfeed68a · inbound

Direct Advantage Estimation for Scalable and Sample-efficient Deep Reinforcement Learning cites this paper.

Direct Advantage Estimation for Scalable and Sample-efficient Deep Reinforcement Learning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:19:31.238331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-26T18:12:00.111067Z digest=sha256:d266593aae5cb38300e7d047fd5ec8a0f0aa7f3f8a7a8254f4f7d9b16448ddc4

Observation b3a18fbf-78e6-49f4-a2f1-ebcb1b4f1ffd · inbound

PAMD: Structured Adaptive Distances for Bisimulation Representations in Visual Reinforcement Learning cites this paper.

PAMD: Structured Adaptive Distances for Bisimulation Representations in Visual Reinforcement Learning Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T16:26:17.964188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:26:17.964188Z digest=sha256:ca85762eb26c840a048fb6a5fd062f8b275e8fb78b439443d9c9c154425db51f

Observation 5328c020-c40e-4471-a32f-1bef29fd3f08 · inbound

TAPO: Transition-Aware Policy Optimization for LLM Agents cites this paper.

TAPO: Transition-Aware Policy Optimization for LLM Agents Data-Efficient Reinforcement Learning with Self-Predictive Representations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-31T21:44:39.632400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T21:44:39.632400Z digest=sha256:eaa41c367515df8f92a3a31483b6a25d6bfb94c18f403456cf03e1bd06c660ae