Pith. sign in

Paper Citation Record · LEDGER

Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2303.01488.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2303.01488 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-30T15:27:16.820824Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T12:28:07.239829Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45fa6957-2665-4124-b484-043ae071944e · inbound

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning cites this paper.

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T23:46:32.301737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:46:32.301737Z digest=sha256:7547c4328d416e007effc99970d087d4c528fe8cd199ac9d00d8c55138b3bd2e

Observation d50584b1-f0dc-4ca5-a381-81e28b556d5b · inbound

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning cites this paper.

ViVa: A Video-Generative Value Model for Robot Reinforcement Learning Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:25:59.129860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:12:08.970164Z digest=sha256:a50c9b31512a037cac749565c8f9c7e09ff6fe07eb6491a48a47be3f1509062c

Observation b52f1d98-2a82-4d4b-9351-362aef7d8bde · inbound

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence cites this paper.

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-13T00:03:53.609175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T00:03:53.609175Z digest=sha256:65eb6a068878081b1a7d7b9c687ff5702449602c1c7d1221c3facc6afbd009b4

Observation a71df6cf-c6ff-4ca7-9d35-a20716002618 · inbound

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon cites this paper.

VLA-Corrector: Lightweight Detect-and-Correct Inference for Adaptive Action Horizon Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T12:28:07.241567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T12:20:11.540651Z digest=sha256:1fa1e1403e4218c6c8224df90ee257d1db5544f186cbb8bfe79e467aa8020edf

Observation 6b74c874-491d-4805-93b7-6d6604e54b54 · inbound

Try Once, Then Optimal: De-Redundified Procedure Memory for Cross-Episode Exploration Amortization cites this paper.

Try Once, Then Optimal: De-Redundified Procedure Memory for Cross-Episode Exploration Amortization Self-Improving Robots: End-to-End Autonomous Visuomotor Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-30T15:27:16.820824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T15:27:16.820824Z digest=sha256:cd414d25d4562dd8cc6c61197040d96a0428cd961200a301994edbed573c6de0