Pith. sign in

Paper Citation Record · LEDGER

Real-World Offline Reinforcement Learning from Vision Language Model Feedback

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2411.05273.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.05273 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T16:10:12.689957Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:59:44.651976Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ae592448-af17-411a-9e61-6674e22f573c · inbound

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning cites this paper.

SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning Real-World Offline Reinforcement Learning from Vision Language Model Feedback

Reference 55

Resolution
unresolved
no resolver link, observed 2026-07-13T16:10:12.689957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:10:12.689957Z digest=sha256:d2e3b4c01e9e65c3ed6ace9d565cd9a7245cfee0120646144d15c2ec9fc973da

Observation dc91aa6c-be91-4f74-8d4a-f7b956d7035e · inbound

Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations cites this paper.

Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations Real-World Offline Reinforcement Learning from Vision Language Model Feedback

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-22T05:34:40.299502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-22T05:32:17.780012Z digest=sha256:7bc336bc416c1393eab8a45c73491d758f585573a8b8cf4fd9d64bb3d346f851

Observation c0ad28e5-6960-412d-a405-ab12c7b0422d · inbound

Learning Process Rewards via Success Visitation Matching for Efficient RL cites this paper.

Learning Process Rewards via Success Visitation Matching for Efficient RL Real-World Offline Reinforcement Learning from Vision Language Model Feedback

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:59:44.653591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T09:20:35.062060Z digest=sha256:076a40cd625f683ca4942aaaf4c207b963a69fa114f449321910e521feb7da0a