Pith. sign in

Paper Citation Record · LEDGER

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement

As of 8 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 0 inbound Pith citation observations for arXiv:2507.06701.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.06701 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:06:06.951316Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a27a6172-f7a0-46cf-b531-e4a07e871936 · outbound

This paper cites Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Wish you were here: Hindsight Goal Selection for long-horizon dexterous manipulation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.116005Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.116005Z digest=sha256:32ce2d7d279008070e81a41c357360f3196c2aedfe74be549bda65efce01304c

Observation 35e2ba22-e5ea-4da3-aa74-c9507fab7691 · outbound

This paper cites A Generalist Agent.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement A Generalist Agent

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.602130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.602130Z digest=sha256:ef69714b49fb2844e523c2b09a414ffdd3d53c947342ad40975dfb6cd8e3a3e2

Observation 1d3d755d-d3b6-4273-a0a9-604db3f09052 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.444819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:06:06.743130Z digest=sha256:1f620fc9b90cf39f06471a84f11d4f499711c08200b0a8cc66f471a6c89a43cf

Observation 99c2d47e-7fff-4ee7-b1fd-7501c86f79c6 · outbound

This paper cites Offline Learning from Demonstrations and Unlabeled Experience.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Offline Learning from Demonstrations and Unlabeled Experience

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.902037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.902037Z digest=sha256:2d381a2960490efb36780e82c8e85958ecd1d69749c32cdeeb4bac51bd9cf565

Observation 9c306700-d951-45f0-8fea-32d6a439cce9 · outbound

This paper cites an unresolved cited work.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:06:07.318510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:06:06.951316Z digest=sha256:91ca591dba5119e4fb93fc7c3b5a2cbc7f4703a5589d8a511f3e5bd92084480c

Observation 379b2b3e-48a3-4756-a055-b3f92a1131a7 · outbound

This paper cites Large Language Models Can Self-Improve.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Large Language Models Can Self-Improve

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.248583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.248583Z digest=sha256:6c21258e88f3feee97a833ba17ba7d8ea9701b6dae3036ee11e39fde6165fd7e

Observation 29bfdde9-336a-4003-b743-3b7accb25f0d · outbound

This paper cites Maximum Entropy Deep Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Maximum Entropy Deep Inverse Reinforcement Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.805435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.805435Z digest=sha256:aa73dcfc820e170e01367a4f606b9eebc66b86e8d3d280ac4ca55ca86c3c769d

Observation 4a047633-9609-4812-8932-74bf9035e92d · outbound

This paper cites Kwon, T., Palo, N.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Kwon, T., Palo, N

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.454510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.454510Z digest=sha256:0bdbdd8927efc3364620a330c8daeb0137db59b8337cc9929b703c8ec6dbd06d

Observation 98b7e95c-a3b6-4886-ab6d-1702aa029fe1 · outbound

This paper cites Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Actionable Models: Unsupervised Offline Reinforcement Learning of Robotic Skills

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.043894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.043894Z digest=sha256:5f78e4652a8d64245d50737e4fc5abafe36fb3272fe781d186d865f1090aeafa

Observation a695a859-fe68-432b-9b3f-443714294e47 · outbound

This paper cites Imitation Learning via Off-Policy Distribution Matching.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitation Learning via Off-Policy Distribution Matching

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.332378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.332378Z digest=sha256:8e8e5d432c917b9f2f7e32e8301fe4f4b8f0206da7b2ebaf3e94c85ddbcbbf78

Observation d3dbb16b-69ea-4a80-9af1-4462da4e71b3 · outbound

This paper cites Imitating Language via Scalable Inverse Reinforcement Learning.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement Imitating Language via Scalable Inverse Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.864052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.864052Z digest=sha256:7ca3ce1892d8af8ecd64baba85866b7bdba1122822a7e830afb8d569820fd04f

Observation a681f6c3-2c41-460e-9cd3-9279e8e30a02 · outbound

This paper cites org/CorpusID:269605913.

Value from Observations: Towards Large-Scale Imitation Learning via Self-Improvement org/CorpusID:269605913

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:06.677382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:06:06.677382Z digest=sha256:9f94b331e7844881397395fb9ea374f2fb3bcc0491c3349b3f242ee8e815df80

Pith citing papers

No inbound Pith citation observations are available.