Pith. sign in

Paper Citation Record · LEDGER

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation

As of 11 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2606.24633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.24633 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T12:30:30.343964Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ace7589c-730d-456a-be09-383ed9a72a9e · outbound

This paper cites and Ng, A.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation and Ng, A

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:046371a7558ed4a22fda41f105aa0d60b61a46b8a4e53f4aec0b0522a17c0b29

Observation 0c04440d-91ba-49a3-b6b4-2ca4814c94f8 · outbound

This paper cites Video- language critic: Transferable reward functions for language-conditioned robotics.arXiv preprint arXiv:2405.19988, 2024.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Video- language critic: Transferable reward functions for language-conditioned robotics.arXiv preprint arXiv:2405.19988, 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:7fbe1637fd5f9a6c3903d3ea8b09cc6dcebf06381f8208329456823e8f18f99b

Observation 0f09f7c4-0e16-4960-a2ed-22f23c9bfbd0 · outbound

This paper cites $\pi^{*}_{0.6}$: a VLA That Learns From Experience.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation $\pi^{*}_{0.6}$: a VLA That Learns From Experience

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:3e545bb5de70c64fd24160fbb8c9a86c564d0b33fbf4a796972f9ccb8480ddb8

Observation 97defe0d-ee31-48e0-bdb8-3d6284358b5d · outbound

This paper cites Qwen3-VL Technical Report.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Qwen3-VL Technical Report

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:7b1c2a840f6aa5fac64496a7d3a9d597eee68e7641de6cbd5d8697a8bf49e6eb

Observation d3850b1f-2ddd-486f-99d4-f315dc8c43fd · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:ef6c13871338c43a2a14b34c8eb36362c4e42ba78b91f412bd11b8dcf60c9c38

Observation 7892df81-7c3a-49ce-8bd4-800c375b701c · outbound

This paper cites an unresolved cited work.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:022e8beedb70609b7c7747b2fba89730f849e8face77ddde0c5dcb193ac205eb

Observation a43d8a70-d384-4b7a-97ed-7f5998e65674 · outbound

This paper cites RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:ef69cff3d73f9b08f981e095f9a05dd7587422ebe08cf20e0299148445fd2463

Observation 706630a9-4bd6-43b4-80f8-7e5330057c1a · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation RT-1: Robotics Transformer for Real-World Control at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:d1bc62b8e78a60429620ba0c249a1719e7b68633fad226d2c0f0b5b34e5733b2

Observation 5686a67d-fa47-42f9-bece-556852ed2576 · outbound

This paper cites S., Goo, W., Nagarajan, P ., and Niekum, S.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation S., Goo, W., Nagarajan, P ., and Niekum, S

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:5a3ad48a9541e3f631b5f77a1c50fcb86124745274c2128287bca0c0bbb2dc9c

Observation bfd25655-aa02-463b-a041-36b7eda252d1 · outbound

This paper cites S., Goo, W., and Niekum, S.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation S., Goo, W., and Niekum, S

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:bc1b6414e76ebefc8dfc41c0320e60362d80d328503e8bae9a8c8ca103435927

Observation eb3bf747-7bf2-4589-ba75-8c70f5dbffe7 · outbound

This paper cites AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:4b25849ec68aad8fe38b19d42e99e933fc17514da56031343d7c49eba30c9dce

Observation 92c95d16-81dc-4f75-a65a-5b39f35f5938 · outbound

This paper cites in-the-wild.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation in-the-wild

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:2867d39bb0c39546d2958c29cf8b840d4ef8adf75e511776c0d58a949ff3e436

Observation 11c85dd2-be5a-40b3-8777-b7cf0a266560 · outbound

This paper cites SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation SARM: Stage-Aware Reward Modeling for Long Horizon Robot Manipulation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:9be2d60783f79f3e213799e9a831045b3f02055fa7494fd9a8400d778815b693

Observation 52ef0489-4b15-48e6-9672-7e7f09a30710 · outbound

This paper cites J., Ren, Z., Ratliff, L.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation J., Ren, Z., Ratliff, L

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:a23a977f228275497a0afadcb22f556ed85182c8a1de704878a4a6798c7bdea6

Observation dd2906fa-6342-47e8-b223-603786c674e7 · outbound

This paper cites villa-x: Enhancing latent action modeling in vision-language-action models,.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation villa-x: Enhancing latent action modeling in vision-language-action models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:0429b95dc1e4027036c579725822a77d78518d9b46d43b9b594def6d4a4b35f1

Observation 1a01366b-5cef-4dd4-a157-9cd3764d51c8 · outbound

This paper cites villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation villa-X: Enhancing Latent Action Modeling in Vision-Language-Action Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:bb90bce2c77cfe27674f47c2346ad54d0ee8b3d2169d3e9d0576510af399b1c8

Observation 5d0437b3-b89d-4eb6-ad0e-bab2f0f17ae5 · outbound

This paper cites F., Leike, J., Brown, T., Martic, M., Legg, S., and Amodei, D.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation F., Leike, J., Brown, T., Martic, M., Legg, S., and Amodei, D

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:f9a752fe6f022d5f1cdd45cd24ac26b5dd8a6690e828311c638bd8f8861da9a6

Observation c1eea95e-ef41-4908-82f1-6199a4c0d6f4 · outbound

This paper cites F., Leike, J., Brown, T.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation F., Leike, J., Brown, T

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:ef51a7bd2437aeb9cea0c19fc1b89c41fb618011c0db1cad766946697a59c09d

Observation 5e4bca3d-8cc9-49ba-9d02-18afe07802ac · outbound

This paper cites Guided cost learning: Deep inverse optimal control via policy optimization.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Guided cost learning: Deep inverse optimal control via policy optimization

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:3e211c06218856030863839cc23db36e1d919e7e53186705b0521f9d6481ded5

Observation 2c2097b0-5df5-4119-8f04-d7d547b11ed2 · outbound

This paper cites Awr: Adaptive weighting regression for 3d hand pose estimation.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Awr: Adaptive weighting regression for 3d hand pose estimation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:cd5b8c65e7fa3b23222e9a0b8f26c6501e821e03cd0a78d24782522dafdec9b1

Observation 656a11d9-4747-4d47-bd46-94c13a3535d6 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:36a440466a7ae57d24cf29ef1169586afdcec4f56c1345951342791fc4876a61

Observation d80ff42a-9493-4239-b43f-01db1f169ef7 · outbound

This paper cites VIMA: General Robot Manipulation with Multimodal Prompts.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VIMA: General Robot Manipulation with Multimodal Prompts

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:55f33249515819450896520a1a7a90aef044cfdde5062945cde2b5e93cd54f4a

Observation 125d4e37-966a-4bd5-af77-4c2bd580b06e · outbound

This paper cites an unresolved cited work.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:4e16f46976becae8769dd23de07ba15d521c55913eca79bb87d354b4cf232cfa

Observation 2b63f625-0be3-4fbd-aeb7-6a80a701323b · outbound

This paper cites Demodice: Offline imitation learning with supplementary imperfect demonstrations.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Demodice: Offline imitation learning with supplementary imperfect demonstrations

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:cb9e26d0cb538ba865c71c649d1e4adf77007ecb155e314cb409197c788504a4

Observation 09e5c63b-d132-4748-86fb-d415e3fabde3 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation OpenVLA: An Open-Source Vision-Language-Action Model

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:da3e3d87501876825dd4e39991804f941216beb48a1ecdf41cee869e76ecbe2a

Observation 161cf680-1135-4ec1-a8ff-567063c259d3 · outbound

This paper cites DART: Noise Injection for Robust Imitation Learning.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation DART: Noise Injection for Robust Imitation Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:86977d35ab72a7607e8cc3ba47a1b8c67b4bbae3492d66567184072462297e6c

Observation 7ce8d7b1-8dab-4563-93ae-811f2113bbc2 · outbound

This paper cites Roboreward: General- purpose vision-language reward models for robotics.arXiv preprint arXiv:2601.00675, 2026.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Roboreward: General- purpose vision-language reward models for robotics.arXiv preprint arXiv:2601.00675, 2026

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:2ea559bb23ce677ec51cebfad64b7abb1bc85ab8065ffa032ddda74078d58bc1

Observation 7a3258de-2525-4a15-9e57-d0506bcccfad · outbound

This paper cites Bridgevla: Input-output alignment for efficient 3d manipulation learning with vision-language models, 2025.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Bridgevla: Input-output alignment for efficient 3d manipulation learning with vision-language models, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:e24f27ce79c0d79cc90b1dc5dc2fcd054651610a2df4cfdd31543e1480098865

Observation 79fae587-cfd9-4504-8db6-2cc2ee36ce3b · outbound

This paper cites Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.arXiv preprint arXiv:2512.01801, 2025.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Gr-rl: Going dexterous and precise for long-horizon robotic manipulation.arXiv preprint arXiv:2512.01801, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:152ddc7e924def1249311e217a985aa71debc78141f0985eccb77103102a9e8c

Observation 54d95685-0b36-433b-831b-7a183e34f551 · outbound

This paper cites Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:baeb423d3bd3674900a0c10b6279a1398720e2e3b5be8407338e4ba43c9b7489

Observation 0ac7198c-b9ac-4522-8439-cbab5e3b4681 · outbound

This paper cites RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:ffae9713a9e92601a57574336e47f535ed1a15e0df53bb205e5469b6b729c6b4

Observation 00d8c645-ad61-4659-89ef-ac542031fa31 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:a4ffd3dd96cdf2a537d1c1f7268a32b49c389b3d9ec576a56dde2729e072371b

Observation 5f290982-b9b1-43db-8e04-e1be75755d7e · outbound

This paper cites J., Liang, W., Som, V ., Kumar, V ., Zhang, A., Bastani, O., and Jayaraman, D.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation J., Liang, W., Som, V ., Kumar, V ., Zhang, A., Bastani, O., and Jayaraman, D

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:30834b9fcf1c15249a9224582603fe26b9f796ce005ede50fab3c1165fa27915

Observation 49ae4da6-514f-4de5-a04c-4e9af9525754 · outbound

This paper cites ARM: Advantage Reward Modeling for Long-Horizon Manipulation.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation ARM: Advantage Reward Modeling for Long-Horizon Manipulation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:b312a6074a7e3fa4bc370172ddf6a5c0cbab42938913774fa7f532c19c1c39f4

Observation 59290d0f-9db2-4cf9-9fdc-ec9e6e8f2570 · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:a9bd63c2bce3e23d97e0caa7366472aebd36bbfae884043f99c29ff753a015fd

Observation 3e50c19c-a1bd-4263-b8bd-f2819d3eb9a4 · outbound

This paper cites an unresolved cited work.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:bded39990a6ca9386ab6c53d4aa0730d8be16b81b59733f48b1d06539d0bd9ab

Observation 33927d87-2415-4034-be0f-bffaf3d99703 · outbound

This paper cites C., Shevchuk, G., and Sadigh, D.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation C., Shevchuk, G., and Sadigh, D

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:7e2acab36120acf0a02273dff4cefff10c727f50f8f9c9f556ddef6c39e0079e

Observation 314d1b5d-376b-4f9d-89d3-7c8f5b5005b6 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:85c5dccfe14b630a9d6ae16965eccf44c3b7e0e8224c136eca1f5b2002f8c293

Observation 89e0a031-5eb3-40b8-8130-90ab90743394 · outbound

This paper cites D., Sastry, S.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation D., Sastry, S

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:98acc8c549576c72887c32353d8bdc41aec099ca258885d0adc28b6e9a7dcf73

Observation c0f0e566-7447-4899-a7e8-860a8471c2b0 · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:44f54d70e3782f573e7f3c196027bc9dad428fc945906d16dd94ac9ff7624195

Observation 802b361c-610c-4bbf-b581-dea9e2abbe77 · outbound

This paper cites Robo- dopamine: General process reward modeling for high-precision robotic manipulation.arXiv preprint arXiv:2512.23703, 2025.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Robo- dopamine: General process reward modeling for high-precision robotic manipulation.arXiv preprint arXiv:2512.23703, 2025

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:0331193dbb82a90e6dbbab61424822a0d1750cf8460ab74a9b1e32243fc0860e

Observation a1f57171-794d-4ad8-89f6-de15535d2472 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Octo: An Open-Source Generalist Robot Policy

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:c389d9832e9eaa41e5aac5c4e06a5b4498d5f46f74ea144a150826895a5fe28c

Observation 4d6aa5ce-235d-4267-a8e8-747698a5f2d0 · outbound

This paper cites DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:e99dca973069f2bb8e39823e2de44d340c0525c49d094ba180cab502aad9b03c

Observation 3eded503-c32d-4c40-8245-d6d42669d4e3 · outbound

This paper cites Imitation Learning from Imperfect Demonstration.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Imitation Learning from Imperfect Demonstration

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:fff23babd4234764e2be0478b1253d8045beeb9d7395eb1debff6f42e257ad61

Observation 533c2f52-62c8-42e2-a976-2c3a929bb956 · outbound

This paper cites Imitation learning from imperfect demonstration.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Imitation learning from imperfect demonstration

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:01146f6b82819bd130859b18179464122f9cc3e834ececc9b16950c334b5ae00

Observation a7d71092-d6b9-487d-9ea9-582477c9e2a3 · outbound

This paper cites Discriminator-weighted offline imitation learning from suboptimal demonstrations.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Discriminator-weighted offline imitation learning from suboptimal demonstrations

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:27fda3673080721e560b33b51f22ec18606e40c93e77c554415d19984364cddd

Observation 336a06ef-28e6-42a1-a98e-397e5012b31c · outbound

This paper cites Compliant residual dagger: Improving real-world contact- rich manipulation with human corrections.Advances in Neural Information Processing Systems, 38: 139559–139581, 2026.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Compliant residual dagger: Improving real-world contact- rich manipulation with human corrections.Advances in Neural Information Processing Systems, 38: 139559–139581, 2026

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:2365a066a446349361e89b3f9d3795b1e607c81e9560c1d6bcb371ac56827fd6

Observation aac45401-531f-4b8a-87de-33b212ec2d1e · outbound

This paper cites RISE: Self-Improving Robot Policy with Compositional World Model.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation RISE: Self-Improving Robot Policy with Compositional World Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:bbdb35d43e2fa5b13f6b8a87caae065b17ddb3acc952601ee345365a5858ccb5

Observation b02d2d26-146d-4627-b0b8-80a66061f62d · outbound

This paper cites ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:4ad1dea012b381f68a0c45735a2186f804deccf4cd7ffc85040e4c4b20e3c41d

Observation 5b1fb27c-0926-49bd-9336-684ff68f7f8e · outbound

This paper cites Confidence-aware imitation learning from demonstrations with varying optimality.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation Confidence-aware imitation learning from demonstrations with varying optimality

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:9cc9790a61306fcc80ab15923d3e6be638b551da2dfe5b8be9c7c0dcedaeb37d

Observation d0f4b36f-b5a1-4bcb-ae7a-03fd37fc19de · outbound

This paper cites VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation VLAS: Vision-Language-Action Model With Speech Instructions For Customized Robot Manipulation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:d8299a954de907f19ed3a7376558219cb73bfd56465cb84e0db0977d34412812

Observation ea73a968-811c-4fbc-97ba-44d9d0ebfa8b · outbound

This paper cites D., Maas, A.

Beyond Monotonic Progress: Retry-Supervised Value Learning for Robot Imitation D., Maas, A

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T12:30:30.343964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T12:30:30.343964Z digest=sha256:cd2a87a4494b48379cfad0f86b435c1e82208dc250f04c127861664f54ece101

Pith citing papers

No inbound Pith citation observations are available.