Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:37:10.302475Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2411.17764.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:37:10.302475Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-21T21:57:15.285757Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T22:00:41.924058Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19ecd7f7-c150-4da4-9afa-467d9eef3a44 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 11dd6156-200d-4326-8beb-5c8979e7f81a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Learning reward functions for robotic manipulation by observ- ing humans, 2023
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation b83cd651-b7cc-41d5-b329-d6d0bcc33a1c · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Human-to-Robot Imitation in the Wild
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c812fbd-383b-46b2-9b55-153ca65ac008 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Video pre- training (VPT): Learning to act by watching unlabeled online videos
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e05c97db-108e-459b-97a6-8d1941bf3494 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement The perils of trial-and-error reward design: Misdesign through overfitting and invalid task specifications
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bb7f932b-2ae9-4327-83b2-b5baf845e580 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Scaling ego- centric vision: The EPIC-KITCHENS dataset
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ee15b13c-1732-4860-8815-935e210c386a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Rescaling egocentric vision: Collection, pipeline and challenges for EPIC-KITCHENS-100
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 82e1517c-2e2f-4a9e-8371-aa9d3e544c37 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Video predic- tion models as rewards for reinforcement learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 711d2e16-250e-481b-b1eb-14581d5bf553 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Contrastive learning as goal-conditioned reinforcement learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 908bc309-970b-4a76-a7ae-68e01e0dd785 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Guided Cost Learning: Deep Inverse Optimal Control via Policy Optimization
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944a3854-6209-45e7-8593-b15ea686d2a2 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Learning Robust Rewards with Adversarial Inverse Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da74d825-d443-43a8-bd2a-f30e958f31b7 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Domain- adversarial training of neural networks
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 963ddfdb-43cb-4490-ae74-5e8ac04635fb · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Generative adversarial nets
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5f248316-4948-4d16-b633-39295f429797 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Ego4D: Around the world in 3,000 hours of egocentric video
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0c25a2c0-35f0-4577-acf5-a791b0b5542b · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Inverse reward design
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation de714324-14be-4d2a-87e0-ef81573709ea · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Deep residual learning for image recognition
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d94d9fea-05f5-42da-a25e-4c1a199c4a68 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Generative adver- sarial imitation learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5d735659-4254-420a-b394-07e71a116252 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Generative Adversarial Imitation Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4177153f-b177-4552-a35b-723fa3cd5a81 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Diffusion Reward: Learning Rewards via Conditional Video Diffusion
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba02e57a-0222-4fd8-99ee-4cd0ed8209a7 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Auto-Encoding Variational Bayes
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b900e1a-230b-4a4d-ba7b-797012caff4b · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement InfoGAIL: Interpretable Imitation Learning from Visual Demonstrations
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc06a059-759c-403d-bd36-5874af830ede · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adba2c85-f6dd-4350-bf1c-3c1b7d81d907 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement 9 Vip: Towards universal visual reward and representa- tion via value-implicit pre-training, 2023
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 04eab034-e282-4624-9b20-0fd582b6e75a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Towards Theoretical Understanding of Inverse Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 246b6b1a-a689-4a57-aa9b-a2a9b7a97eb5 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement R3M: A Universal Visual Representation for Robot Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 004c2f5f-ee50-415c-94cb-c222c43dc6b1 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecd888a8-d3b6-4ca2-bbc9-1fbec814783c · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Reinforcement learning by reward-weighted regression for operational space control
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 357bb53a-dc1d-43ca-8bbc-c6a30e6aa68e · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement DexMV: Imitation learning for dexterous manipula- tion from human videos
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 645b6f49-f943-47a1-9e58-929e12831184 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Artificial Intelli- gence: A Modern Approach
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation e2b7d4e4-2577-413c-b68b-9dcdcaf96fed · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Time-contrastive networks: Self-supervised learning from video
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6cadf0c8-47d7-4eba-a540-378e34d050c8 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Time-contrastive networks: Self-supervised learning from video, 2018
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 15a2298d-e747-4f54-ba67-b433d350bbf6 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Lewis, and Andrew G
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 6d51ecbd-82a2-4075-97eb-1b64fc379c50 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Reinforcement Learning: An Introduction
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 936a2caa-1bd3-4ba7-a870-c0edf194c34f · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement DeepMind Control Suite
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5cd4583-7686-4fab-b620-431b38323e79 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Wozniak, Andrea Gasparri, and Danica Kragic
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c8c79295-a259-4903-b880-ab88d7646851 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Maximum entropy deep inverse reinforcement learning, 2016
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 996fff09-c9d9-4953-846a-8b72577ef331 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Rank2Reward: Learning Shaped Reward Functions from Passive Video
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1970e3be-72f5-45f6-9243-e36475cf76c4 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Representation Matters: Offline Pretraining for Sequential Decision Making
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a483121f-877f-4c80-abf5-f4462187d123 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Im- age augmentation is all you need: Regularizing deep reinforcement learning from pixels
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 48f19083-4e99-40ac-9f03-7518cbbe2b43 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Meta-World: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 9e42a5d4-a8b3-4084-a30d-bea6d3ddc0b0 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Learning to drive by watching YouTube videos: Action-conditioned contrastive policy pretraining
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c9badd64-d63e-4c23-ac7c-b2f32411452a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3dec94b-e16f-4063-979e-2dd7f4cabf69 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Ziebart, Andrew Maas, J
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4010f532-83d2-4b61-a1d5-8994dd33c5ff · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 59b62d96-d08e-40bc-b3a6-5d739e0ac71a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation dc209de8-e71d-49c8-b68d-be4ee4327b38 · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Robotic Experiment Setup The real-robot experiments are performed using a Universal Robots UR5 robot arm equipped with a Robotiq 3-Finger Gripper (Figure 8)
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d005df41-9d71-4167-b01f-87516338559a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement The case of β = 0(PROGRESSOR with- out Push-back) is discussed in the main paper
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5422f123-94a9-4e8a-b7a6-299b5d1cba3a · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement This figure serves as an extension to Figure 7 for complete- ness
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d855ac28-3147-47c0-9bba-c525d45f80fb · outbound
PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement Unresolved cited work
Reference 2004
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7eaceeb8-4290-46cd-aa61-6ea87d041e5b · inbound
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance PROGRESSOR: A Perceptually Guided Reward Estimator with Self-Supervised Online Refinement
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.