Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T23:44:37.937519Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 2 inbound Pith citation observations for arXiv:2603.10282.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-14T23:44:37.937519Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T07:49:04.825693Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-04T09:59:44.623252Z
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation b458c64e-d34e-4084-bbc3-8dc903467df9 · outbound
Update-Free On-Policy Steering via Verifiers Rt-1: Robotics transformer for real-world control at scale,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec083a6f-bb15-4d7c-b2af-3a7d0a568bf0 · outbound
Update-Free On-Policy Steering via Verifiers Rt-2: Vision-language-action models transfer web knowledge to robotic control,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 251fe879-9c30-4fb8-a078-3081649ed30f · outbound
Update-Free On-Policy Steering via Verifiers Diffusion policy: Visuomotor policy learning via action diffusion,
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4352bd92-01bf-48d6-8248-5a25c8b1c5be · outbound
Update-Free On-Policy Steering via Verifiers $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99d45fd4-3fee-4dc3-b3a3-55fd57257ecb · outbound
Update-Free On-Policy Steering via Verifiers Openvla: An open-source vision-language-action model,
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e30a4cdb-9f72-4481-8535-86653da50461 · outbound
Update-Free On-Policy Steering via Verifiers What matters in learning from offline human demonstrations for robot manipulation,
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e5db4ea-44f2-4bd4-9dbb-4b3a3387be0f · outbound
Update-Free On-Policy Steering via Verifiers A reduction of imitation learning and structured pre- diction to no-regret online learning,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5b73c3a-838c-44c6-88b8-2c2a465a8337 · outbound
Update-Free On-Policy Steering via Verifiers Fighting copycat agents in behavioral cloning from ob- servation histories,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96a4faa8-1147-44b4-88eb-7cdd86131131 · outbound
Update-Free On-Policy Steering via Verifiers Implicit behavioral cloning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3a8be01-3ea7-4dd0-bda0-397501c4f50c · outbound
Update-Free On-Policy Steering via Verifiers Human-in-the-Loop Imitation Learning using Remote Teleoperation
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b3a935e-b90a-4643-9a07-42843266dd71 · outbound
Update-Free On-Policy Steering via Verifiers Behavior transformers: Cloningkmodes with one stone,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb49b78-be5f-4876-b2cd-498b3ac10b03 · outbound
Update-Free On-Policy Steering via Verifiers Learning fine-grained bimanual manipulation with low-cost hardware,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c994b07-11c8-4b0a-93d3-b178d2bbed71 · outbound
Update-Free On-Policy Steering via Verifiers Data quality in imitation learning,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a131d0-61fa-4890-8c20-fe387551de5f · outbound
Update-Free On-Policy Steering via Verifiers Hg-dagger: Interactive imitation learning with human experts,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fe8249a-0028-45ec-a2e1-fffd94c92ce8 · outbound
Update-Free On-Policy Steering via Verifiers Dart: Noise injection for robust imitation learning,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a103dd7d-0ca2-437e-b755-e5519c68a125 · outbound
Update-Free On-Policy Steering via Verifiers Inference-Time Scaling for Diffusion Models beyond Scaling Denoising Steps
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81fd325d-c2c4-4af3-9915-e57bacce70b6 · outbound
Update-Free On-Policy Steering via Verifiers Inference-time scaling of diffusion models through classical search,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9b1c9e6-4caa-4a2f-a85b-4257d267a118 · outbound
Update-Free On-Policy Steering via Verifiers Universal guidance for diffusion models,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 136da44f-7fcf-4eb0-97a1-c3d8e39f21fb · outbound
Update-Free On-Policy Steering via Verifiers Dynaguide: Steering diffusion polices with active dynamic guidance,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d2003d0-bd8d-4867-b0ae-fd22179a5421 · outbound
Update-Free On-Policy Steering via Verifiers Inference-time policy steering through human interac- tions,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0223b433-cf1f-4b7f-9c24-937c7dce20cd · outbound
Update-Free On-Policy Steering via Verifiers From foresight to forethought: Vlm-in-the-loop policy steering via latent alignment,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 333d83c1-d898-4070-9cff-ddcff4b97d7e · outbound
Update-Free On-Policy Steering via Verifiers Fine-tuning reinforcement learning models is secretly a forgetting mitigation problem,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6768c286-cdfa-4714-aa92-7e99f1afcc3c · outbound
Update-Free On-Policy Steering via Verifiers Robocat: A self-improving generalist agent for robotic manipulation,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d19371eb-a5cb-4f21-81de-07130a9a0bf5 · outbound
Update-Free On-Policy Steering via Verifiers Sime: Enhancing policy self-improvement with modal- level exploration,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1792a51-8ef0-4483-9753-e4280d59978b · outbound
Update-Free On-Policy Steering via Verifiers Soe: Sample-efficient robot policy self-improvement via on-manifold exploration,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26de344e-0a88-4b24-b07d-a11605ab8029 · outbound
Update-Free On-Policy Steering via Verifiers Selfi: Autonomous self-improvement with reinforce- ment learning for social navigation,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bec5b23b-0178-42f0-ac97-d33b2a8c5717 · outbound
Update-Free On-Policy Steering via Verifiers Awac: Accelerating online reinforcement learning with offline datasets,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 177b5958-2f6d-4caf-bd64-c000e2a447fb · outbound
Update-Free On-Policy Steering via Verifiers Self-improving embodied foundation models,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3f34b0d-c82e-4882-8232-b0ced5d43091 · outbound
Update-Free On-Policy Steering via Verifiers Scaling Instructable Agents Across Many Simulated Worlds
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a91c47b-ddc0-4bb8-886a-204ae4ff0e0f · outbound
Update-Free On-Policy Steering via Verifiers Seil: Simulation-augmented equivariant imitation learn- ing,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d544eb13-b496-4c01-b33c-8b239665a5c8 · outbound
Update-Free On-Policy Steering via Verifiers Self-augmented robot trajectory: Efficient imitation learn- ing via safe self-augmentation with demonstrator-annotated precision,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd9fcc78-4b49-476c-a6e3-fec1e2facd9b · outbound
Update-Free On-Policy Steering via Verifiers Curating demonstrations using online experience,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfa3b500-bf5c-4b2a-950e-7c42c4aca8d6 · outbound
Update-Free On-Policy Steering via Verifiers Is conditional generative modeling all you need for decision-making?
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99916ee1-1754-44d0-ac18-3fa2ad6c56d5 · outbound
Update-Free On-Policy Steering via Verifiers Diffusion Guidance Is a Controllable Policy Improvement Operator
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc5198ba-3e43-4280-8cea-1eeee4990551 · outbound
Update-Free On-Policy Steering via Verifiers Bidirectional decoding: Improving action chunking via guided test-time sampling,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95cb89bc-bd91-4549-980e-0c8003013f87 · outbound
Update-Free On-Policy Steering via Verifiers Steering your generalists: Improving robotic foun- dation models via value guidance,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 815357fc-fe62-4997-abb1-ac14439895e2 · outbound
Update-Free On-Policy Steering via Verifiers Steering your diffusion policy with latent space reinforcement learning,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88977c50-aab6-4d56-b40a-eba76f4b4f04 · outbound
Update-Free On-Policy Steering via Verifiers Behavioral cloning from observation,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f832458-e998-4444-9632-c5872fa4ea9a · outbound
Update-Free On-Policy Steering via Verifiers Diffusion models beat gans on image synthesis,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a428a7c-712a-4176-953d-daa0d2d510f2 · outbound
Update-Free On-Policy Steering via Verifiers Denoising diffusion probabilistic models,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e9bd8bf-a950-4289-b188-a9d51f0deec5 · outbound
Update-Free On-Policy Steering via Verifiers Denoising diffusion implicit models,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c47ee81d-eab6-48a2-8679-a87ef6beb72f · outbound
Update-Free On-Policy Steering via Verifiers Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf62e0ac-6c58-417e-b5ff-85a99dfabc72 · outbound
Update-Free On-Policy Steering via Verifiers Implicit generation and generalization in energy-based models,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6c36d89-dcc2-475b-b267-023495955fb2 · outbound
Update-Free On-Policy Steering via Verifiers Attention is all you need,
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 574db489-54dc-426f-9ce1-3fc2ca63b8cb · outbound
Update-Free On-Policy Steering via Verifiers Dimensionality reduction by learning an invariant mapping,
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 243b8656-6784-4354-a18a-a96244682ddf · outbound
Update-Free On-Policy Steering via Verifiers A smooth sea never made a skilled SAILOR: Robust imitation via learning to search,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ca6822d-0c68-49c1-adfe-7083820cd263 · outbound
Update-Free On-Policy Steering via Verifiers Libero: Benchmarking knowledge transfer for lifelong robot learning,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bc3b7c0-3cff-48f0-b901-bc1f141e678b · outbound
Update-Free On-Policy Steering via Verifiers Flow matching for generative modeling,
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5505ee8e-646f-4dc9-999a-d6e526b0fb0f · outbound
Update-Free On-Policy Steering via Verifiers Unsupervised Machine Translation Using Monolingual Corpora Only
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 136942d0-f594-4853-a099-165c3cac9fee · outbound
Update-Free On-Policy Steering via Verifiers Interval estimation for the difference between independent proportions: Comparison of eleven methods,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78780e70-87c6-4ce7-8b3d-444b53de9727 · outbound
Update-Free On-Policy Steering via Verifiers U-net: Convolutional networks for biomedical image segmentation,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ffd1e20-75a3-429c-ac98-4640ad659f07 · outbound
Update-Free On-Policy Steering via Verifiers Deep residual learning for image recognition,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7b92bc1-0050-48c4-acbe-3b3b9dc122a4 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL Update-Free On-Policy Steering via Verifiers
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5cc91ec5-46c6-491c-8dc0-3db97183c7cf · inbound
Behavior Uncloning: Distilling Mode Redirection into Policy Weights without Inference-Time Steering Update-Free On-Policy Steering via Verifiers
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.