Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T04:01:18.219745Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 58 of 58 outbound references and 3 inbound Pith citation observations for arXiv:2502.07645.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T04:01:18.219745Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T05:33:10.067648Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-01T23:26:21.964199Z
58 of 58 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 964cd3fd-39e8-40e1-9d66-6a0c487874bc · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Implicit behavioral cloning
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b5e9a2a9-18f6-4870-ae7e-ae3248b98a45 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback An algorithmic perspective on imitation learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7a2b61ae-21ad-462f-b6eb-5bb707a46b7b · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Re- cent advances in robot learning from demonstration
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6e52e327-5476-42bf-a208-0babcff3e854 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback A survey of imitation learning: Algorithms, recent developments, and challenges
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4d8cd410-1bf0-4a88-b347-4ce66a16bb1c · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback A survey of communicating robot learning during human-robot in- teraction
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f0102c90-28b6-4948-9b31-15004c8c25a9 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Diffusion policy: Visuomotor policy learning via action diffusion
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation eb8c8d1e-5f02-44e9-b020-6b4a09b48835 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Conditional Energy-Based Models for Implicit Policies: The Gap between Theory and Practice
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 243dd2ad-2b1a-4ea1-b287-2df0d0d59de0 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Goal conditioned imitation learning using score-based diffusion policies
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fc40ee59-04d6-4b29-95e2-0fbb9556b9dd · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Fast and Robust Visuomotor Riemannian Flow Matching Policy
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 42588c92-d57d-4c93-ad6d-1b9097a961df · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Deep Generative Models in Robotics: A Survey on Learning from Multimodal Demonstrations
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e668243f-4b19-4a67-96b7-3299202612fe · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Interactive imitation learning in robotics: A survey
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6466329d-8bb4-4ef0-8309-3849e1621ecd · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Reinforcement learning of motor skills using policy search and human corrective advice
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ebb5835e-3f67-46bc-9850-5a563261e347 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Contin- uous control for high-dimensional state spaces: An interactive learning approach
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation be6b9431-5d04-47dc-9664-42a41742c599 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback An interactive framework for learning continuous actions policies based on corrective feedback
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ecabbe74-4cd0-461a-8196-136718e51157 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Implicit generation and modeling with energy based models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation debec787-66d9-467d-8f01-8681f52cba43 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Towards tight convex relaxations for contact- rich manipulation
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c1112610-36d7-4848-9694-c19430886e79 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback How to Train Your Energy-Based Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ddb54be8-88f9-4e3d-8291-1ec9026aea20 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Deep unsupervised learning using nonequilibrium thermodynamics
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 092d489a-dede-47e1-ae9d-c139c4ed2cb0 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Denoising diffusion probabilistic models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 79639212-d07a-478f-a7a7-752d95b8c62f · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Score-based generative modeling through stochastic differ- ential equations
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 05895d16-1754-4783-8f62-238ec16da166 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Energy-based contact planning under uncertainty for robot air hockey
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c1e222f6-0be0-4d4a-b2fe-e7c0d0ccdaa5 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Using im- plicit behavior cloning and dynamic movement primitive to facilitate reinforcement learning for robot motion planning
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 99896726-cd49-4fc8-806e-7a39687baf45 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Iifl: Implicit interactive fleet learning from heterogeneous human supervisors
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7e26d9bd-6fed-4c3b-aa50-f3694479ca96 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Diff-DAgger: Uncertainty Estimation with Diffusion Policy for Robotic Manipulation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4f779a3f-4ebb-4480-b4fc-6a46e56b45b5 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Deep reinforcement learning from human preferences
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5ad75fd2-b316-49c3-aff0-73893d665a30 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning preferences for manipulation tasks from online coactive feedback
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b3209018-ef3b-478d-a0cb-f2e5751e542c · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Pebble: Feedback-efficient interac- tive reinforcement learning via relabeling experience and unsupervised pre-training
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a0ec34fc-fedf-4edd-8ba7-b99e7ecdb6da · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning to summarize with human feedback
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 65cfb03f-9577-43b4-a7d9-3b67403c2e53 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Trajectory improvement and reward learning from comparative language feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e222255c-11a0-47dc-bb8e-15530259d435 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Contrastive preference learning: Learning from human feedback without reinforcement learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9c9c8b3d-1676-40c6-a291-6ed798be7e88 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Calibrating sequence likelihood improves conditional language generation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 73525543-999e-45bc-bc28-8bbc904b8cdf · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Direct preference optimization: Your language model is secretly a reward model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f2335496-a74b-43c1-81f4-d65f5c6190e9 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Extrapolating beyond suboptimal demonstrations via inverse reinforcement learning from observations
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a9c171dd-bafe-4655-91d9-343dd65282a9 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Batch active learning of reward functions from human preferences
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a06d8981-b0ff-42b5-b9b7-568aa68d52f4 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Hindsight PRIORs for reward learning from human preferences
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation da478840-5636-4894-bcc8-cd74dda27ed3 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning robot objectives from physical human interaction
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 64de3116-b100-45b6-96a0-503bedadf55a · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Including uncertainty when learning from human corrections
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 74b451cf-77e4-4641-a2d3-98b5d88d4b35 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning from human directional corrections
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e17b6781-8eb2-4ab9-bb3a-92467feb6b4e · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Interactive learning with corrective feedback for policies based on deep neural networks
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 98b8ad1d-1520-4a8c-b0f7-8caa720d7426 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Towards corrective deep imitation learning in data intensive environments: Helping robots to learn faster by leveraging human knowledge
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0d2477d9-1e3f-4768-a24a-2503d7cee6d5 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Interactive imitation learning in 18 state-space
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9f81d2b3-d00b-4db0-90e2-507acbb91217 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning from active human involvement through proxy value propagation
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation aa954b4c-ffd8-40f6-a8d5-34e68021ccc2 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Reinforcement learning with deep energy-based policies
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 25705914-2d45-4510-aa97-8b2063ef7060 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 55c9025f-9c63-4bca-a578-251354d22685 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Aligning human intent from imperfect demonstrations with confidence-based inverse soft-q learning
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3d2a8f9a-4dff-4cf4-812c-6bf2d94286fc · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Bayesian reparameteri- zation of reward-conditioned reinforcement learning with energy-based models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e5dbdc9a-fb2c-4090-9407-851c218d5b01 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Inverse preference learning: Preference-based rl without a reward function
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4a03b651-0daa-4acb-80aa-be86c190394a · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Learning from interventions: Human-robot interaction as both explicit and implicit feedback
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1f037782-f950-4156-a6d3-9ed14b01aa8d · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Flow contrastive estimation of energy-based models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 770ffb4f-80b6-4a50-8f1c-a06bcd9a1c38 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Hard negative mixing for contrastive learning
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a738eff6-1346-4475-adc3-bae0ff835626 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Representation Learning with Contrastive Predictive Coding
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c0a7a3bf-31e8-46b7-aa1d-4129418c2e1a · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Revisiting energy based models as policies: Ranking noise contrastive estimation and interpolating energy models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9190a2c3-5710-4598-92e9-5f3212d7a73d · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback A reduction of imitation learn- ing and structured prediction to no-regret online learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7f986c3d-15de-40c1-b223-7a2a946e923f · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Hg- dagger: Interactive imitation learning with human experts
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5c93ad7d-28c3-4f4e-b295-26d68cdb257a · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Ambient diffusion: Learning clean distributions from corrupted data
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 94a20fcf-5862-414a-9899-71cf91debbf9 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback robosuite: A Modular Simulation Framework and Benchmark for Robot Learning
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3f9023d4-a2b2-4504-89e5-5742e15f9c2a · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Interactive learning of temporal features for control: Shap- ing policies and state representations from human feedback
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3b05a8d8-7cfc-482d-9554-525e7165c0c5 · outbound
From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback Bayesian learning via stochastic gradient langevin dynamics
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b9f6d8ea-b5a9-4042-893e-8e764753b7b3 · inbound
Wavelet Policy: Imitation Learning in the Scale Domain with World Prior Memory From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8ab9caa0-8478-4c6a-891e-f52a3913ffd6 · inbound
CLASS: Contrastive Learning via Action Sequence Supervision for Robot Manipulation From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a4d691f-7451-45c8-9af4-c098c2b28288 · inbound
Set-Supervised Diffusion Policy: Learning Action-Chunking Diffusion through Corrections From Action Labels to Sets: Rethinking Action Supervision for Imitation Learning from Corrective Feedback
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.