Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:59.972964Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.04730.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:59.972964Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
26 of 26 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7d324a6c-9d4a-45b2-936b-55379f7455ae · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Learning Complex Dexterous Manip- ulation with Deep Reinforcement Learning and Demonstrations,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 275b77ec-66e8-4bb4-8d72-6d13e66314f9 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep q-learning from demonstrations,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2d57bf9a-6545-4fc5-be86-c6979e74f76e · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Action advising with advice imitation in deep reinforcement learning,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e2361b3e-3fde-43c7-a388-63ab0619c523 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Dqn-tamer: Human-in-the-loop reinforcement learning with intractable feedback,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dda975d2-e957-4c35-b55f-5cd2c1ae4252 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Combining manual feedback with sub- sequent mdp reward signals for reinforcement learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ee2d89a7-e9a9-4d9b-ab0d-6fb0937dee98 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79ff446e-8da7-4570-8fae-dcbd72a0333b · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning with corrective feedback for policies based on deep neural networks,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4e9d4548-a904-42cf-901c-2dbf6c0cd36a · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning of motor skills using policy search and human corrective advice,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82101eaf-4938-4dc9-8a2a-bec3b2fdefd3 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback No, to the right: Online language corrections for robotic manipulation via shared autonomy,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 147de395-621b-457c-a72f-cab7571b85ec · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Yell At Your Robot: Improving On-the-Fly from Language Corrections
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4283aa42-c4de-43e0-b529-6e8cb74fc9f8 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback An interactive framework for learning continuous actions policies based on corrective feedback,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 935d72a3-3053-488f-8139-1c70e4f5a339 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Algorithms for inverse reinforcement learning,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d80a485-d838-4d3f-821e-7b0a46936b8b · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactively shaping agents via human reinforcement: The TAMER framework,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff3522ad-c95c-4bdf-8712-fad249077b2e · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning from policy-dependent human feedback,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f5a78cb-696a-4f5e-af77-b26a72e8fb48 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning from human preferences,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 444caa52-0700-4b11-bc43-9d1ed8f5799b · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Few-shot preference learning for human- in-the-loop rl,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ff089631-31ee-44d3-b19f-88e63eda4736 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Integrating behavior cloning and reinforcement learning for improved performance in sparse reward environments,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c843d469-34c9-46df-b213-c841243fe7fa · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Agent- advising approaches in an interactive reinforcement learning scenario,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0e479ed-1cf6-425e-8774-279f96601423 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Modem: Accelerating visual model-based reinforcement learning with demonstrations,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 00d80b20-ee01-4682-8da9-061fd5a1db2f · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive robot learning from verbal correction,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 367dae9a-9353-44d3-8ad8-533c927286e6 · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback A reduction of imitation learning and structured prediction to no-regret online learning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f182bb5a-b9af-4189-93f1-5341d9f71a8b · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Orbit: A unified simulation framework for interactive robot learning environments,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 441a8885-0384-41ee-98cf-3ecda6a3a4fc · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Anymal - a highly mobile and dynamic quadrupedal robot,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1de215f1-4210-429d-8fea-5d79092ad54e · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning with double q-learning,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba101753-c005-411e-ad8a-d94ef21d471f · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Probabilistic roadmaps for path planning in high-dimensional configuration spaces,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af2ff169-3898-4f90-a720-2dc33086bd1c · outbound
CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning: An introduction,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.