Pith. sign in

Paper Citation Record · LEDGER

CueLearner: Bootstrapping and local policy adaptation from relative feedback

As of 8 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.04730.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.04730 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:59.972964Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7d324a6c-9d4a-45b2-936b-55379f7455ae · outbound

This paper cites Learning Complex Dexterous Manip- ulation with Deep Reinforcement Learning and Demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Learning Complex Dexterous Manip- ulation with Deep Reinforcement Learning and Demonstrations,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.954722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:56.890664Z digest=sha256:5c4217e3d7b30e095443000d61193380c30010a0bb15996c88b7bdb506be64d4

Observation 275b77ec-66e8-4bb4-8d72-6d13e66314f9 · outbound

This paper cites Deep q-learning from demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep q-learning from demonstrations,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.809877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:56.972368Z digest=sha256:9ecb0e2bfa00ff06b7e632aca6a535c926faaf29e1ca832f324700fd3dd687de

Observation 2d57bf9a-6545-4fc5-be86-c6979e74f76e · outbound

This paper cites Action advising with advice imitation in deep reinforcement learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Action advising with advice imitation in deep reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.689134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.065963Z digest=sha256:2eed869df9c1b52d31bf4a8effe31cfc81792af742f993867d7d20fe4e3b7716

Observation e2361b3e-3fde-43c7-a388-63ab0619c523 · outbound

This paper cites Dqn-tamer: Human-in-the-loop reinforcement learning with intractable feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Dqn-tamer: Human-in-the-loop reinforcement learning with intractable feedback,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.508637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.188486Z digest=sha256:275ea2b127f8191d3f6f0277de5de7b804b0d90973ff9e7da88fa20299354a26

Observation dda975d2-e957-4c35-b55f-5cd2c1ae4252 · outbound

This paper cites Combining manual feedback with sub- sequent mdp reward signals for reinforcement learning.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Combining manual feedback with sub- sequent mdp reward signals for reinforcement learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.329570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.290909Z digest=sha256:54507a2ee56a54bf65f0093bf579c17aefb3110d0a6fefbf9ec688ba1f38e0b9

Observation ee2d89a7-e9a9-4d9b-ab0d-6fb0937dee98 · outbound

This paper cites PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training.

CueLearner: Bootstrapping and local policy adaptation from relative feedback PEBBLE: Feedback-Efficient Interactive Reinforcement Learning via Relabeling Experience and Unsupervised Pre-training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:57.368857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:57.368857Z digest=sha256:004c19ca4ff0226be7224e5eb09efb3c2d0ea20438131fe747ba5b8b60d0b1ab

Observation 79ff446e-8da7-4570-8fae-dcbd72a0333b · outbound

This paper cites Interactive learning with corrective feedback for policies based on deep neural networks,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning with corrective feedback for policies based on deep neural networks,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:03.113145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.460776Z digest=sha256:c0a290648acfd7a87b2cc345e5797fa010900497043c7b57c2280349c0e2db0a

Observation 4e9d4548-a904-42cf-901c-2dbf6c0cd36a · outbound

This paper cites Reinforcement learning of motor skills using policy search and human corrective advice,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning of motor skills using policy search and human corrective advice,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.916151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.539086Z digest=sha256:3baab93230c9734855161fa8c56fe3850cc4c1bfb4232c54a29dc02be930e009

Observation 82101eaf-4938-4dc9-8a2a-bec3b2fdefd3 · outbound

This paper cites No, to the right: Online language corrections for robotic manipulation via shared autonomy,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback No, to the right: Online language corrections for robotic manipulation via shared autonomy,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.701103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.686163Z digest=sha256:de5ef0e8c41d9cd6aa1aed9e4f67e8d851313d68fa93a5bd4c06787822140622

Observation 147de395-621b-457c-a72f-cab7571b85ec · outbound

This paper cites Yell At Your Robot: Improving On-the-Fly from Language Corrections.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Yell At Your Robot: Improving On-the-Fly from Language Corrections

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:57.823707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:57.823707Z digest=sha256:2267c26449ab8d179f5629973e782b1319229e81ebfae9dd1eb1d6166aaf4d4c

Observation 4283aa42-c4de-43e0-b529-6e8cb74fc9f8 · outbound

This paper cites An interactive framework for learning continuous actions policies based on corrective feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback An interactive framework for learning continuous actions policies based on corrective feedback,

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T19:47:00.199565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:57.999108Z digest=sha256:db05d8986819e4b43d6c7d7bbb54119da14dcbe303954dbb879689c26f8cbdad

Observation 935d72a3-3053-488f-8139-1c70e4f5a339 · outbound

This paper cites Algorithms for inverse reinforcement learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Algorithms for inverse reinforcement learning,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.503736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.116908Z digest=sha256:aeb495fac10ee5f35eece5250dda284b55f2e7f37a1a3b0cfcea7b053bcdc576

Observation 6d80a485-d838-4d3f-821e-7b0a46936b8b · outbound

This paper cites Interactively shaping agents via human reinforcement: The TAMER framework,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactively shaping agents via human reinforcement: The TAMER framework,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.326424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.292369Z digest=sha256:0f6be9cc0effc33a1d6253a6255a08e1da495772206f768d1da139609082858c

Observation ff3522ad-c95c-4bdf-8712-fad249077b2e · outbound

This paper cites Interactive learning from policy-dependent human feedback,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive learning from policy-dependent human feedback,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:02.079797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.448656Z digest=sha256:43ebe90d6b5160fc3c70d283009ff1dd92ec8d6b4f6a20654707fbf20c8b2cae

Observation 5f5a78cb-696a-4f5e-af77-b26a72e8fb48 · outbound

This paper cites Deep reinforcement learning from human preferences,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning from human preferences,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.823057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.646118Z digest=sha256:859b6892f3730e559ce9acd9f12561ffe043e790113727fdbcd8e03c94a5257c

Observation 444caa52-0700-4b11-bc43-9d1ed8f5799b · outbound

This paper cites Few-shot preference learning for human- in-the-loop rl,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Few-shot preference learning for human- in-the-loop rl,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.663685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.820859Z digest=sha256:3476e6078fb4e6ab7aa95053884a009bf9aacd942eac09a1926ee6d8402dc2aa

Observation ff089631-31ee-44d3-b19f-88e63eda4736 · outbound

This paper cites Integrating behavior cloning and reinforcement learning for improved performance in sparse reward environments,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Integrating behavior cloning and reinforcement learning for improved performance in sparse reward environments,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.411300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:58.938867Z digest=sha256:0bec7c832be842ab1aa45be8c90dccac14af7c8bcbd986ab64ae1553f9ce20fa

Observation c843d469-34c9-46df-b213-c841243fe7fa · outbound

This paper cites Agent- advising approaches in an interactive reinforcement learning scenario,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Agent- advising approaches in an interactive reinforcement learning scenario,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.218622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:59.082151Z digest=sha256:00061a190d74f22c800275d88f8132afffb590776c76fd07bddd842311516218

Observation a0e479ed-1cf6-425e-8774-279f96601423 · outbound

This paper cites Modem: Accelerating visual model-based reinforcement learning with demonstrations,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Modem: Accelerating visual model-based reinforcement learning with demonstrations,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:01.035675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:59.166682Z digest=sha256:4b160daf85e3daaacc688da2e586bd5177f1807fa9316dbb2e6b43b6b4084a99

Observation 00d80b20-ee01-4682-8da9-061fd5a1db2f · outbound

This paper cites Interactive robot learning from verbal correction,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Interactive robot learning from verbal correction,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.268947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.268947Z digest=sha256:083175d2a980077e2be749ff654b7fb41994e1eac9372fd6dfa99b762e014eda

Observation 367dae9a-9353-44d3-8ad8-533c927286e6 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback A reduction of imitation learning and structured prediction to no-regret online learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.801971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:59.375451Z digest=sha256:7527a5edb627a03242e2f877a59ec8424f14138ed82c40585f9203c740b15b40

Observation f182bb5a-b9af-4189-93f1-5341d9f71a8b · outbound

This paper cites Orbit: A unified simulation framework for interactive robot learning environments,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Orbit: A unified simulation framework for interactive robot learning environments,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.507262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.507262Z digest=sha256:fbbc686906964346a64c9f47c301f4f95310e122bb091554ea7eb58efbcc2c0e

Observation 441a8885-0384-41ee-98cf-3ecda6a3a4fc · outbound

This paper cites Anymal - a highly mobile and dynamic quadrupedal robot,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Anymal - a highly mobile and dynamic quadrupedal robot,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.562384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:59.632097Z digest=sha256:5939da7b9f90f0f792c38eeb0fcdb1ee021cc66acf508f8b2bfef17472d500d5

Observation 1de215f1-4210-429d-8fea-5d79092ad54e · outbound

This paper cites Deep reinforcement learning with double q-learning,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Deep reinforcement learning with double q-learning,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.739184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.739184Z digest=sha256:54110b137c4434bcc80754a2c757581c6042a366a1e96234fa2a3b5daa9e9ebe

Observation ba101753-c005-411e-ad8a-d94ef21d471f · outbound

This paper cites Probabilistic roadmaps for path planning in high-dimensional configuration spaces,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Probabilistic roadmaps for path planning in high-dimensional configuration spaces,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:46:59.839291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:46:59.839291Z digest=sha256:50cdf367a0eeead8141a11e92f1b654b7c4b0cea444d3a4acb3af82f3e7d81df

Observation af2ff169-3898-4f90-a720-2dc33086bd1c · outbound

This paper cites Reinforcement learning: An introduction,.

CueLearner: Bootstrapping and local policy adaptation from relative feedback Reinforcement learning: An introduction,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:47:00.386268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T19:46:59.972964Z digest=sha256:f0512ccc8395938c0e70c30908ffba4e1c75befda4c8b55483f4725f3d872376

Pith citing papers

No inbound Pith citation observations are available.