Pith. sign in

Paper Citation Record · LEDGER

Training People to Reward Robots

As of 21 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2505.10151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10151 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:20:49.524942Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fab6465b-5489-4f14-9553-1c9a66e5b8ad · outbound

This paper cites A survey of demonstration learn- ing,.

Training People to Reward Robots A survey of demonstration learn- ing,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.776554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.457423Z digest=sha256:56cea6db9131eaf35be85f92c1325e17da544af98332dc6eb4e06e6031f54ba6

Observation a0473a5d-3524-4155-9722-5270bb3de240 · outbound

This paper cites An algorithmic perspective on imitation learning,.

Training People to Reward Robots An algorithmic perspective on imitation learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.764492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.461759Z digest=sha256:021d742ec6abab4b1c543bd78f396904074f4a6b6335d800e4acc828651409c6

Observation 3d93bc03-f1f0-4144-8f77-1bc21c468f73 · outbound

This paper cites How can everyday users efficiently teach robots by demonstrations?.

Training People to Reward Robots How can everyday users efficiently teach robots by demonstrations?

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.753040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.465344Z digest=sha256:224a4270308eef52419be5d90b3d56fd9b91ac77cea9b2ce7792efb646226111

Observation de6544aa-23e7-460d-b126-9989be965878 · outbound

This paper cites Using machine teaching to boost novices’ robot teaching skill,.

Training People to Reward Robots Using machine teaching to boost novices’ robot teaching skill,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.741618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.469342Z digest=sha256:dbab4964d07196821c2253eae07b1574ea2ecc46320e0982af40112a0d53aadc

Observation 5d32237d-abc9-4e21-8eb0-23d49c916353 · outbound

This paper cites When should we prefer offline reinforcement learning over behavioral cloning?.

Training People to Reward Robots When should we prefer offline reinforcement learning over behavioral cloning?

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.729841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.473368Z digest=sha256:5155d0f3e022fdb170df36b2810ab7488d2e2e4299139feff2ee4544ce3d514d

Observation 13ad6d3c-59b4-41a0-9d27-48281d9f41b9 · outbound

This paper cites Where machines could re- place humans-and where they can’t (yet),.

Training People to Reward Robots Where machines could re- place humans-and where they can’t (yet),

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.716650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.477355Z digest=sha256:aabf07ef94ebe9a206fce74ec76eed837db63a193c08a94de4be9daf16905416

Observation 7dfedc2f-28f3-4b29-a41c-d4571079989b · outbound

This paper cites Artificial intelligence and work: A critical review of recent research from the social sciences,.

Training People to Reward Robots Artificial intelligence and work: A critical review of recent research from the social sciences,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.702870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.481433Z digest=sha256:798f33432a58dd56b89cd3dae6c791f28e639c5cc617f860d59de3a3372a1158

Observation ce79350c-b6d9-4edf-8ef6-a8058a64e666 · outbound

This paper cites Ai, robotics, and the future of jobs,.

Training People to Reward Robots Ai, robotics, and the future of jobs,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.690110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.485147Z digest=sha256:2875b066e2b8628f9ef75a64be53af6cb280d48028627f3e8dfdf306d0314d43

Observation fa767e03-5d89-422d-a13f-828790dede54 · outbound

This paper cites An Overview of Machine Teaching.

Training People to Reward Robots An Overview of Machine Teaching

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:20:49.489374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:20:49.489374Z digest=sha256:b587f485eb15a34c46a64b0484b64a21df60c6e46ecd03fd93f28dfc9b531769

Observation 8566d782-fa0e-4e36-be7c-e8d3821b875e · outbound

This paper cites Least-squares policy iteration algorithms for robotics: Online, continuous, and automatic,.

Training People to Reward Robots Least-squares policy iteration algorithms for robotics: Online, continuous, and automatic,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.677275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.493380Z digest=sha256:f1c276093d58d40d6eb32c20f65fe92e2938fe77b6e9641d254deb06d23d1edf

Observation bcdf8e0e-d600-4688-ac4f-6bb8b7d26f4e · outbound

This paper cites Locally weighted least squares policy iteration for model-free learning in uncertain environments,.

Training People to Reward Robots Locally weighted least squares policy iteration for model-free learning in uncertain environments,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.663974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.497064Z digest=sha256:6a3e3e5754807bb8b5f37d00746402868350dbb8a44ceb2a5aa4802caf5ca05c

Observation 1740758e-fb04-462b-aa2a-63bc33429e43 · outbound

This paper cites The teaching dimension of linear learners,.

Training People to Reward Robots The teaching dimension of linear learners,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.651547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.500655Z digest=sha256:0e8c8ce7321252781a2440c892360611afc524c8e2f99ec0cc520da6292f5abd

Observation 8586b3d3-00cc-4cae-bff0-7be9ec709418 · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:20:49.504106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:20:49.504106Z digest=sha256:7794b00c00e4bf9fd3a02553262aeb344b7e110012eed1b70bafbb39786748a8

Observation 5b1aa9dd-06e6-4854-84e0-49ba7d619d8c · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:20:49.630636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.507906Z digest=sha256:49fd71f3d9b1860822a57a26f8b55083aec3c785c22bd7a8ef7e78a3f0bea35c

Observation 0070a10b-c597-4142-8f8a-c467a7c02bc6 · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:20:49.617789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.511509Z digest=sha256:8f772ed1e4a50d146fd30c8fbc33e66f3770ccdfac0977b707a244dc092932df

Observation 5b408ec6-884a-45ed-b422-bfa0c5dd2add · outbound

This paper cites A taxonomy of mixed reality visual displays,.

Training People to Reward Robots A taxonomy of mixed reality visual displays,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.605083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.514945Z digest=sha256:5eef31680219b158b7b7cfeb320fca14cc5e9788220f516d2ca8cd89eccf1728

Observation 541f4c8c-c008-4a38-9b17-cc1dd4388171 · outbound

This paper cites How do humans teach: On curriculum learning and teaching dimension,.

Training People to Reward Robots How do humans teach: On curriculum learning and teaching dimension,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.593634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.518163Z digest=sha256:5a5675f0a2954d82dc1f8f7d29ae739c4e2bcd041e4c7dd1323743adc3f5e6ab

Observation b00b1ef2-b393-4509-bf01-d523b3fc5145 · outbound

This paper cites Curriculum learning,.

Training People to Reward Robots Curriculum learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.581713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.521552Z digest=sha256:7fe01e5bb85537c84916dc52d23fedfd9abd21c5bba6b24a03ba05c2e3aad8f0

Observation 80a4f14b-711e-4372-b97d-34888a808bc6 · outbound

This paper cites LQR-trees: Feedback motion planning on sparse ran- domized trees,.

Training People to Reward Robots LQR-trees: Feedback motion planning on sparse ran- domized trees,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.569583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-15T21:20:49.524942Z digest=sha256:074880185daf253d16d71dd32be252d98aa601853227122f9d2fc2d27a1eac85

Pith citing papers

No inbound Pith citation observations are available.