Pith. sign in

Paper Citation Record · LEDGER

Training People to Reward Robots

As of 20 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2505.10151.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10151 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:20:49.524942Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy15
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fab6465b-5489-4f14-9553-1c9a66e5b8ad · outbound

This paper cites A survey of demonstration learn- ing,.

Training People to Reward Robots A survey of demonstration learn- ing,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.776554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.457423Z digest=sha256:a7a01c7bf1d1bf2ddbe4e21352c02a65a9d59e4c74afc66853a6c7b005c9506b

Observation a0473a5d-3524-4155-9722-5270bb3de240 · outbound

This paper cites An algorithmic perspective on imitation learning,.

Training People to Reward Robots An algorithmic perspective on imitation learning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.764492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.461759Z digest=sha256:1442559be2e761f50caf58146fad6363a7c0993a6214e9e66368565e998d0cbb

Observation 3d93bc03-f1f0-4144-8f77-1bc21c468f73 · outbound

This paper cites How can everyday users efficiently teach robots by demonstrations?.

Training People to Reward Robots How can everyday users efficiently teach robots by demonstrations?

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.753040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.465344Z digest=sha256:f7d9f456d4d87dcc501a911ce54899e9452a6baba8f0217734aab711d00b4fd5

Observation de6544aa-23e7-460d-b126-9989be965878 · outbound

This paper cites Using machine teaching to boost novices’ robot teaching skill,.

Training People to Reward Robots Using machine teaching to boost novices’ robot teaching skill,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.741618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.469342Z digest=sha256:f0455a6ec0106036fba46a507cd15d6d85d9e355a1c2a46a23095a1496a968f6

Observation 5d32237d-abc9-4e21-8eb0-23d49c916353 · outbound

This paper cites When should we prefer offline reinforcement learning over behavioral cloning?.

Training People to Reward Robots When should we prefer offline reinforcement learning over behavioral cloning?

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.729841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.473368Z digest=sha256:5e66e6fc923125318248c25c3048a4128da5570993f7dd5346b21324ac874ce3

Observation 13ad6d3c-59b4-41a0-9d27-48281d9f41b9 · outbound

This paper cites Where machines could re- place humans-and where they can’t (yet),.

Training People to Reward Robots Where machines could re- place humans-and where they can’t (yet),

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.716650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.477355Z digest=sha256:83c15e1e5f00f2b716525535fa5801220f6be696380a85b6eea583a9d575b4c0

Observation 7dfedc2f-28f3-4b29-a41c-d4571079989b · outbound

This paper cites Artificial intelligence and work: A critical review of recent research from the social sciences,.

Training People to Reward Robots Artificial intelligence and work: A critical review of recent research from the social sciences,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.702870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.481433Z digest=sha256:457d8fdb554952f7e47e067895cf923fb428f9879244962dde425c369e39d6e4

Observation ce79350c-b6d9-4edf-8ef6-a8058a64e666 · outbound

This paper cites Ai, robotics, and the future of jobs,.

Training People to Reward Robots Ai, robotics, and the future of jobs,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.690110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.485147Z digest=sha256:3711b22039a6e376024f43d9366e7860fef508e276a5ee33a11614abe4b7abef

Observation fa767e03-5d89-422d-a13f-828790dede54 · outbound

This paper cites An Overview of Machine Teaching.

Training People to Reward Robots An Overview of Machine Teaching

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T21:20:49.489374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:20:49.489374Z digest=sha256:b587f485eb15a34c46a64b0484b64a21df60c6e46ecd03fd93f28dfc9b531769

Observation 8566d782-fa0e-4e36-be7c-e8d3821b875e · outbound

This paper cites Least-squares policy iteration algorithms for robotics: Online, continuous, and automatic,.

Training People to Reward Robots Least-squares policy iteration algorithms for robotics: Online, continuous, and automatic,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.677275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.493380Z digest=sha256:eb039488ad4254155a5f20c752cc8bd6ddf10e769e4dcf3002661c4656fb65b4

Observation bcdf8e0e-d600-4688-ac4f-6bb8b7d26f4e · outbound

This paper cites Locally weighted least squares policy iteration for model-free learning in uncertain environments,.

Training People to Reward Robots Locally weighted least squares policy iteration for model-free learning in uncertain environments,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.663974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.497064Z digest=sha256:70cf087e774aede7b03a7785dac20d3d26bee73e267744f1c6d6d467c9fa5bd6

Observation 1740758e-fb04-462b-aa2a-63bc33429e43 · outbound

This paper cites The teaching dimension of linear learners,.

Training People to Reward Robots The teaching dimension of linear learners,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.651547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.500655Z digest=sha256:39fb505d02bdcdea07536fc3b71e6b1a1c71ea3fe1d1c34244c26c4f598cb6f3

Observation 8586b3d3-00cc-4cae-bff0-7be9ec709418 · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T21:20:49.504106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:20:49.504106Z digest=sha256:7794b00c00e4bf9fd3a02553262aeb344b7e110012eed1b70bafbb39786748a8

Observation 5b1aa9dd-06e6-4854-84e0-49ba7d619d8c · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:20:49.630636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.507906Z digest=sha256:347d111dfeca29655b4a80384bb32f1fe9fd1b2e8b2bd6d9a6bd3ba3b222d4db

Observation 0070a10b-c597-4142-8f8a-c467a7c02bc6 · outbound

This paper cites an unresolved cited work.

Training People to Reward Robots Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:20:49.617789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.511509Z digest=sha256:5782af5f6e031752400a59e37f8d78c153dd7a2a504b41aeadf07e21d2ad3b37

Observation 5b408ec6-884a-45ed-b422-bfa0c5dd2add · outbound

This paper cites A taxonomy of mixed reality visual displays,.

Training People to Reward Robots A taxonomy of mixed reality visual displays,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.605083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.514945Z digest=sha256:a3744d465e7bb663e94077e6da018030b7ec342a83eea0bca6e612491963cce9

Observation 541f4c8c-c008-4a38-9b17-cc1dd4388171 · outbound

This paper cites How do humans teach: On curriculum learning and teaching dimension,.

Training People to Reward Robots How do humans teach: On curriculum learning and teaching dimension,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.593634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.518163Z digest=sha256:2824dacb939596e15c384a98cb5fc1728552d12db664af242f7b6e4830abaaf8

Observation b00b1ef2-b393-4509-bf01-d523b3fc5145 · outbound

This paper cites Curriculum learning,.

Training People to Reward Robots Curriculum learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.581713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.521552Z digest=sha256:343fcfcc37280d5578a0db5801a2f133f0c37473f7b64f5e53807c3e64975fd4

Observation 80a4f14b-711e-4372-b97d-34888a808bc6 · outbound

This paper cites LQR-trees: Feedback motion planning on sparse ran- domized trees,.

Training People to Reward Robots LQR-trees: Feedback motion planning on sparse ran- domized trees,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:20:49.569583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-15T21:20:49.524942Z digest=sha256:2e1e7f43549a4cd1705c17340ec9f2216086de00451a67014c346ac91a780cbb

Pith citing papers

No inbound Pith citation observations are available.