Pith. sign in

Paper Citation Record · LEDGER

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning

As of 12 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2607.08647.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.08647 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T04:00:47.056185Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact6
  • verified fuzzy1
  • unresolved1
  • parse uncertain0
  • malformed identifier4
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 208e9110-1738-43dc-8668-0ea9739d0691 · outbound

This paper cites URLhttps: //doi.org/10.1007/978-3-642-00982-2_1.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning URLhttps: //doi.org/10.1007/978-3-642-00982-2_1

Reference 1

Resolution
verified exact
doi, observed 2026-07-10T04:06:44.494310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:b837d561a0e8254e3b831282ec14ed31595d28559d24e52563e149982a19d060

Observation 724bce8f-8f9f-4345-bdfe-4a3af69c8bed · outbound

This paper cites Understanding the Power and Limitations of Teaching with Imperfect Knowledge.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Understanding the Power and Limitations of Teaching with Imperfect Knowledge

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:06:44.795817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:4737a96fc152423d50d570fc731965009294a9bdd86902aba2869c26836ed182

Observation 89ae9dcb-9b29-4a60-bfdc-084100c778b2 · outbound

This paper cites Learning Robust Rewards with Adversarial Inverse Reinforcement Learning.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Learning Robust Rewards with Adversarial Inverse Reinforcement Learning

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:06:44.799707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:113f4369d0b393720f09f67be80c551b7d8b5413159d88b3d373929cd7c09b32

Observation 0cfe838c-0d66-413a-bbbf-cf1ecf6df5d8 · outbound

This paper cites The effect of modeling human rationality level on learning rewards from multiple feedback types.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning The effect of modeling human rationality level on learning rewards from multiple feedback types

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T04:06:45.255728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:cbf4af6c8daf4bd57e65c10013a738243db312622ad38ad68ab31741c0c437a0

Observation 4bbd38d5-4249-47c0-bb47-2893f642f869 · outbound

This paper cites Assisted Robust Reward Design.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Assisted Robust Reward Design

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:06:44.786673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:d51bad97a3479782898c9a5ad1523a287de0b13a708634ca4beaba667971da52

Observation eab54314-22fd-4d18-8bc3-a2d7e765b742 · outbound

This paper cites Interactive Teaching Algorithms for Inverse Reinforcement Learning.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Interactive Teaching Algorithms for Inverse Reinforcement Learning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:06:44.792911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:7fc41a403f20598da4a59a512a326236651b2cd898df10f2fe0cf28aa8028168

Observation 165abf34-3469-4981-a613-e4638b6b1fc0 · outbound

This paper cites Mehta and Dylan P.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Mehta and Dylan P

Reference 7

Resolution
metadata mismatch
doi, observed 2026-07-10T04:06:44.496762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:96e3856676dc9f211b22006de12dbe9a02fee490c7caa0d49de96c4c3d8b037b

Observation f31d1d4c-ea9e-4b6a-856b-c57a00822c25 · outbound

This paper cites Effects of Robot Competency and Motion Legibility on Human Correction Feedback.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Effects of Robot Competency and Motion Legibility on Human Correction Feedback

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-10T04:06:44.789465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:1b26eaae8c822e8977c50ad612d470df7c5b880401272cb835ab80b1aebdc021

Observation 05e0804c-f802-43d4-803d-94241f7b81be · outbound

This paper cites An Overview of Machine Teaching.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning An Overview of Machine Teaching

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T04:06:44.803111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:82dc43df8090a06fdf6b1f9faacf0856ce203a80c0e76358cec2be3bae0385e1

Observation 395203d5-155b-44ed-8db4-dd4d596c1fd9 · outbound

This paper cites an unresolved cited work.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Unresolved cited work

Reference 10

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T04:06:45.253692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:96d88a1e112bbcdf12df4778b1e552faa1f81a9a73aedaf97f58b014def37479

Observation db90686e-f1e6-4046-9935-59cf7cb9866c · outbound

This paper cites an unresolved cited work.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-07-10T04:06:45.250108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:065ac646f6b7227020c74680bc1d35525bf30f813408ee3f2788e4c14ecc9f0f

Observation f4b88810-7a0b-417d-ab89-bd2594532659 · outbound

This paper cites We use2×3gridworlds with two cell features (drawn gray and white) and a randomly placed terminal cellTthat may occupy either feature.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning We use2×3gridworlds with two cell features (drawn gray and white) and a randomly placed terminal cellTthat may occupy either feature

Reference 12

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T04:06:45.246599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:8da7c641545114b18b210c58c6a746b14e83fc07f7b53379b96e2fe5963ecb5f

Observation 8a7cab3b-aa36-43cc-a292-133c09631019 · outbound

This paper cites an unresolved cited work.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning Unresolved cited work

Reference 13

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T04:06:45.248495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:dc3ced9b5c597985e2606ad93664da81e93e29de6c0a399948a01f2facdb480b

Observation 5a762f50-e6a9-4847-9796-477fb4079fb4 · outbound

This paper cites S1) 2:Restrict candidate atoms to those in environmentsK 3:D←Greedy Atom Selection(K,U)(Alg.

Multi-Modal, Multi-Environment Machine Teaching for Robust Reward Learning S1) 2:Restrict candidate atoms to those in environmentsK 3:D←Greedy Atom Selection(K,U)(Alg

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-07-10T04:06:45.252086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-10T04:00:47.056185Z digest=sha256:183055ea27554d31e6a27dc26eea75beff190ca59e031a028aa27ec4603d43fa

Pith citing papers

No inbound Pith citation observations are available.