Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T17:33:56.918863Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:1906.09624.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-25T17:33:56.918863Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-13T14:35:55.709732Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T04:50:57.028932Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f9a6e9bf-9498-4f53-ac25-5f938ad5b4a3 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference write newline
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 26e2a86f-0512-48c5-8098-12ac539c5ded · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Ng, A
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 77b5de5f-6063-47de-a3d7-8c51bc312049 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Learning from human preferences
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3c05d9c6-d592-418d-8463-a6316bf42b7a · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Mindermann, S
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 193bf705-5cf7-4b4b-94b0-bd5ed66ddeb5 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e8f101c9-350d-4682-95a6-d59e15f2a6a0 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 767671ce-bfbb-4b2e-a31e-8246bca7043f · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference planning fallacy
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f6fbc452-286f-4db7-8fdf-c5ca48d683bc · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Kim, K.-E
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1cbc551f-505a-4dc1-8ed9-abd5bfc600d7 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference The easy goal inference problem is still hard
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 735b2c2e-e947-4ac1-a3c1-2e3ce8f481dc · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Rothkopf, C
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a5013674-fbe8-4304-b25d-19923ddaa0ad · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Goodman, N
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1af2b87f-67d2-43f9-a3f8-20d5c74df4cd · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f514c3cb-791d-4a11-8073-adf486d44e67 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Guided cost learning: Deep inverse optimal control via policy optimization
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d1604d1d-1c1d-4c7b-80d0-c23ebf32d61d · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Time discounting and time preference: A critical review
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e782b411-6eb1-48c9-8348-74b46bb0e138 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Multi-task Maximum Entropy Inverse Reinforcement Learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 693ac56f-77bf-4623-802d-636ac2ea1880 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Learning to Search with MCTSnets
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3b4dbd0b-5ccd-4530-932f-e887ac9ad469 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference J., and Dragan, A
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 23bd822e-2431-4e82-9a7c-5b779ed17794 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference A perspective on judgment and choice: mapping bounded rationality
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b887dcf2-7868-4167-bb9c-d3d702316007 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Specification gaming examples in ai
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation be275a38-9860-4fad-8a6e-eeb5ec6ff86c · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Risk-sensitive inverse reinforcement learning via coherent risk models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 56950a82-5761-4c44-a5aa-65a238e37afd · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Y., Russell, S
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b9913298-4bcc-41a7-84c7-14f154df20e8 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Feature visualization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 61ab85df-20b3-4953-a8d2-da32723ea0e3 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Learning model-based planning from scratch
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 667b0124-3235-4e53-8c45-e53c9379e5a2 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation eab746cc-61e7-4515-ac5f-983210fa1049 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Machine Theory of Mind
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 6c41499b-dd99-4e74-924a-d000df67156f · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Where do you think you're going?: Inferring beliefs about dynamics from behavior
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3646b5bd-076a-41c9-b119-87b7eb57b9ce · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Learning agents for uncertain environments
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 92b912de-f170-4396-939e-e930cb1c4b8a · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Inverse reinforcement learning from failure
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation f9cdf10e-d340-41ea-8a20-a6999f8d3d82 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Universal Planning Networks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1af7be0c-c56d-45ce-b586-0bdd95715b6b · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Latent variables and model mis-specification
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 3efd5067-b471-40a5-b3da-e36d7c46f49d · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Evans, O
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0f0ce870-7300-4a21-a28a-4061421fa776 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Value iteration networks
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 1f648b7b-962e-4888-ae17-5b08e458f425 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference and Kahneman, D
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b36a48f2-a461-4f32-9728-0737055dbcde · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Learning a Prior over Intent via Meta-Inverse Reinforcement Learning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a4a2fe69-97b4-4d29-b5f0-32e14b136c93 · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 24f237e6-58b5-48dc-99c6-42b6b278deca · outbound
On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference D., Maas, A
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c999d903-9ad4-4a83-8059-3382c556d2bf · inbound
Mitigating Cognitive Bias in RLHF by Altering Rationality On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation e084806b-b6e8-4988-bfbb-7f96090ddea7 · inbound
Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction On the Feasibility of Learning, Rather than Assuming, Human Biases for Reward Inference
Reference 141
Source-reported events for the cited work
Unavailable: canonical work link unavailable.