Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:06:17.543845Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2607.13172.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T06:06:17.543845Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
75 of 75 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 5872679e-69fb-4871-8442-faea6ecf55b5 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 645e471f-4904-4ee9-99fb-f067f3f6c5c9 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models CRC Press (1999)
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 635f6ba8-d718-4b37-944c-2bde0560d2d1 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Constrained Policy Optimization via Bayesian World Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57b6cea2-4dba-4902-a553-fedd4f52f7b6 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 358affb7-4d68-4030-ad32-1c145fad9b7a · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21ae2d28-a47f-4457-b042-6a54ef43b148 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3482d3ab-8f75-4c3b-8830-71a2cb140a5f · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Handbook of statistics, vol
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aca5c713-9cad-4fe8-8fce-72390c668a99 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models The method of paired comparisons.Biometrika39(3/4), 324–345 (1952)
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5016fb8d-b9da-4ee7-b6cd-c63212b73bc5 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models OpenAI Gym
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f90a88f-d5d7-4b16-bcdf-454ff0b300cf · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02e51200-a361-4bce-bf94-993176c2b2c6 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13e8c098-c7b3-4f83-a797-a91f43c26da0 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ad9ae3-6f99-4754-b6a8-deff92f61d24 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7e97992-2418-4da3-b72c-fa52d17f9619 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e99214a-5d7b-4b27-9085-d06858a07401 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Twelfth International Conference on Learning Representations (2024)
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 003ee4ba-50a2-4782-81a2-488d3e064660 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pacific Rim Inter- national Conference on Artificial Intelligence
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c24d43-d88f-4a7c-9c83-2358e42272e6 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d04bc555-6b81-4871-af4f-e414589ed257 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Parenting: Safe Reinforcement Learning from Human Input
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18b43076-150c-4c7d-b569-2580dd35b431 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Automatica25(3), 335–348 (1989)
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7509fb55-fde4-4126-b2ea-a2f6b0179fa0 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning Research16(1), 1437–1480 (2015)
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fe89d64-3774-41e7-b02b-f175fe6d44c0 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access7, 165007–165017 (2019)
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc6f1a8b-7806-4288-8df3-8424ba291284 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53896ce6-c72f-468f-a090-9c6c9ccbec8b · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Frontiers in Neurorobotics17, 1280341 (2023)
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fda8422-b934-4416-b670-1ccb4585d4b8 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Pattern Analysis and Machine Intelligence (2024)
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d319fb4-ab68-45ab-bf88-a348bdfff61d · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Principles and Practice of Multi-Agent Systems
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6b6c9fc-913c-436f-b37e-247da2248c9e · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 32nd International Conference on Neural Information Process- ing Systems
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e415062-a077-4100-88da-d721a118e24d · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Dream to Control: Learning Behaviors by Latent Imagination
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56aad13e-629d-4680-994f-87849a73a0f7 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 943c0f4b-72e2-4ba6-8660-fa074d8bb6d0 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Mastering Atari with Discrete World Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ca26ccb-06a1-42ec-a9ea-d7fafb19234a · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Nature640(8059), 647–653 (2025)
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3e44c1-3dd1-4b81-bb01-b02016f7f152 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Evolutionary Computation9(2), 159–195 (2001)
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f757f143-a240-4854-bd5f-c0c208396c7e · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Neural Computation 9(8), 1735–1780 (1997)
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 950401ec-3c38-4b08-98ad-faaad4904bb9 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 65–70 (1979)
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f56b354d-fb38-49c7-8a45-b4d22fd99058 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Learning Represen- tations
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2ea76fe-7146-47a8-80f5-a3c04e4d718c · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access (2025)
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a0bf936-427a-4af9-ae7a-9d934677ae93 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf643320-0fe1-44d1-ae41-7106519c751a · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c6c5dc8-7912-4fdc-98bb-884db763a9c0 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c188e54-b49a-433a-bdf0-6619003d1220 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 18th International Conference on Agents and Artificial Intelligence
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f49963fc-ef84-4b43-a42b-96086028000d · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Confer- ence on Learning Representations (2023)
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d09fb77e-9940-4330-9a07-8ccdd633b873 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Auto-Encoding Variational Bayes
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9441392-03eb-47ea-88a2-4281bd825553 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the fifth International Conference on Knowledge Capture
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1330e3f-97ac-4d75-97ab-7f057f5b8c36 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Jour- nal of the American statistical Association47(260), 583–621 (1952)
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8397e02c-a90f-4027-8cbd-54bbb6192d5c · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 2, 2022-06-27
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1848782-22fe-4e43-8502-a06ec1a2c853 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Inter- national Conference on Machine Learning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcfcf58d-0322-4365-a1e0-dcd6d8635859 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models B-Pref: Benchmarking Preference-Based Reinforcement Learning
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee0177bf-6ff1-4f8b-9caa-d4173551a443 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Scalable agent alignment via reward modeling: a research direction
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24c366b6-3ec2-4869-ad26-a94856ca2f87 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Cognitive Computation17(5), 1–16 (2025)
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 026496bb-fa81-43a1-9971-17339f61fec6 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Vehicular Technology (2026)
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d9527da-7e3d-4c08-8d4e-ff9b119897a6 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 2023 IEEE International Conference on Robotics and Automation (ICRA)
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1160be37-5f4b-4398-84d8-04913db487b9 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 23rd International Conference on Autonomous Agents and Mul- tiagent Systems
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c63f1e5-a823-43a7-bfe6-55c99a5222bb · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 776cf99f-9ddf-4ee4-8df3-3f8ee777ce29 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning research9(Nov), 2579–2605 (2008)
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 957e170d-daf2-41a7-8f9a-86a380d4fbe1 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Fourteenth International Conference on Learning Representations (2026)
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fc26d47-98bb-47c6-b80c-6a3d8d2b6478 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 50–60 (1947)
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3442ae48-deb0-4e1e-bd19-aa877764b219 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af83dc68-0a49-4aec-b7bf-e314cb8d7e4d · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pro- ceedings of the Seventeenth International Conference on Machine Learning
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3885f59c-497c-4dc9-ad4c-283c9b24e31a · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0aaabe6b-ca77-476c-97bc-8734150627b3 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 10th International Conference on Learning Representa- tions, ICLR 2022 (2022)
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 074b221b-1036-4f30-aa36-8b18886db844 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in neural information processing systems36, 53728–53741 (2023)
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90203423-ee24-41a8-b5a0-79d12db101fe · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Learning for Dynamics and Control
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b552c757-6c46-45d1-99d0-e6dd2803d23b · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Safe Deep RL in 3D Environments using Human Feedback
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3250ad68-b736-426a-81d7-5b520946a97e · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learn- ing
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3a928fa-c4bb-47e8-a354-044fada585b5 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 17th Interna- tional Conference on Autonomous Agents and MultiAgent Systems
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92f2741f-5159-475e-8371-d7a4aaa3fcfd · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d76b9425-38b3-4529-aeeb-cb462a2feff1 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff335f8f-a034-48d1-9cce-9f925edaf422 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 12151–12162 (2020)
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20b95e36-b785-4b08-ad8e-5b64e3732c0b · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e991026e-41f1-4826-b093-b6be673af8c7 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems36, 29252–29272 (2023)
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08edd2ea-965a-4942-b0fb-1261872a5187 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems34, 20759–20771 (2021)
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3072b85-af77-423a-892d-e7035bf9f231 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI conference on artificial intelligence
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7cbe25a-5236-4692-a1e6-7365ea47ed3f · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 9th International Conference on Learning Representations, ICLR 2021 (2021)
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 693bc6dd-be15-4198-bb61-a348d4c92b05 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Deep RL Workshop NeurIPS 2021 (2021)
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 173032a1-72fe-4220-aeee-4099c2a0edf1 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems35, 2608–2621 (2022)
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff9edad7-bbf7-4f5f-bd6a-993e64a61a78 · outbound
Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.