Pith. sign in

Paper Citation Record · LEDGER

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models

As of 8 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 0 inbound Pith citation observations for arXiv:2607.13172.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.13172 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:06:17.543845Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

75 of 75 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved73
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5872679e-69fb-4871-8442-faea6ecf55b5 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.407075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.407075Z digest=sha256:6c03f171622757aa0f671a91dcc71be48af16916f1b1558494af22d693b5b118

Observation 645e471f-4904-4ee9-99fb-f067f3f6c5c9 · outbound

This paper cites CRC Press (1999).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models CRC Press (1999)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.482844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.482844Z digest=sha256:69d445437dcefd6e49ae4eb2924f5909a3ed41d8262525cca858f09ba973a367

Observation 635f6ba8-d718-4b37-944c-2bde0560d2d1 · outbound

This paper cites Constrained Policy Optimization via Bayesian World Models.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Constrained Policy Optimization via Bayesian World Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.559482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.559482Z digest=sha256:31d02a0e1f89141d0d2bbcf9beaa71582181f5d7219e043f1085e33e4a346fef

Observation 57b6cea2-4dba-4902-a553-fedd4f52f7b6 · outbound

This paper cites V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.637436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.637436Z digest=sha256:7b83b50ef1284adef804d178f994134ad347f43da859cf52e2c37bf0f962d230

Observation 358affb7-4d68-4030-ad32-1c145fad9b7a · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.779702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.779702Z digest=sha256:318ac702badd270a36a8a06f355abcb371b859cf2e5b1fc1aa9eeaed03c5c1cc

Observation 21ae2d28-a47f-4457-b042-6a54ef43b148 · outbound

This paper cites In: Conference on Robot Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.861850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.861850Z digest=sha256:151daf1121a4b752f25445ad7417096cb4c6ac043719d751f7b2b72350612584

Observation 3482d3ab-8f75-4c3b-8830-71a2cb140a5f · outbound

This paper cites In: Handbook of statistics, vol.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Handbook of statistics, vol

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:10.943797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:10.943797Z digest=sha256:a75abf14acdd4ae96b8945b4b6a71a8d8fab6b669885c228e85c599d7e8c9824

Observation aca5c713-9cad-4fe8-8fce-72390c668a99 · outbound

This paper cites The method of paired comparisons.Biometrika39(3/4), 324–345 (1952).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models The method of paired comparisons.Biometrika39(3/4), 324–345 (1952)

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.032054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.032054Z digest=sha256:311c2ed87bfec640930ce692293cbc6716f6b38c676dbbff787a3c7036dc91d8

Observation 5016fb8d-b9da-4ee7-b6cd-c63212b73bc5 · outbound

This paper cites OpenAI Gym.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models OpenAI Gym

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.152578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.152578Z digest=sha256:edb98ae8b0ffff59a1e6164bffaae071debb9b79dc3478c7735da35610cdb003

Observation 8f90a88f-d5d7-4b16-bcdf-454ff0b300cf · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.262813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.262813Z digest=sha256:f8c097c06dada168a07ff41a48bf65385150641a66a17b4033e0b7c4befd97dd

Observation 02e51200-a361-4bce-bf94-993176c2b2c6 · outbound

This paper cites In: Conference on Robot Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Conference on Robot Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.371963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.371963Z digest=sha256:eb5ed708e6c776f3e17b2d672228ce8f50e983d0f48079eb0de45cc713d2497b

Observation 13e8c098-c7b3-4f83-a797-a91f43c26da0 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.572284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.572284Z digest=sha256:7982b32ec247deaae19e78502bd26958c83381163f405c29c1350053305e575b

Observation a4ad9ae3-6f99-4754-b6a8-deff92f61d24 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.680141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.680141Z digest=sha256:bdfef950d1ffe89ef34ea5d10c18a1d7408489cf4db35b9ee8c012d7eead899c

Observation d7e97992-2418-4da3-b72c-fa52d17f9619 · outbound

This paper cites In: Advances in Neural Information Processing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Advances in Neural Information Processing Systems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.803531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.803531Z digest=sha256:f51a37136d382e703d37e892be89857571181daca3a2429837e3ed57a0fe3770

Observation 3e99214a-5d7b-4b27-9085-d06858a07401 · outbound

This paper cites In: The Twelfth International Conference on Learning Representations (2024).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Twelfth International Conference on Learning Representations (2024)

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.914470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.914470Z digest=sha256:f0e5646054bba02e4012125cf0c3506f14eed5ef159a09b596dc2b52fd10d9bd

Observation 003ee4ba-50a2-4782-81a2-488d3e064660 · outbound

This paper cites In: Pacific Rim Inter- national Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pacific Rim Inter- national Conference on Artificial Intelligence

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:11.985701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:11.985701Z digest=sha256:5e11b4947dae9446a8854435fdbb138ab3f4a145a1080ec7d7fa2993335f61e8

Observation a4c24d43-d88f-4a7c-9c83-2358e42272e6 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.044195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.044195Z digest=sha256:5ba58bed3559f757c603453bedcf2f8244e2cd6713d16f0b18097f2bc5918f50

Observation d04bc555-6b81-4871-af4f-e414589ed257 · outbound

This paper cites Parenting: Safe Reinforcement Learning from Human Input.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Parenting: Safe Reinforcement Learning from Human Input

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.149527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.149527Z digest=sha256:caa11d83205c5a60b5e33989fd6fcadc54b633825b9e9bc362d4b96f95e14092

Observation 18b43076-150c-4c7d-b569-2580dd35b431 · outbound

This paper cites Automatica25(3), 335–348 (1989).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Automatica25(3), 335–348 (1989)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.208187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.208187Z digest=sha256:30f99fe1bce460256e47aced2e0acbdfad813bd92c5980c5907e93b1a7df1dee

Observation 7509fb55-fde4-4126-b2ea-a2f6b0179fa0 · outbound

This paper cites Journal of Machine Learning Research16(1), 1437–1480 (2015).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning Research16(1), 1437–1480 (2015)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.261956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.261956Z digest=sha256:6dd0ee2a0c5a432bf001fb8b7fffcd83bbbf20d6046cc1b672dbcf16d6332235

Observation 7fe89d64-3774-41e7-b02b-f175fe6d44c0 · outbound

This paper cites IEEE Access7, 165007–165017 (2019).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access7, 165007–165017 (2019)

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.301697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.301697Z digest=sha256:d95b47c294ae7d3bc0d22249e115c68b66ee9d6506a02bca646c62ddd2a5953c

Observation bc6f1a8b-7806-4288-8df3-8424ba291284 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.410489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.410489Z digest=sha256:430566f36f95ec39cb7872074a32aa9738f096bb11f1c75ec47d330ea678e884

Observation 53896ce6-c72f-468f-a090-9c6c9ccbec8b · outbound

This paper cites Frontiers in Neurorobotics17, 1280341 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Frontiers in Neurorobotics17, 1280341 (2023)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.461979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.461979Z digest=sha256:8f30d0344812bbe77fb7960043fa20c931cde6a29abd5da62c7052e60bdb02e7

Observation 4fda8422-b934-4416-b670-1ccb4585d4b8 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence (2024).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Pattern Analysis and Machine Intelligence (2024)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.513304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.513304Z digest=sha256:882124d1aaced728d1c79f2797b6ba42f75c1379bd890610f30d166e13526197

Observation 2d319fb4-ab68-45ab-bf88-a348bdfff61d · outbound

This paper cites In: International Conference on Principles and Practice of Multi-Agent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Principles and Practice of Multi-Agent Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.563032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.563032Z digest=sha256:8ff5d6a96dae7a3b0fd854d6ed0456f4e3d666804cb668470c440e731c55ded6

Observation c6b6c9fc-913c-436f-b37e-247da2248c9e · outbound

This paper cites In: Proceedings of the 32nd International Conference on Neural Information Process- ing Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 32nd International Conference on Neural Information Process- ing Systems

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.615304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.615304Z digest=sha256:053a7c5ad39e2bf1ee891e1a0a73b09abbd3009f1f06c3976df1ae201f20561b

Observation 6e415062-a077-4100-88da-d721a118e24d · outbound

This paper cites Dream to Control: Learning Behaviors by Latent Imagination.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Dream to Control: Learning Behaviors by Latent Imagination

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.660894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.660894Z digest=sha256:931bfe0ce4e740c826f9a942552e72315f14a71e795826f8a134632826bc9ab5

Observation 56aad13e-629d-4680-994f-87849a73a0f7 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.724121Z digest=sha256:184826ad75a16313236b4e10f8385a888cb9d7525e937050ff0c17ba071eed1b

Observation 943c0f4b-72e2-4ba6-8660-fa074d8bb6d0 · outbound

This paper cites Mastering Atari with Discrete World Models.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Mastering Atari with Discrete World Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.788064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.788064Z digest=sha256:b8ad109ee310e7e0210562cb6bb0ead3bf47fafdde69452f48d7ea6f3cd8628d

Observation 0ca26ccb-06a1-42ec-a9ea-d7fafb19234a · outbound

This paper cites Nature640(8059), 647–653 (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Nature640(8059), 647–653 (2025)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.900210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.900210Z digest=sha256:a8953991940f419d8f87306892b7ca6c424271117cbc1d4f279d3d2e4334f3fd

Observation fa3e44c1-3dd1-4b81-bb01-b02016f7f152 · outbound

This paper cites Evolutionary Computation9(2), 159–195 (2001).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Evolutionary Computation9(2), 159–195 (2001)

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:12.958241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:12.958241Z digest=sha256:2bbdb750c5f92aad03691a5b40ff89227ab816dbea72cf1c6a8ef463a11b8999

Observation f757f143-a240-4854-bd5f-c0c208396c7e · outbound

This paper cites Neural Computation 9(8), 1735–1780 (1997).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Neural Computation 9(8), 1735–1780 (1997)

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.048117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.048117Z digest=sha256:8f8233f9147151c5ec4dcf7c960fd0a3279ae71757cdbb0670708492ca9d96fa

Observation 950401ec-3c38-4b08-98ad-faaad4904bb9 · outbound

This paper cites 65–70 (1979).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 65–70 (1979)

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.146254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.146254Z digest=sha256:2a06c62180c83a8f226c5f4334e9f8935daf892178032f5a0ad0e88ca1c9f05c

Observation f56b354d-fb38-49c7-8a45-b4d22fd99058 · outbound

This paper cites In: International Conference on Learning Represen- tations.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Learning Represen- tations

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.201380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.201380Z digest=sha256:de9f161d089fdc95634a8e90209b134e414f3fcdc23eaa51aa5b2700352fae8a

Observation f2ea76fe-7146-47a8-80f5-a3c04e4d718c · outbound

This paper cites IEEE Access (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Access (2025)

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.274243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.274243Z digest=sha256:a4110ade91e3e36a67b8fa994f57a49e9e544017d23818e5b2c4a42f24a540f2

Observation 1a0bf936-427a-4af9-ae7a-9d934677ae93 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.294083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.294083Z digest=sha256:0719075690534937e1eb849097d0e52c42a2eec74ebd6da197b7aa839b34da9d

Observation cf643320-0fe1-44d1-ae41-7106519c751a · outbound

This paper cites In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.415296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.415296Z digest=sha256:bfeedb0cd9e7242d0a58eb7abb677fdd6359ffacb2bdb5218ae928640c9e932c

Observation 4c6c5dc8-7912-4fdc-98bb-884db763a9c0 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 38

Resolution
verified exact
doi, observed 2026-08-02T06:09:24.475711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-02T06:06:13.517450Z digest=sha256:d633a657ce15e098375383670e14339ba0b879eeb48c118a14b26b171b0312f6

Observation 7c188e54-b49a-433a-bdf0-6619003d1220 · outbound

This paper cites In: Proceedings of the 18th International Conference on Agents and Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 18th International Conference on Agents and Artificial Intelligence

Reference 39

Resolution
verified exact
doi, observed 2026-08-02T06:09:24.301079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-02T06:06:13.622953Z digest=sha256:25c4e1c54161b2e085badd3dce7997d72e0ffcfc3cb55290ca9f783263cf8723

Observation f49963fc-ef84-4b43-a42b-96086028000d · outbound

This paper cites In: International Confer- ence on Learning Representations (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Confer- ence on Learning Representations (2023)

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.763095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.763095Z digest=sha256:cbb504e0d8bf2bf07d1a4ad547255701311e0404706b1f75039e7bee0badb0a3

Observation d09fb77e-9940-4330-9a07-8ccdd633b873 · outbound

This paper cites Auto-Encoding Variational Bayes.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Auto-Encoding Variational Bayes

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:13.913636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:13.913636Z digest=sha256:c0b428f5a94b78bdcc197de99f3f7674c5959920488d6132ce0e2b00422a6bff

Observation f9441392-03eb-47ea-88a2-4281bd825553 · outbound

This paper cites In: Proceedings of the fifth International Conference on Knowledge Capture.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the fifth International Conference on Knowledge Capture

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.043032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.043032Z digest=sha256:749358334a7a06b7217aa545cff49b147f213f76625f5bfc6b4e1a5f94a482b7

Observation e1330e3f-97ac-4d75-97ab-7f057f5b8c36 · outbound

This paper cites Jour- nal of the American statistical Association47(260), 583–621 (1952).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Jour- nal of the American statistical Association47(260), 583–621 (1952)

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.089341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.089341Z digest=sha256:d1c494a81a78979301d2d5fdd22c9d313737bdcebb69c123b4504aeaeaf83e1b

Observation 8397e02c-a90f-4027-8cbd-54bbb6192d5c · outbound

This paper cites 2, 2022-06-27.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 2, 2022-06-27

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.194382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.194382Z digest=sha256:d4ff9e0edfa69216a4eac40d2a62ba06782f978a29ae233ebafaf6f23d831ef8

Observation e1848782-22fe-4e43-8502-a06ec1a2c853 · outbound

This paper cites In: Inter- national Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Inter- national Conference on Machine Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.325447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.325447Z digest=sha256:fe08d1c87eee367663fbd1502e6c3af37e9f9ad70983d4defbda47c1e3b7a2d5

Observation dcfcf58d-0322-4365-a1e0-dcd6d8635859 · outbound

This paper cites B-Pref: Benchmarking Preference-Based Reinforcement Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models B-Pref: Benchmarking Preference-Based Reinforcement Learning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.354404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.354404Z digest=sha256:39fce74233ebed4e3e24788afd88b07f81b4f55a9dd08c1aba66695a27b41658

Observation ee0177bf-6ff1-4f8b-9caa-d4173551a443 · outbound

This paper cites Scalable agent alignment via reward modeling: a research direction.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Scalable agent alignment via reward modeling: a research direction

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.400874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.400874Z digest=sha256:ff6c9aeed82aeb226344b94e44eacfef79d681b46ec9ad361962fa95f5b12660

Observation 24c366b6-3ec2-4869-ad26-a94856ca2f87 · outbound

This paper cites Cognitive Computation17(5), 1–16 (2025).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Cognitive Computation17(5), 1–16 (2025)

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.457876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.457876Z digest=sha256:2acbb72092dc5f726e33070fc183782bf4af2b28cbe7735c05f2883817647226

Observation 026496bb-fa81-43a1-9971-17339f61fec6 · outbound

This paper cites IEEE Transactions on Vehicular Technology (2026).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models IEEE Transactions on Vehicular Technology (2026)

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.505942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.505942Z digest=sha256:d6b57c0ac229c6239b5fa60354854301e0357c9dff6c088973d2d07d3deb7f8b

Observation 0d9527da-7e3d-4c08-8d4e-ff9b119897a6 · outbound

This paper cites In: 2023 IEEE International Conference on Robotics and Automation (ICRA).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 2023 IEEE International Conference on Robotics and Automation (ICRA)

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.562784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.562784Z digest=sha256:a7f12dcb9a0388f2db44c2a587cc7000543f29bb0802b97339cc1a46ad65686c

Observation 1160be37-5f4b-4398-84d8-04913db487b9 · outbound

This paper cites In: Proceedings of the 23rd International Conference on Autonomous Agents and Mul- tiagent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 23rd International Conference on Autonomous Agents and Mul- tiagent Systems

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.618551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.618551Z digest=sha256:9eb1f6df637c62358a36df4248ef2497ddde732e78e3032930dab1191c646b06

Observation 4c63f1e5-a823-43a7-bfe6-55c99a5222bb · outbound

This paper cites Advances in Neural Information Processing Systems pp.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.667389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.667389Z digest=sha256:19769862bf1025690a32e7644ebd26b6911a7be25eed598939bf1160fce2826d

Observation 776cf99f-9ddf-4ee4-8df3-3f8ee777ce29 · outbound

This paper cites Journal of Machine Learning research9(Nov), 2579–2605 (2008).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Journal of Machine Learning research9(Nov), 2579–2605 (2008)

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.786967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.786967Z digest=sha256:d387370b0c56f85d5be52489c452e39833d7d08397c649ac5bf75978c19bf6a7

Observation 957e170d-daf2-41a7-8f9a-86a380d4fbe1 · outbound

This paper cites In: The Fourteenth International Conference on Learning Representations (2026).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: The Fourteenth International Conference on Learning Representations (2026)

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.845611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.845611Z digest=sha256:b611fa916774d0e37caa481064fcfa64fd8b3b2d5913e7393157f906adeafb25

Observation 0fc26d47-98bb-47c6-b80c-6a3d8d2b6478 · outbound

This paper cites 50–60 (1947).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 50–60 (1947)

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:14.967182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:14.967182Z digest=sha256:1f2f9bf3e3c97d21be2b455ff56009a46ef877d5b788a37b9f2195bbebfaba11

Observation 3442ae48-deb0-4e1e-bd19-aa877764b219 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.086757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.086757Z digest=sha256:c036ecac001d6d345192df0b134f536aaf079f5b2c693cf948b7c68030b79885

Observation af83dc68-0a49-4aec-b7bf-e314cb8d7e4d · outbound

This paper cites In: Pro- ceedings of the Seventeenth International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Pro- ceedings of the Seventeenth International Conference on Machine Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.267967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.267967Z digest=sha256:bad511a0224cfba52095577b51a8de0afc55110969874dcb083af2249de48d80

Observation 3885f59c-497c-4dc9-ad4c-283c9b24e31a · outbound

This paper cites Advances in Neural Information Processing Systems pp.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems pp

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.330406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.330406Z digest=sha256:100c45459c3585d1e1dd68314ed2a14f2e06d230e9624b7cef50d4d4a4dce3b8

Observation 0aaabe6b-ca77-476c-97bc-8734150627b3 · outbound

This paper cites In: 10th International Conference on Learning Representa- tions, ICLR 2022 (2022).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 10th International Conference on Learning Representa- tions, ICLR 2022 (2022)

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.433542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.433542Z digest=sha256:3e9a928ccc197216b33bafb434873192b09e052a8449c96fa6fb137161bf1626

Observation 074b221b-1036-4f30-aa36-8b18886db844 · outbound

This paper cites Advances in neural information processing systems36, 53728–53741 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in neural information processing systems36, 53728–53741 (2023)

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.548909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.548909Z digest=sha256:49b6adf612026466a2c3a6ad56ca3d4c6a1c6ec6c720f2db4312b45e67a3ef49

Observation 90203423-ee24-41a8-b5a0-79d12db101fe · outbound

This paper cites In: Learning for Dynamics and Control.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Learning for Dynamics and Control

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.674885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.674885Z digest=sha256:59850cd015fb5a6eef7de2f3196b8396843f0c05aa6f5b56bc95bb050c0ecfbe

Observation b552c757-6c46-45d1-99d0-e6dd2803d23b · outbound

This paper cites Safe Deep RL in 3D Environments using Human Feedback.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Safe Deep RL in 3D Environments using Human Feedback

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.779669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.779669Z digest=sha256:979835124f8fe9de82f93e665781893a670aa6297dda2076262a315f60efc873

Observation 3250ad68-b736-426a-81d7-5b520946a97e · outbound

This paper cites In: International Conference on Machine Learn- ing.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learn- ing

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:15.872266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:15.872266Z digest=sha256:0a6093acc4b3cdfc9ccd79e648e5106c4b34f71972ee63ffaaf1f0b81d2eeb60

Observation c3a928fa-c4bb-47e8-a354-044fada585b5 · outbound

This paper cites In: Proceedings of the 17th Interna- tional Conference on Autonomous Agents and MultiAgent Systems.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the 17th Interna- tional Conference on Autonomous Agents and MultiAgent Systems

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.034204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.034204Z digest=sha256:6258f12f94cf20cfe55377e4937c8f3852c84b39e1afdcb6cc14b7fde284f2e8

Observation 92f2741f-5159-475e-8371-d7a4aaa3fcfd · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.157222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.157222Z digest=sha256:0b22dbc594cc135c38201d429d02ba56efcc0a5e66cfa49051a5a3238438495c

Observation d76b9425-38b3-4529-aeeb-cb462a2feff1 · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.287126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.287126Z digest=sha256:cca8df0d7428915b285e7006e2b02a72b31ab0eb555afc7c37ee4dddba6bb755

Observation ff335f8f-a034-48d1-9cce-9f925edaf422 · outbound

This paper cites 12151–12162 (2020).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models 12151–12162 (2020)

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.398947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.398947Z digest=sha256:7d633d2127a1c448a0ab2148b5abafa797e938957c183bc005a71e7719861953

Observation 20b95e36-b785-4b08-ad8e-5b64e3732c0b · outbound

This paper cites an unresolved cited work.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Unresolved cited work

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.496029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.496029Z digest=sha256:9623cd1cb9b83795fa8b60616f2cfccd364d9f9f5a7899cf8145307ed40ea6c1

Observation e991026e-41f1-4826-b093-b6be673af8c7 · outbound

This paper cites Advances in Neural Information Processing Systems36, 29252–29272 (2023).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems36, 29252–29272 (2023)

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.638621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.638621Z digest=sha256:a9f2ceca2a86a71b1060271a65ac61609d37e47076935cadc0b4a79075bbbb96

Observation 08edd2ea-965a-4942-b0fb-1261872a5187 · outbound

This paper cites Advances in Neural Information Processing Systems34, 20759–20771 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems34, 20759–20771 (2021)

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.836785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.836785Z digest=sha256:41a1ca371c8b35d7b5fe54f93deb435659d2a775e26a73f247a06061926b2d59

Observation f3072b85-af77-423a-892d-e7035bf9f231 · outbound

This paper cites In: Proceedings of the AAAI conference on artificial intelligence.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Proceedings of the AAAI conference on artificial intelligence

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:16.994533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:16.994533Z digest=sha256:b1b8dcc69a531728bf31d5e9bda753e05b7320ea171671484ea5502da08b939c

Observation e7cbe25a-5236-4692-a1e6-7365ea47ed3f · outbound

This paper cites In: 9th International Conference on Learning Representations, ICLR 2021 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: 9th International Conference on Learning Representations, ICLR 2021 (2021)

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.157234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.157234Z digest=sha256:c0a3681178d2fd67fba0ca67521aadf0685f429c6cfd8bb653b75cfb6035a415

Observation 693bc6dd-be15-4198-bb61-a348d4c92b05 · outbound

This paper cites In: Deep RL Workshop NeurIPS 2021 (2021).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: Deep RL Workshop NeurIPS 2021 (2021)

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.253761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.253761Z digest=sha256:0d05a06903c629ac8d0199b5cb1a188aa4f021a085c913d1b10d261156b04bf0

Observation 173032a1-72fe-4220-aeee-4099c2a0edf1 · outbound

This paper cites Advances in Neural Information Processing Systems35, 2608–2621 (2022).

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models Advances in Neural Information Processing Systems35, 2608–2621 (2022)

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.437226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.437226Z digest=sha256:f48a6081236f99f91c30c353c04de6ded52c72014263a9af298ccc45dfce96e2

Observation ff9edad7-bbf7-4f5f-bd6a-993e64a61a78 · outbound

This paper cites In: International Conference on Machine Learning.

Learning Safe Agent Behaviour from Human Preferences and Justifications via World Models In: International Conference on Machine Learning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-02T06:06:17.543845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:06:17.543845Z digest=sha256:8b1d0afb7ac682bafae2226e77574300b2b38372f8b2ef0e56c38fb02389be9a

Pith citing papers

No inbound Pith citation observations are available.