Pith. sign in

Paper Citation Record · LEDGER

The Limits of Predicting Agents from Behaviour

As of 8 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 1 inbound Pith citation observation for arXiv:2506.02923.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02923 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:25:27.476820Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T10:12:19.552548Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:17:57.278786Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact1
  • verified fuzzy6
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7950fef2-220c-46ef-b56c-baca23b22dfb · outbound

This paper cites assumption-free.

The Limits of Predicting Agents from Behaviour assumption-free

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.677826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.400658Z digest=sha256:a49c675f6271f5b46af67780d6df5071406e06edf426cf5c62fc638d83d5d38a

Observation 62022ce5-b7dd-458f-885e-073ae0566f57 · outbound

This paper cites Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?.

The Limits of Predicting Agents from Behaviour Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.095118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.095118Z digest=sha256:5c49dcc0912e6627149bdb9afbaeb2084a8d9e84fd869334a6eebc99b221920e

Observation 2638d28b-679f-41d6-8d07-0b41a562c222 · outbound

This paper cites Does ChatGPT Have a Mind?.

The Limits of Predicting Agents from Behaviour Does ChatGPT Have a Mind?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.213047Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.213047Z digest=sha256:2190c56d69656b25324ce0636a3355379cffa83c36a9536d63e0b9857cea793c

Observation 2b3eb09e-176d-4597-af94-b601178faaa1 · outbound

This paper cites Language Models Represent Space and Time.

The Limits of Predicting Agents from Behaviour Language Models Represent Space and Time

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.217058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.217058Z digest=sha256:bd4d257d7c352bb4f1d571b39d402a99127a2e42d170c70e95ad93f26f9c1c47

Observation 1ed8737b-4b4e-4a77-88f8-27b58863fe40 · outbound

This paper cites Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task.

The Limits of Predicting Agents from Behaviour Emergent World Representations: Exploring a Sequence Model Trained on a Synthetic Task

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.253885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.253885Z digest=sha256:76b6e9cefad505ebfeeed07db599c63077f48ec6d7489719ec5ccb4445e742c5

Observation 9d97dfa0-8ecb-42d8-ac3d-f806e2f31504 · outbound

This paper cites Robust agents learn causal world models.

The Limits of Predicting Agents from Behaviour Robust agents learn causal world models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.272210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.272210Z digest=sha256:341456ad67098b87068f9edc7a659df851c54e390b9d174d42ac4d044ea043a0

Observation 70e9df7c-176a-40f9-9fba-98c2a26d905f · outbound

This paper cites Bounds on the conditional and average treatment effect with unobserved confounding factors.

The Limits of Predicting Agents from Behaviour Bounds on the conditional and average treatment effect with unobserved confounding factors

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.359955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.359955Z digest=sha256:74f96c5519d49307d717d54ffe6f32b7de9dd7cfa2ad23e068bd73fee257a90c

Observation 3870e3e7-28a4-42d7-a433-b05dc9e897ae · outbound

This paper cites Were I to intervene in the environment, what action do you believe is optimal?.

The Limits of Predicting Agents from Behaviour Were I to intervene in the environment, what action do you believe is optimal?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.527939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.414471Z digest=sha256:dd78ef153e00f2cf5647e475dbace80e3f10fdafc6893d28dfa0d711c55dccf0

Observation a8974cc0-590c-4fa8-9617-de36a99e1a0b · outbound

This paper cites 𝐴 is a difference of two terms written𝐴(𝒓)=𝐴 1(𝒓)−𝐴 2(𝒓).

The Limits of Predicting Agents from Behaviour 𝐴 is a difference of two terms written𝐴(𝒓)=𝐴 1(𝒓)−𝐴 2(𝒓)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.339013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.427661Z digest=sha256:cd3d13f0ce1638d963f0dd957ba0a1090013f6ae223e9f6733f710e21c6fc2d4

Observation b63fbb39-5b9a-465d-af40-336d39e097fd · outbound

This paper cites The nature of the modification is unknown but we are told that after modification, the expected probability of𝑪 is given by𝑃𝜎,𝑑(𝑪), assumed to be known and internalised by the A.

The Limits of Predicting Agents from Behaviour The nature of the modification is unknown but we are told that after modification, the expected probability of𝑪 is given by𝑃𝜎,𝑑(𝑪), assumed to be known and internalised by the A

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:28.205541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.436669Z digest=sha256:06e00d84f90c1842bcb64997e32bb4ad585feca7dda9dd29f5304181628f339e

Observation 645d3b13-869c-4516-ade9-0137cb168b1f · outbound

This paper cites good” or “beneficial.

The Limits of Predicting Agents from Behaviour good” or “beneficial

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:27.983506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.476820Z digest=sha256:359ad02d44411a04305e0e8a6e063f1a1fb89951591a7aac157b3c4057b7f2a8

Observation 2ca736b1-553e-40ca-b613-b5481fe69588 · outbound

This paper cites Towards Resolving Unidentifiability in Inverse Reinforcement Learning.

The Limits of Predicting Agents from Behaviour Towards Resolving Unidentifiability in Inverse Reinforcement Learning

Reference 1967

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.084038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.084038Z digest=sha256:2c4a3713b7557e99223bcf53db6e1be46849e6228ebd591149eca5d8635a8bc0

Observation 41c195e1-68e0-4163-b20d-f9b5f395fb77 · outbound

This paper cites an unresolved cited work.

The Limits of Predicting Agents from Behaviour Unresolved cited work

Reference 1972

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:25:28.846022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.314884Z digest=sha256:87daec825375977c045c1811def14b59f234ce7e71a4323a8d07f803b13302a7

Observation bf64d97c-13b1-43df-979f-acdf4f64c7c6 · outbound

This paper cites Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems.

The Limits of Predicting Agents from Behaviour Towards Guaranteed Safe AI: A Framework for Ensuring Robust and Reliable AI Systems

Reference 1996

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.207549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.207549Z digest=sha256:87ed519ad1d6feb6665bb365f9b3415e9b5c969f308353a89e4d8969a39a2dd4

Observation 6133a42e-5572-4e9d-aeb2-e72494a0ea22 · outbound

This paper cites Risks from Learned Optimization in Advanced Machine Learning Systems.

The Limits of Predicting Agents from Behaviour Risks from Learned Optimization in Advanced Machine Learning Systems

Reference 2008

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.238331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.238331Z digest=sha256:c9978e51842060f811062f423d7eb13ed2cb953b6573dff87a84622f7b750829

Observation 12220a8d-aaf4-4f1a-a636-0d8397a58051 · outbound

This paper cites Preference elicitation and inverse reinforcement learning.

The Limits of Predicting Agents from Behaviour Preference elicitation and inverse reinforcement learning

Reference 2010

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:25:29.038826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.294578Z digest=sha256:323fc7870332630f5cb0d34c003be50c0861e561f10f1622974c0587a2746876

Observation dd62d694-1fe9-4ecd-a76a-240c481816de · outbound

This paper cites Partial Counterfactual Identification from Observational and Experimental Data.

The Limits of Predicting Agents from Behaviour Partial Counterfactual Identification from Observational and Experimental Data

Reference 2021

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:25:27.666163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:25:27.384918Z digest=sha256:a3a6be024a4414cc0a240993e633a8a42e362d95bb0b483d5a95e0da5cb60159

Observation 6ddb955a-133d-4239-b92a-3487ea86450d · outbound

This paper cites Evaluating the World Model Implicit in a Generative Model.

The Limits of Predicting Agents from Behaviour Evaluating the World Model Implicit in a Generative Model

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.336907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.336907Z digest=sha256:624699e07581dd9184e1aadae5f411ba26435417b0d3cf92445157a4ac3dd22f

Observation 821071bc-3983-4688-a749-f6c0789e8f8c · outbound

This paper cites Subjective Causality.

The Limits of Predicting Agents from Behaviour Subjective Causality

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.223716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.223716Z digest=sha256:469e0bd06b494fc6572c6bd831409ac8a5851fa69b6da98fd19810d17c69ed69

Observation b2ea569c-aed3-423c-b25d-2431cbad024b · outbound

This paper cites Can a Bayesian Oracle Prevent Harm from an Agent?.

The Limits of Predicting Agents from Behaviour Can a Bayesian Oracle Prevent Harm from an Agent?

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T11:25:27.089122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:25:27.089122Z digest=sha256:cdde13c6c00bb5aa183935cfacd9acf6a8cc4739bd7349c691961a987f225205

Pith citing papers

Observation dda7c65b-5681-463d-9868-0b2b5ddbf6cd · inbound

The Impossibility of Eliciting Latent Knowledge cites this paper.

The Impossibility of Eliciting Latent Knowledge The Limits of Predicting Agents from Behaviour

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.280043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T10:12:19.552548Z digest=sha256:b1f8504092e78e81922f5a2480f756f57773cdba87e8a1b75db102a1a9c14a99