Pith. sign in

Paper Citation Record · LEDGER

What AI evaluations for preventing catastrophic risks can and cannot do

As of 12 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 3 inbound Pith citation observations for arXiv:2412.08653.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.08653 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T11:56:21.102873Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T09:01:07.210356Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved17
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7468784d-5e67-48ad-aee4-9c0c36b67029 · outbound

This paper cites Declare and Justify: Explicit assumptions in AI evaluations are necessary for effective regulation.

What AI evaluations for preventing catastrophic risks can and cannot do Declare and Justify: Explicit assumptions in AI evaluations are necessary for effective regulation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:20.959799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:20.959799Z digest=sha256:f15d027a93009d5a0be8ada395cfb0f1f9194d24343a4766e6dce223eab1a23f

Observation 4c07cfaa-96fd-4d42-82f1-afecfaf2b215 · outbound

This paper cites Safety Cases: How to Justify the Safety of Advanced AI Systems.

What AI evaluations for preventing catastrophic risks can and cannot do Safety Cases: How to Justify the Safety of Advanced AI Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:20.965769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:20.965769Z digest=sha256:ec5f6c57ea98d02a9f22d7b867256f1269b753f3d629e52c4d96ba6e916c93a3

Observation bc3b5182-caf2-481a-8daa-223e31c20c3e · outbound

This paper cites Safety case template for frontier AI: A cyber inability argument.

What AI evaluations for preventing catastrophic risks can and cannot do Safety case template for frontier AI: A cyber inability argument

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:20.970943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:20.970943Z digest=sha256:ce249712b898c02df30a7265a137f0229356c9d84c25a6d998e6257ea2aa8293

Observation c19967bf-392a-4674-a7af-9dbf3aee01c5 · outbound

This paper cites Anthropic’s Responsible Scaling Policy Version 1.0, 2023.

What AI evaluations for preventing catastrophic risks can and cannot do Anthropic’s Responsible Scaling Policy Version 1.0, 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.577496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:20.976513Z digest=sha256:77f99234734fe68b9c57dedd1748a2a3d287b67716aa19a43d266f9a8559370e

Observation 817b2439-fa05-4b40-b1e5-e808212ae3f6 · outbound

This paper cites Preparedness Framework (Beta), 2023.

What AI evaluations for preventing catastrophic risks can and cannot do Preparedness Framework (Beta), 2023

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.562724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:20.981314Z digest=sha256:6ee6cef4b8e52d11be070f3aebbb647b452c3167ae14fb666b45823527852da9

Observation e81ae6be-27ad-4d7b-9c40-76a12851e49a · outbound

This paper cites Frontier Safety Framework, 2024.

What AI evaluations for preventing catastrophic risks can and cannot do Frontier Safety Framework, 2024

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.546647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:20.986041Z digest=sha256:a0615a0841e4bba96fa372f7d291f33151ce454558eeab554b340bc7ceffab97

Observation d7d3fc15-ed60-4b0e-984f-dca0c53661e7 · outbound

This paper cites Evaluating Frontier Models for Dangerous Capabilities.

What AI evaluations for preventing catastrophic risks can and cannot do Evaluating Frontier Models for Dangerous Capabilities

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:20.991463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:20.991463Z digest=sha256:7328a63a1e314ccec5bd59bc37b5e75d194fa9dd4f9cc5076e2b7fd05a76b1ed

Observation ba8390aa-6224-4285-90d1-2c691ef4fdf2 · outbound

This paper cites LLM Agents can Autonomously Hack Websites.

What AI evaluations for preventing catastrophic risks can and cannot do LLM Agents can Autonomously Hack Websites

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:20.996657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:20.996657Z digest=sha256:2dbf38df102a86f9e3120b6cedad6e6ef00410f34415b84861d096348d20df59

Observation e9a0ca4c-7e19-4ec5-af0f-07ede74845fd · outbound

This paper cites LLM Agents can Autonomously Exploit One-day Vulnerabilities.

What AI evaluations for preventing catastrophic risks can and cannot do LLM Agents can Autonomously Exploit One-day Vulnerabilities

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.001715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.001715Z digest=sha256:87f56e3a45b98784381989178b881f11994cead79374e8d561d3ee822ac6dbf4

Observation 8630ca99-e29e-44ea-960c-c5679d8e56af · outbound

This paper cites Teams of LLM Agents can Exploit Zero-Day Vulnerabilities.

What AI evaluations for preventing catastrophic risks can and cannot do Teams of LLM Agents can Exploit Zero-Day Vulnerabilities

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.006726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.006726Z digest=sha256:75c1333f5edd5f3c006ceec452403f49d3ec6cd56c930eeabdaf41668a03f99a

Observation 2d3b758c-01ec-4f00-8584-2d1dbeda2f49 · outbound

This paper cites On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial.

What AI evaluations for preventing catastrophic risks can and cannot do On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.012376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.012376Z digest=sha256:b12dd3f8de5d47083ba29c89da8be4f0317432c8aa7903f7893f38a8bf514fe4

Observation 3ac8c7fa-0f30-479f-8f82-fc30170e4a39 · outbound

This paper cites an unresolved cited work.

What AI evaluations for preventing catastrophic risks can and cannot do Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-12T11:56:21.530325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.017500Z digest=sha256:80cb7a129c23f166da6e16851ef94e41a1a2c514c5fbb7e66a18418d9a97be89

Observation 12a40c67-8ce5-441d-8cc3-e7c4a83cb7fd · outbound

This paper cites Black-Box Access is Insufficient for Rigorous AI Audits.

What AI evaluations for preventing catastrophic risks can and cannot do Black-Box Access is Insufficient for Rigorous AI Audits

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.024293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.024293Z digest=sha256:4ad4cbac2c8fece6723d495aefef150b82cc9d2e216926d55a490dd06a152da6

Observation 300b0777-7082-439e-ae89-b4811143885e · outbound

This paper cites Number 1.

What AI evaluations for preventing catastrophic risks can and cannot do Number 1

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.514424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.029232Z digest=sha256:5db1f698e96b0500d3d0fdccb3b1bbfabf756352ab1ec9dc5e8616eb85bd223e

Observation 26f9a1da-deb7-4799-9710-10215a022e79 · outbound

This paper cites Mistral CEO confirms ‘leak’ of new open source AI model nearing GPT-4 performance.

What AI evaluations for preventing catastrophic risks can and cannot do Mistral CEO confirms ‘leak’ of new open source AI model nearing GPT-4 performance

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.499309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.035546Z digest=sha256:4f2c56e43269d41a52492530da40225a6a978c27dccc606adefbbfc78bb80d81

Observation b269c1b4-c725-4567-b27c-298ddd69fc68 · outbound

This paper cites Coordinated pausing: An evaluation-based coordination scheme for frontier AI developers.

What AI evaluations for preventing catastrophic risks can and cannot do Coordinated pausing: An evaluation-based coordination scheme for frontier AI developers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.045148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.045148Z digest=sha256:7c5e24251e406ed2b0985495fc6addba16927cc6b99c1a452b2f01b5fad8207e

Observation 32c574b4-ab9c-4204-806c-30cee30ae79d · outbound

This paper cites CyberSecEval 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language Models.

What AI evaluations for preventing catastrophic risks can and cannot do CyberSecEval 2: A Wide-Ranging Cybersecurity Evaluation Suite for Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.049719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.049719Z digest=sha256:1631da32286437f80811193d099174a673783a83b3669df41ed8d9e1ba619adc

Observation 9b452253-cdb7-4371-bea3-85ac17dd387d · outbound

This paper cites Project Naptime: Evaluating Offensive Security Capabili- ties of Large Language Models.

What AI evaluations for preventing catastrophic risks can and cannot do Project Naptime: Evaluating Offensive Security Capabili- ties of Large Language Models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.467303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.054573Z digest=sha256:94c89101376875f0c83390a378e152c6212dce4b579c09898f97101025cc0315

Observation 2cdbc31e-6fd0-40f6-9683-586b08b5b7df · outbound

This paper cites AI capabilities can be significantly improved without expensive retraining.

What AI evaluations for preventing catastrophic risks can and cannot do AI capabilities can be significantly improved without expensive retraining

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.059369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.059369Z digest=sha256:29c36779c1d04a486040af46775c6d039bae1d513b7a09b4fdb71d51351536d8

Observation 0b1e790e-94f8-468b-85ec-2729c0a77599 · outbound

This paper cites SWE-bench: Can language models resolve real-world github issues? In The Twelfth International Conference on Learning Representations, 2024.

What AI evaluations for preventing catastrophic risks can and cannot do SWE-bench: Can language models resolve real-world github issues? In The Twelfth International Conference on Learning Representations, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.064339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.064339Z digest=sha256:3ea70bb7952f5604f5ea997b7d5ec71b991b2213a7a1eb6dd842d2be882c5cb8

Observation 0be6524a-f8b9-4584-9f82-8045a94b9f74 · outbound

This paper cites SWE-bench leaderboard.

What AI evaluations for preventing catastrophic risks can and cannot do SWE-bench leaderboard

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.440729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.069025Z digest=sha256:d1d2bbcd90745fc8a724a71d90033cc32fc2534182266f17ac269394be3817e7

Observation 6b0b18a3-90b8-43a3-bb7f-733cfd2ffecf · outbound

This paper cites GAIA Leaderboard - a Hugging Face Space by gaia-benchmark.

What AI evaluations for preventing catastrophic risks can and cannot do GAIA Leaderboard - a Hugging Face Space by gaia-benchmark

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.424786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.073748Z digest=sha256:1b8ca14abe524c84bf03371e05e806eaed6295ae1862d8c55a40769466e69402

Observation 4975690c-760e-4c98-ae24-983f7d988b7e · outbound

This paper cites HumanEval Benchmark (Code Generation).

What AI evaluations for preventing catastrophic risks can and cannot do HumanEval Benchmark (Code Generation)

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.408538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.078515Z digest=sha256:f3dbe8b8eda05edf63a32e9e03c3453abeef6a3f259b4d55214bfe5fe4f10708

Observation 39d36d8a-4b55-4336-922f-1c3e625b3269 · outbound

This paper cites A survey on in-context learning.

What AI evaluations for preventing catastrophic risks can and cannot do A survey on in-context learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.083324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.083324Z digest=sha256:e49be5abf6c14a4769f9d1570b6f0a42b14297e87206adf9f07001086185a0fa

Observation f7345fa4-c59e-4de9-8973-045ab5437f15 · outbound

This paper cites Anthropic’s Responsible Scaling Policy, 2024.

What AI evaluations for preventing catastrophic risks can and cannot do Anthropic’s Responsible Scaling Policy, 2024

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.382641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.087823Z digest=sha256:98f2e028df4af83e61174f6298732af1148ca8a5d092ad00885b4486ae7e0db6

Observation f59e5af6-936f-44b7-9435-3081475a4b5f · outbound

This paper cites We need a Science of Evals.

What AI evaluations for preventing catastrophic risks can and cannot do We need a Science of Evals

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T11:56:21.365857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.092632Z digest=sha256:577e61a587035bdfc9f0a208d59848b85824d30dedae2deb319c291ac55caca1

Observation 7c19e7bd-3cd7-4c9b-bfa1-4a682fc8b45e · outbound

This paper cites AI Sandbagging: Language Models can Strategically Underperform on Evaluations.

What AI evaluations for preventing catastrophic risks can and cannot do AI Sandbagging: Language Models can Strategically Underperform on Evaluations

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.097843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.097843Z digest=sha256:0813b7919709282954d3ec17383e7361982d7b04d9722b137447ae2d1f2edcd5

Observation 7d36241c-6419-43a1-a959-e4f407471c27 · outbound

This paper cites Stress-Testing Capability Elicitation With Password-Locked Models.

What AI evaluations for preventing catastrophic risks can and cannot do Stress-Testing Capability Elicitation With Password-Locked Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T11:56:21.102873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:56:21.102873Z digest=sha256:f715899944fb1fbac1bd6830c0c8729a43bd7112699aa694912304456e20f40c

Observation 8c2a2164-8318-49f5-bb5a-3d9fa2172135 · outbound

This paper cites an unresolved cited work.

What AI evaluations for preventing catastrophic risks can and cannot do Unresolved cited work

Reference 2024

Resolution
parse uncertain
raw_fallback, observed 2026-08-12T11:56:21.483430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-12T11:56:21.040583Z digest=sha256:967974d164f3df6caba5ff0d60010c90cbb439b6288bc2cea64795c8d18ac1bf

Pith citing papers

Observation 8f23a12e-520b-4ff7-b021-a0d8eb82b270 · inbound

From Disclosure to Self-Referential Opacity: Six Dimensions of Strain in Current AI Governance cites this paper.

From Disclosure to Self-Referential Opacity: Six Dimensions of Strain in Current AI Governance What AI evaluations for preventing catastrophic risks can and cannot do

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:00:22.254888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T11:55:50.822764Z digest=sha256:3d2426dc0f85c5613eb55ad9a699d1b4745b6180873c5bf6cf99b105e7f30dc1

Observation c9e9d58b-a850-459d-9472-b294307149e7 · inbound

Scaffold Effects on GAIA: A Controlled Comparison cites this paper.

Scaffold Effects on GAIA: A Controlled Comparison What AI evaluations for preventing catastrophic risks can and cannot do

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:07:27.088498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-27T18:26:58.469351Z digest=sha256:a15738acf34e658ab1ce6fed98dcc00df929be01754a2f4ecd66298c4b120fff

Observation 7629eaf4-63ee-48b1-8620-7afde37db9c6 · inbound

Verifying Restrictions on Frontier AI Research cites this paper.

Verifying Restrictions on Frontier AI Research What AI evaluations for preventing catastrophic risks can and cannot do

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.211462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-30T09:01:07.210356Z digest=sha256:a035ced84bc719ef4866bf04d85af406e7a226526955d903e9f149f1c7e412d6