Pith. sign in

Paper Citation Record · LEDGER

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report

As of 7 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 1 inbound Pith citation observation for arXiv:2507.16534.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.16534 v2

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:23.018672Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T09:18:59.008107Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

51 of 51 outbound references displayed

  • verified exact0
  • verified fuzzy8
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 87e7b001-350c-4616-af54-89f5e6b64167 · outbound

This paper cites BreachSeek: A Multi-Agent Automated Penetration Tester.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report BreachSeek: A Multi-Agent Automated Penetration Tester

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.557548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.557548Z digest=sha256:9c6c9ded0ab837633c3f986c5200cb50f17141a7740a1cd261cb95cc165699c2

Observation 1caa89ce-d703-4800-bc51-1c1fc2ae7ed8 · outbound

This paper cites RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.791055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.791055Z digest=sha256:f8b22c33b50746589daf2b2d89e5f87a146587b9c9aaf503cb01977613b58ea6

Observation 96732b36-dfbd-4014-a33f-523f5a0f6fc3 · outbound

This paper cites Accessed: 2025-06-19.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Accessed: 2025-06-19

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:24.027670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:21.939704Z digest=sha256:a1bffb037b442169f631f3ae67bcbecb7f630785c10bbbe12cdc0ca48bed4563

Observation d3c26d2a-fd3f-4ecd-94c6-92a3d98e3565 · outbound

This paper cites DeepSeek-V3 Technical Report.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report DeepSeek-V3 Technical Report

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.020933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.020933Z digest=sha256:537855404b753f7f01510b8caab6910910c0e49843e499fee9c1c4a7f601e6a3

Observation d0e9f8ec-2569-4156-86f9-5e008b821167 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.121008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.121008Z digest=sha256:2a36bdd8be7e38f31865386c71d58322d10e8efcae0a6c3d5b2ee02dbf5cca5d

Observation 278e80e6-66d3-409e-b85d-a4c1c121a7ef · outbound

This paper cites Sciknoweval: Evaluating multi-level scientific knowledge of large language models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Sciknoweval: Evaluating multi-level scientific knowledge of large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.185824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.185824Z digest=sha256:033cd45169befa5e7f404804c17df16d8b3324e076019d098d1bb010a0bd396f

Observation 25ee9542-ef06-462f-a86a-b32797ca7a00 · outbound

This paper cites Virology Capabilities Test (VCT): A Multimodal Virology Q&A Benchmark.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Virology Capabilities Test (VCT): A Multimodal Virology Q&A Benchmark

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.300029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.300029Z digest=sha256:1b332f91474b9f42c45e0a1b3241abea00bcb705878ef11643ab38f83dec1d01

Observation 1f1df01c-3831-4f67-be33-83ff168ce5f8 · outbound

This paper cites The Llama 3 Herd of Models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.418574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.418574Z digest=sha256:3115ea6a330eb418df59d5791c08c94d86790f9b4f36e35cd8c1971d6e04da19

Observation e4758feb-304a-485c-934f-522f1505f8f2 · outbound

This paper cites Alignment faking in large language models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Alignment faking in large language models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.519969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.519969Z digest=sha256:88f1cd3f91b1f1f066db4ebc9073ee94a0546bb30eafdab6df5900fc407ab20c

Observation c9f31935-def7-4d11-ab8f-44771815d2c5 · outbound

This paper cites Multi-Agent Risks from Advanced AI.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Multi-Agent Risks from Advanced AI

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.652580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.652580Z digest=sha256:c9ae310f7748b669c956da92af0f8fc4529298a928eda1e8d88f06c5900f7c43

Observation 52f89928-a294-41b3-b949-9b5b8faae16c · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Measuring Massive Multitask Language Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.743190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.743190Z digest=sha256:b5fb49bd587a7c71114bda578861059f298eefa9fa8df6b7b6db2d9912e29259

Observation a6528721-9175-4293-aa54-96cae918c15e · outbound

This paper cites Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.834304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.834304Z digest=sha256:d8924333d7312539b4a6dd368ab8338ab148b938fe246399591496b2d54fe433

Observation c6c98204-f467-4bbd-be04-871d7c2dbfe1 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.844018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.844018Z digest=sha256:bb537af39afa32b2cc233f5f6509fb91157668e85d144c305775f4b3c28f92d7

Observation 7cee2df6-4214-4911-bf5f-530698648edc · outbound

This paper cites PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report PKU-SafeRLHF: Towards Multi-Level Safety Alignment for LLMs with Human Preference

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.849329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.849329Z digest=sha256:6eaec4771275622f634d7d0feaf89b3db2959edac9a71588b22f3e2a0a63453a

Observation 50541677-0d41-417a-94bd-ff30d9a909d5 · outbound

This paper cites Mitigating Deceptive Alignment via Self-Monitoring.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Mitigating Deceptive Alignment via Self-Monitoring

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.854339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.854339Z digest=sha256:58a0e29f4797c07a1e963988e030551841ed4b71a9520b3c7210e0bfc8064286

Observation c1dfdbb7-93ca-471b-a80f-efbeddb31a2d · outbound

This paper cites SoSBench: Benchmarking Safety Alignment on Six Scientific Domains.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.859306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.859306Z digest=sha256:3d9ae8c74fa395206f94b5523f97ff5c20fe32b8512d9d6610247ca75b771a9c

Observation 63ae91d0-92bf-4670-bb24-cbd6f4184d1b · outbound

This paper cites SecBench: A Comprehensive Multi-Dimensional Benchmarking Dataset for LLMs in Cybersecurity.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report SecBench: A Comprehensive Multi-Dimensional Benchmarking Dataset for LLMs in Cybersecurity

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.863925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.863925Z digest=sha256:fea5a3a409fe1590da3aa4150bb8e1ca3137cab51c3d53a5a1c612637bd12fa5

Observation b3adbd70-7617-4799-bb22-0bfc8b9628b4 · outbound

This paper cites LAB-Bench: Measuring Capabilities of Language Models for Biology Research.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report LAB-Bench: Measuring Capabilities of Language Models for Biology Research

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.869124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.869124Z digest=sha256:e09f26da38f193e40f5926067083d4cd677a2a8ec77e040595c4dacdf9641145

Observation f19f1ff8-ae1f-48e5-9ba7-d3826bd5e83a · outbound

This paper cites SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.874665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.874665Z digest=sha256:b86722bc3748fc0ba151c066472139ec52bf0054206b574a21ebdff19044c85c

Observation 06b97683-5d85-4068-aa1c-34b691cebdb7 · outbound

This paper cites American invitational mathematics examination 2024,.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report American invitational mathematics examination 2024,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.879999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.879999Z digest=sha256:ca6f7211d523a576650f900c5beb297d3b0af89136b0a0671dff2bf8fa1d4b97

Observation 6fe81f37-ab14-4d38-9829-7575cc8dc8ca · outbound

This paper cites Contest problems from the 2024 AIME competition.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Contest problems from the 2024 AIME competition

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:23.978516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.884756Z digest=sha256:60cc045861c102686ce4edb01dc732fc81d5338d5eaa8d86061b96d89ebc347d

Observation cb2efb7f-7418-4292-8608-0d639b571c51 · outbound

This paper cites CAI: An Open, Bug Bounty-Ready Cybersecurity AI.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report CAI: An Open, Bug Bounty-Ready Cybersecurity AI

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.889495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.889495Z digest=sha256:727328ebc1f5f71273fb052f5f710ed270865560cc1e3a8287afaf901fffcf67

Observation fdf74127-cb50-416c-9d5e-bcff8fc39895 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Frontier Models are Capable of In-context Scheming

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.894587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.894587Z digest=sha256:a7a3227e2c1505819757f56a152ebaa3fce75cfa1facd0ff499a2eacc010611e

Observation ba52e359-94d8-4d62-8dd2-626fc61de117 · outbound

This paper cites Responsible scaling policies (rsps).https://metr.org/blog/2023-09-26-rsp/, 09.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Responsible scaling policies (rsps).https://metr.org/blog/2023-09-26-rsp/, 09

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:23.963015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.899057Z digest=sha256:e12ebcd1d89e7b87201f260240a1a548491e4a48f5e6765ff67e9302795e75f3

Observation 93fac917-a50b-44f4-876f-5feedcef89f4 · outbound

This paper cites GAIA: a benchmark for General AI Assistants.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report GAIA: a benchmark for General AI Assistants

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.907474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.907474Z digest=sha256:5a852103986a827a14b4b67d6d9f586f453cc7d3d7163d46d3a5c9b5558758c8

Observation be4d7f1d-a0a9-4166-83dd-5f37f87b8450 · outbound

This paper cites an unresolved cited work.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T15:14:23.946869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.912298Z digest=sha256:dcaf54b001769cad016a596a357cce2cc180f0b81b379b9f1b40259e90c726ba

Observation 6acdd052-728a-470b-a7f7-8b50efb7cddb · outbound

This paper cites Accessed: 2025-07-13.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Accessed: 2025-07-13

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:23.932093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.917782Z digest=sha256:7b6335a566ed553579b28873727b223852623481850efe0c9744ae04dd085ffa

Observation 83e52089-9a2b-4838-b162-475dc803f59e · outbound

This paper cites Large language model-powered AI systems achieve self-replication with no human intervention.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Large language model-powered AI systems achieve self-replication with no human intervention

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.922224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.922224Z digest=sha256:cb63b71583e6bd32b3280db956d3e8f86f00608c89834994728ec236eb10fc29

Observation ded1746d-a757-43f2-ae1c-421e4d68f658 · outbound

This paper cites Evaluating Frontier Models for Dangerous Capabilities.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Evaluating Frontier Models for Dangerous Capabilities

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.926909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.926909Z digest=sha256:2ac913423322a1fa05ae0ff84fb09f401562a541b142aa212d20061ecafb014d

Observation eb9e64a6-a0eb-46a9-9d7e-d14847fea8cc · outbound

This paper cites Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Towards Empathetic Open-domain Conversation Models: a New Benchmark and Dataset

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.932039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.932039Z digest=sha256:88e6d4efd5a04cd58ef5c1c90a423a8636207f742e2b6f218d71d6fcb13dd3ac

Observation 8cb89a31-cc0e-4952-adca-938f89183e31 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.942276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.942276Z digest=sha256:48093898cb1bd83ed60aef460e2a5a62d34ea895a0fa0bf736cd7390df99c86d

Observation 0f26c0a8-a4e0-4f35-b9db-a6df662b0251 · outbound

This paper cites When Autonomy Goes Rogue: Preparing for Risks of Multi-Agent Collusion in Social Systems.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report When Autonomy Goes Rogue: Preparing for Risks of Multi-Agent Collusion in Social Systems

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.946739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.946739Z digest=sha256:fb8a11c35c9075f150148df84e96b6e118f0605eefb5b92fec187f78a8339adc

Observation 1a1a5579-dcf3-487f-b11b-651db1286378 · outbound

This paper cites Towards Understanding Sycophancy in Language Models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Towards Understanding Sycophancy in Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.951866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.951866Z digest=sha256:ca5b09b9056182ac66e2b8f87d4a4f86f733bb58d1f38e0b09e65af87e870cdd

Observation 31cf8291-b415-4b89-b249-67b52b43b67d · outbound

This paper cites Can Language Models Solve Olympiad Programming?.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Can Language Models Solve Olympiad Programming?

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.961513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.961513Z digest=sha256:af1107a099fe82030a1a880f5a532aae675218f4493b71b36ab0bff82a1f27b7

Observation cf9f1a7b-47d0-49cd-b3f6-9e0ee04b9197 · outbound

This paper cites github.io/blog/qwq-32b/.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report github.io/blog/qwq-32b/

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:23.915454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.976376Z digest=sha256:b51acac8a58fe003a7228c78401acce63d9b1702f5e017d78bf26659d8197869

Observation f49cb79f-7a76-4936-848c-20bd29528720 · outbound

This paper cites Teun van der Weij, Felix Hofstätter, Ollie Jaffe, Samuel F Brown, and Francis Rhys Ward.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Teun van der Weij, Felix Hofstätter, Ollie Jaffe, Samuel F Brown, and Francis Rhys Ward

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.981987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.981987Z digest=sha256:6c8f73523f1d1df8612515ae4a3f8f9658dc2e5d85f53e692300551c1a3131d5

Observation 979b8ec1-35af-4b39-9e3c-5a1b481ba118 · outbound

This paper cites Forewarned is Forearmed: A Survey on Large Language Model-based Agents in Autonomous Cyberattacks.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Forewarned is Forearmed: A Survey on Large Language Model-based Agents in Autonomous Cyberattacks

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.991593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.991593Z digest=sha256:a654c4fabe2ccef838d1dd3831a8da055c12cd46dfb33aee745f651765b2a50f

Observation 304d5447-6a79-48f7-af13-9fd530fe6518 · outbound

This paper cites Is Your LLM-Based Multi-Agent a Reliable Real-World Planner? Exploring Fraud Detection in Travel Planning.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Is Your LLM-Based Multi-Agent a Reliable Real-World Planner? Exploring Fraud Detection in Travel Planning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.996359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.996359Z digest=sha256:293490817bbdd44b69f15c9dafb4aded15a5bf9d7014c13cd79c53d5b0f0eb9a

Observation cbe1dbe2-9e4d-41a5-bc10-fff805cfe8bb · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:23.001640Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:23.001640Z digest=sha256:b365d21dfdd7cc7f107669b34ee53ccef47612efa92e476222884ba689be29bc

Observation ca2ca46d-83e1-4bf8-b634-cbf1ce3a42f3 · outbound

This paper cites Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risks of Language Models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Cybench: A Framework for Evaluating Cybersecurity Capabilities and Risks of Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:23.006196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:23.006196Z digest=sha256:fefee020e80ba3067d3db9e595139355063610c35b28d24927fce4523408c61e

Observation 5f8fbb08-a161-444e-9aa8-0823e81e849d · outbound

This paper cites Instruction-Following Evaluation for Large Language Models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Instruction-Following Evaluation for Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:23.014971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:23.014971Z digest=sha256:3eda463709dd55fb281c144f5fe7b2e422b9abbe688412b7d9cfd6afa2fc6b06

Observation 6c2b8b6a-ebd0-4314-91f3-d600ba8c1487 · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:23.018672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:23.018672Z digest=sha256:96e21685d502d68c0cc8d1ce724a3830aa0f4b2fdccdab58519ab43b3d5b3b9b

Observation 59418bc6-79b4-4a4d-b07a-1f5276087639 · outbound

This paper cites GPT-4o System Card.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report GPT-4o System Card

Reference 1953

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.804368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.804368Z digest=sha256:79173ee66e5c65eca5d4df560fd1134e282306c83546ff59127869be2c8d5fc6

Observation 59b84338-8ab4-482f-a89f-9306667dde6e · outbound

This paper cites Scheming AIs: Will AIs fake alignment during training in order to get power?.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.887761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.887761Z digest=sha256:fb47f380e155200090acb16f36631548dc84d0c246fec76ccf2acfe60a376568

Observation 118bef26-7bfa-44a8-bc95-f54a0330bc81 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.986497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.986497Z digest=sha256:0be7243ad44c9d4b53d579f417d321670f58518135b07d7e83db7f5beba6cd6c

Observation edd18540-d5f1-4a2b-8964-d4ab795144cc · outbound

This paper cites Biolp-bench: Measuring understanding of biological lab protocols by large language models.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Biolp-bench: Measuring understanding of biological lab protocols by large language models

Reference 2014

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:24.009701Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:22.839502Z digest=sha256:a42f77ac180b06ac99e7581ccf29510d5aeb5e6a5cd65d261d76948d03c15f1e

Observation b88639e1-0404-4a6a-9840-6186f29b8de5 · outbound

This paper cites International AI Safety Report.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report International AI Safety Report

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.709311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.709311Z digest=sha256:3973c65318643436fdb6ec00e3f98e210eda4b2e296eff914ae18362a7b832e0

Observation c4be9b7e-cc32-4fee-8169-6250a3f5348c · outbound

This paper cites Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.971435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.971435Z digest=sha256:7210e5f7922007fa4b96d0dd346fc68cf988666483db028c8d1b6c0a1ab40208

Observation ce5731ec-f81b-4dee-a162-84c9f9a26377 · outbound

This paper cites Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Fast-DetectGPT: Efficient Zero-Shot Detection of Machine-Generated Text via Conditional Probability Curvature

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.642015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.642015Z digest=sha256:c02e60cd45ff9c47f8bac6d478db77392356907fb64b1d05065c8563af42009f

Observation 52e159df-342b-41ac-b844-e667b42aae8f · outbound

This paper cites Mistral AI.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Mistral AI

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:24.060858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:21.421720Z digest=sha256:f2df568eb3a19914fa551e352ab33ea762a083fd102c25d37337aa785bed415f

Observation c143ecc8-b397-4d13-b15f-036849d2f37e · outbound

This paper cites Accessed: 2025-06-19.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Accessed: 2025-06-19

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T15:14:24.044219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T15:14:21.487379Z digest=sha256:52c3c0566819c919a13e6948c0fee1e190f3e16b37682bc1badf2276503cc4bb

Pith citing papers

Observation f781125e-9a5a-4396-87c3-56b24ecd8e26 · inbound

AI Agents Enable Adaptive Computer Worms cites this paper.

AI Agents Enable Adaptive Computer Worms Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-28T09:21:49.378804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-28T09:18:59.008107Z digest=sha256:2ac8f5340649e99bb2be49a83abbea23c2be8cb20188526d01654216b570e7b1