Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:38.745141Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2507.12872.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:39:38.745141Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-01T05:40:54.002702Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T10:15:44.642321Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3c8a5245-5b7a-4bec-a5f1-ef815bdc27a2 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Towards evaluations-based safety cases for AI scheming
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dd1535e-65ea-4ea5-b03b-12a6419c341d · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Sabotage Evaluations for Frontier Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5daba2b2-56c1-428a-9a14-c27588fa87d8 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework RepliBench: Evaluating the Autonomous Replication Capabilities of Language Model Agents
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c7f23e-f1b2-4122-a1ef-f42783b790ab · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Safety cases for frontier AI
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4fd43903-ab84-4a90-8136-6b47cfd3699f · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Discovering Latent Knowledge in Language Models Without Supervision
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb126926-4866-4039-ba8a-20e9b42e8c9c · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework doi: 10.1126/science
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5945e5-cfdb-47ac-96b1-87d871d8d1b3 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Safety case template for frontier AI: A cyber inability argument
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 092ed59a-3920-4f54-a142-9417237a6709 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Alignment faking in large language models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8251340-9940-4529-9b26-d5f2fa42defb · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Evaluating Large Language Models' Capability to Launch Fully Automated Spear Phishing Campaigns: Validated on Human Subjects
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4514e2c-d7c9-4557-878f-44d85e69c1a3 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework An Overview of Catastrophic AI Risks
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ebb534-918c-4ad6-a927-557843c99c51 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Facade: High-Precision Insider Threat Detection Using Deep Contextual Anomaly Detection
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 35a9ba68-1cf3-48a7-a135-5e6f71fe3865 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework A sketch of an AI control safety case
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca512bf0-6067-4e17-8342-e8b8d5e4dcdc · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Measuring AI Ability to Complete Long Software Tasks
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82d8d8d1-f84e-4a27-b3c1-ece7c72ba992 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f295ef85-315e-4427-ad9d-383551dc10ce · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework arXiv:2410.03768
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab9e34a3-8724-443e-82f5-200cdb825ffd · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework DeepStack: Expert-Level Artificial Intelligence in No-Limit Poker
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a7314a0-47fe-4c6f-85f8-3eaf24427fd8 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d196b84c-0b27-4bf2-a575-2e893d947be3 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Large Language Models Often Know When They Are Being Evaluated
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7579f224-aec7-4c7a-b527-bf7b68bc9725 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework The Alignment Problem from a Deep Learning Perspective
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e3323b6-488b-486c-a837-1b926db7d428 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Linear Probe Penalties Reduce LLM Sycophancy
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e107757b-5e07-42fd-a604-df71d28c0cf0 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Evaluating Frontier Models for Dangerous Capabilities
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f9fe1ac-4046-4eec-97e9-b50e3d3713ca · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework On the Conversational Persuasiveness of Large Language Models: A Randomized Controlled Trial
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbe3c3ce-fde7-47b6-9ef2-b51c5f46b7d1 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Large Language Models can Strategically Deceive their Users when Put Under Pressure
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 254ec2be-34e0-49e3-abd5-1b93ff8b788c · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Melanie Sclar, Jane Yu, Maryam Fazel-Zarandi, Yulia Tsvetkov, Yonatan Bisk, Yejin Choi, and Asli Celikyilmaz
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8217a998-f671-41cf-85db-56b1dbc779ac · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06a423fb-b5be-4a4d-91f9-b1f054111d87 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Cameron Tice, Philipp Alexander Kreer, Nathan Helm-Burger, Prithviraj Singh Shahani, Fedor Ryzhenkov, Jacob Haimes, Felix Hofstätter, and Teun van der Weij
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6142e7d-acb8-4907-bdbb-ae1fba6c8cdd · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Oriol Vinyals, Igor Babuschkin, Wojciech M
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2aa3db-74cd-4152-8fa1-dfcef47d83d0 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework On Targeted Manipulation and Deception when Optimizing LLMs for User Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88aa5c2d-366c-477e-8608-eb2a8a0693c1 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework WebShop: Towards Scalable Real-World Web Interaction with Grounded Language Agents
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d92ccab7-940a-463d-a334-18a3d93c4c69 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework trusted inquiry
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a6467fe5-0cce-4521-b018-c2720d3a5487 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework doi: 10.1017/S0140525X00076512
Reference 1978
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54ff7abf-df0b-4ce8-84da-b5785de4165c · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework doi: 10.1038/s41586-019-1724-z
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ba4898-9e81-40d1-8001-4cef1191468b · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Risks from Learned Optimization in Advanced Machine Learning Systems
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9dbc06b4-a131-4486-9e6b-75de0ea0bbee · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Emergent Abilities of Large Language Models
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5b9cf08-8480-49e9-9190-96519907496a · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Taken out of context: On measuring situational awareness in LLMs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed0ebd9f-1d21-47c4-b787-933009b79d1d · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Anthropic
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 86c9ba4b-f572-46cf-8cae-43602dff42c7 · outbound
Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Ctrl-Z: Controlling AI Agents via Resampling
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5afbc83d-aa33-4eae-bf22-40991007b544 · inbound
Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.