Pith. sign in

Paper Citation Record · LEDGER

Scheming AIs: Will AIs fake alignment during training in order to get power?

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2311.08379.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.08379 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T22:01:11.271010Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 94381514-2069-49f2-87d2-d5a24581ad82 · inbound

Safety case template for frontier AI: A cyber inability argument cites this paper.

Safety case template for frontier AI: A cyber inability argument Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T22:01:11.271010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T22:01:11.271010Z digest=sha256:386e7d2f8d3d5811acb294dbebfcc22672d41efd811661b3caae04fd44e0b38f

Observation 636e1d16-f862-487a-8b2a-317e6825a835 · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T14:22:01.705967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:9b49c75e996f1e6948ade763130045ca1f89b564fd652f727a3dba77a77d02dc

Observation fc547ba4-2dd6-4738-96d0-4e2dc8832389 · inbound

Open Problems in Machine Unlearning for AI Safety cites this paper.

Open Problems in Machine Unlearning for AI Safety Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:24:09.978683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:24:09.978683Z digest=sha256:890697734a36d92227b2bcd211b4df13819719100fed0a1ba63f62fbf53d543d

Observation fedd0b75-5c0d-458e-9d94-51811bd7e2f1 · inbound

Governing AI Agents cites this paper.

Governing AI Agents Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:38.447343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:38.447343Z digest=sha256:361aa83fa5f172b9a295b511f3c94962e9e457e3ef4b91d61715fe5816377284

Observation 7accc103-8857-4485-b457-824a4a74d5a2 · inbound

Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development cites this paper.

Gradual Disempowerment: Systemic Existential Risks from Incremental AI Development Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:34:03.719967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T05:34:03.719967Z digest=sha256:eb826aaf5c6e651759a906f6b9b327109924203828fc1875f266d9ab48ce3c3f

Observation 12bd7929-6ffb-441e-90c7-32071a6903a5 · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.440989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.440989Z digest=sha256:32a5abf8ac9102fb461bc551754cf5f0ebdc6c818b71b09aa26a9f48b7e2e842

Observation d262e28e-47f4-4895-8452-53709453a4fd · inbound

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors cites this paper.

Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.785732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.785732Z digest=sha256:e8fe709d742e4321c7029de1d2dd8caf2ba2decc7941e1ef1eb76a389fc9fd1d

Observation ecc0bc3e-b4bc-4c3b-a1c3-db44e9f26802 · inbound

Why Do Some Language Models Fake Alignment While Others Don't? cites this paper.

Why Do Some Language Models Fake Alignment While Others Don't? Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:27.157853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:27.157853Z digest=sha256:92bdcddc101302838ae41a37f5df5e9ec9579119cfe3e661c8923dedf5d84289

Observation b181277c-52bb-4fc6-b5a1-1e39f6752c58 · inbound

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language cites this paper.

Lessons from a Chimp: AI "Scheming" and the Quest for Ape Language Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T20:18:15.448152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:18:15.448152Z digest=sha256:b3d4a764c60ac2db8591cc757a75167817d1fed050c8d1cb59105864cb978084

Observation 07d99b1f-87ee-4cf2-aa17-3c2e861b04a2 · inbound

Towards Measurement Theory for Artificial Intelligence cites this paper.

Towards Measurement Theory for Artificial Intelligence Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:27:33.107797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:27:33.107797Z digest=sha256:fd828dc87caf9bb2f93b48e078a1a9d1ee87fc5c15caac3ce500b78eac0ec4e0

Observation 4b9e95c8-ae1f-4e1e-94b4-af488562dd9c · inbound

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety cites this paper.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.780078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:03dec1eae47d38ca3876fe3fb77124813fe2359f79fe5b721aa1d9db23fa4e80

Observation 59b84338-8ab4-482f-a89f-9306667dde6e · inbound

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report cites this paper.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2002

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:21.887761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:21.887761Z digest=sha256:f97efb4b093490ff4a135c7b70dd4607424820881545777cb0d30f3192518f2d

Observation 1e25c8df-a869-4f6a-a463-8925d68006c0 · inbound

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare cites this paper.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:55.522992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.522992Z digest=sha256:4ecedbb643bf12346e3aaf6511829f29dcc8e252c75f40c919df9394c01afe72

Observation a43b4a90-2eb6-4333-925f-13e7ba1ddf41 · inbound

Scheming Ability in LLM-to-LLM Strategic Interactions cites this paper.

Scheming Ability in LLM-to-LLM Strategic Interactions Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-18T07:51:03.861101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T07:50:30.597108Z digest=sha256:93de67c6d7ef6bdee0829b2fbe50f78c426cf56538bb6ce1940fde0b69549781

Observation 5dd7bdfb-3129-4812-8929-1e5c6eb8873b · inbound

Emergent Social Intelligence Risks in Generative Multi-Agent Systems cites this paper.

Emergent Social Intelligence Risks in Generative Multi-Agent Systems Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:48:00.845736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T21:45:04.625084Z digest=sha256:00012efbc5fbac6dbec771f2a9e73793ae429e55d9928ff5f77e7d84a5f49dfd

Observation 1c647010-58db-4d5e-93d4-5b626f679129 · inbound

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments cites this paper.

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:20.911520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T11:40:58.120504Z digest=sha256:b630f66d334eaa64505937240fa23a24246945ecb7aff93e1652845f65b0485e

Observation b0ba97ba-c4d4-4d4d-932c-dc70845e3958 · inbound

Deconstructing Superintelligence: Identity, Self-Modification and Diff\'erance cites this paper.

Deconstructing Superintelligence: Identity, Self-Modification and Diff\'erance Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:56:14.089379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T02:34:01.220406Z digest=sha256:fb545acc85751ad97aeea18b70400eb7bb1318dfdefa16492ec0a8be27155d80

Observation 64c14f71-e809-4782-87c6-63ad549b4235 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:59:54.700397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:166b448a496e66c049ebdb6c9e60bc0c590c59a46757bf8bcacc3e49eb85522c

Observation f886069e-f7cc-439f-87c0-5f4c63cce993 · inbound

Consistency Training while Mitigating Obfuscation via Rate Matching cites this paper.

Consistency Training while Mitigating Obfuscation via Rate Matching Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 123

Resolution
verified exact
arxiv_id, observed 2026-06-28T14:32:18.147320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T14:25:43.147442Z digest=sha256:4751dfe99d6d09f110745f47727cd4a90f5eb529e6aac681a7b6d9d2cb4e73fb

Observation 3b60783e-bd76-4676-a4b5-c503757b4ef5 · inbound

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety cites this paper.

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T08:16:47.902337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T06:15:32.762877Z digest=sha256:9a7136c36d388069efbbe70ba9f665b5e8629436100d46534f80c7ec2b562ce1

Observation e99ac314-d84d-4601-836c-9abc1afbd047 · inbound

Building Comparative Motivation Profiles with Instrumental Interventions cites this paper.

Building Comparative Motivation Profiles with Instrumental Interventions Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:27:24.591985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-27T19:45:25.449282Z digest=sha256:ad5edd77d77aa5e2895575f056c179561faf16c07010caf078e573505163c6ad

Observation c0d27a7e-d35a-41ee-8bf1-fc2f3684add6 · inbound

Post-AGI Economies: Superposition and the Second Fundamental Theorem of Welfare Economics cites this paper.

Post-AGI Economies: Superposition and the Second Fundamental Theorem of Welfare Economics Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:27:25.655886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T18:56:37.173132Z digest=sha256:c4d0ca2b33096f0916533f8cbb971b49d16e7fcda199c7130775a3e8b0871169

Observation 9386435d-d62c-45c4-97bd-9dd1b6a4d55d · inbound

Defeat Devices in AI Systems cites this paper.

Defeat Devices in AI Systems Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T08:44:27.905417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-30T08:34:58.879346Z digest=sha256:391dc088bc526aead19f39d88e5a16186893c8a7df6118a550cff12361047189