Pith. sign in

Paper Citation Record · LEDGER

Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2412.12480.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12480 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:26:57.294628Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T08:16:47.896473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3fdd020e-e755-40fa-90bc-d03e4a508367 · inbound

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) cites this paper.

Out of Control -- Why Alignment Needs Formal Control Theory (and an Alignment Control Stack) Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:26:57.294628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:26:57.294628Z digest=sha256:375281a09f83c6a3d62309c590c704c9c8c5d08fa1266525065ce29b9edaa387

Observation 7b890ddb-8501-4c5c-8ebe-256fbb103f47 · inbound

Subversion via Focal Points: Investigating Collusion in LLM Monitoring cites this paper.

Subversion via Focal Points: Investigating Collusion in LLM Monitoring Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:52:29.606862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:52:29.606862Z digest=sha256:94634bb9e586cd5cc6ca49d6fdac6422841ff7ce56131449d96a604df5506bbb

Observation 82d8d8d1-f84e-4a27-b3c1-ece7c72ba992 · inbound

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework cites this paper.

Manipulation Attacks by Misaligned AI: Risk Analysis and Safety Case Framework Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T16:39:36.071018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:39:36.071018Z digest=sha256:7d6df66e53ab344abf2e55cca49bbb73f202ee2919e3dc0d6ce1d1a6eb951d9a

Observation 770be177-6000-4f5a-8c36-0d232562c226 · inbound

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments cites this paper.

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T11:50:20.898113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:40:58.120504Z digest=sha256:6f20243b9107073ecb159fb3cd32ba319965a4770763c2d1ec863a1c44f6a4dd

Observation f4386413-f186-4cdc-b866-736d3c8a7961 · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.422934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:3381d3002d1a0d1202b0514ad49d4229f1101ef0fcd053544d8283d702dae3d8

Observation 39b1371d-14fa-4eb6-ae4b-cf9eef710a14 · inbound

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety cites this paper.

Attack Selection in Agentic AI Control Evaluations Meaningfully Decreases Safety Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T08:16:47.897907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T06:15:32.762877Z digest=sha256:afbe2637181b9410a91108eacfde736ece4bf3627679492ec808dc9e5178a542

Observation 95a36e71-de68-452d-93ad-c64a3c7e6cc7 · inbound

GDM AI Control Roadmap cites this paper.

GDM AI Control Roadmap Subversion Strategy Eval: Can language models statelessly strategize to subvert control protocols?

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T06:53:53.134779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:53:53.134779Z digest=sha256:6ee2834ae06e5bcb746ea880588080560c65fea3cdffe1b596fb2638600ba1d5