Pith. sign in

Paper Citation Record · LEDGER

AI Deception: A Survey of Examples, Risks, and Potential Solutions

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2308.14752.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.14752 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:22:42.672539Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

22
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8fdb77d7-d2dc-4ce1-8537-ae97433be3fa · inbound

The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey cites this paper.

The Landscape of Emerging AI Agent Architectures for Reasoning, Planning, and Tool Calling: A Survey AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:16:41.775976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T23:16:41.679855Z digest=sha256:f1bb65a26a041401f1963e9df0d0ba90b6795392a4a53ac8d3fd8d13f62b1d6f

Observation 7d59b646-c460-4aa0-ab38-03d00e2d72a5 · inbound

Safety case template for frontier AI: A cyber inability argument cites this paper.

Safety case template for frontier AI: A cyber inability argument AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T22:01:11.497659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T22:01:11.497659Z digest=sha256:12c4dcc95c812aed316d62f56078c18cf146c5a645673cbefd121cf63e48e511

Observation 3e154b21-d89c-41c5-86a8-6d1a2f369891 · inbound

Governing AI Agents cites this paper.

Governing AI Agents AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 158

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:38.451939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:38.451939Z digest=sha256:3cd60b59e11b6fa94cafe90e4c91f62a22bdd771ff99769f0dce21cc8109a119

Observation 85c92d42-2d8c-43e8-9bf8-053b56e0a0dc · inbound

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models cites this paper.

Unraveling Token Prediction Refinement and Identifying Essential Layers in Language Models AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-10T14:43:14.797897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:43:14.797897Z digest=sha256:08d01a5c95ddcc434e0947aa3c95f9b929db648662d08f6dda6850e1ab2a5825

Observation 4a0bdd51-1176-47e1-9bf7-e0faea062590 · inbound

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention cites this paper.

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-25T08:30:31.501832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-25T08:30:15.011210Z digest=sha256:d2eb82eb8d8c6e5a00f4cd2b3e942dbb7403fc96b9241fa37e929650137ae77a

Observation d48a54de-5817-475b-84ab-1c65ceeba0f4 · inbound

Evaluating Frontier Models for Stealth and Situational Awareness cites this paper.

Evaluating Frontier Models for Stealth and Situational Awareness AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T04:22:42.672539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T04:22:42.672539Z digest=sha256:d27a42ff77a216af2e0363e0954e929eb07f5a47d6d19009ba061dc51d6ae2d8

Observation 70c9af57-e56a-4e6c-9703-e3cbd36e3cc6 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:51.822116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:51.822116Z digest=sha256:cb7f6ae03ed11589325be6f1b1e303519098906229a18109817cb2aec5461ffa

Observation 55ceeacc-2e29-4b7d-b2c2-ca8e43c16364 · inbound

Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation cites this paper.

Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:19:50.098601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:19:50.098601Z digest=sha256:80e613929adf573b1cb8efeb4a661b50ed4926e2def61bb46b913054212cdfc8

Observation ab6bd055-1c17-4c4c-87f2-16e6d0cd6ba7 · inbound

The Impact of Artificial Intelligence on Human Thought cites this paper.

The Impact of Artificial Intelligence on Human Thought AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T19:58:07.759972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T19:58:07.759972Z digest=sha256:964b70c3bb4ae5a1c61bc046c60c8ecfb50081e8b184a352240a3cec831c47b4

Observation e080a251-6d31-467f-bb62-3332da192d7a · inbound

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia cites this paper.

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:26:24.815133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-18T13:25:17.313704Z digest=sha256:e163b6bfc5885c419ccfa77b208105c6393a741a47744f064bdd03b166dfb5c3

Observation 00847b8c-d90d-42c6-bacf-ac21d8df4758 · inbound

Information Access of the Oppressed: Freirean Design for Emancipatory Information Access cites this paper.

Information Access of the Oppressed: Freirean Design for Emancipatory Information Access AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 148

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:30:27.948925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T07:27:39.494140Z digest=sha256:d7055547e8b018d00af6ef9005a6d46b750fab6541b0ccc023f55a6c59466785

Observation d1618da5-8278-473f-beb6-08d88addf160 · inbound

Detecting Multi-Agent Collusion Through Multi-Agent Interpretability cites this paper.

Detecting Multi-Agent Collusion Through Multi-Agent Interpretability AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:48:22.964745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T22:46:39.337887Z digest=sha256:1bf9d3cce1f7d4c9cc83ddbee63fa7e3b700a759ae013d05b931e1e4013a1e09

Observation 3f59aa94-558a-43c4-87f6-163eadec8634 · inbound

Simulating the Evolution of Alignment and Values in Machine Intelligence cites this paper.

Simulating the Evolution of Alignment and Values in Machine Intelligence AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:05:48.449641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T20:15:46.311347Z digest=sha256:109752244b26283858117d89d9ff6318a0f55ba80e1cd0043ecc6e8e7cf62fc4

Observation 2441a7bf-ce08-4871-96fe-3b5481b42109 · inbound

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling cites this paper.

Linear Probe Accuracy Scales with Model Size and Benefits from Multi-Layer Ensembling AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:29.191771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T13:50:14.400797Z digest=sha256:88803cf89fc8b7b37414f7f238acde824e5f20b507a87ef827136464670b94c7

Observation cdd2fa72-0248-4394-a7df-180d8b9c69ab · inbound

Slot Machines: How LLMs Keep Track of Multiple Entities cites this paper.

Slot Machines: How LLMs Keep Track of Multiple Entities AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:01:04.046214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T23:48:36.019590Z digest=sha256:9877f907d76d1e6f9d6a3b6dc57807e07ace4a1169287cdcd203526a3c6ec016

Observation 2b008938-2cba-41d5-a14b-c9c0530aac54 · inbound

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces cites this paper.

Correcting Influence: Unboxing LLM Outputs with Orthogonal Latent Spaces AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 239

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:55.537487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-14T20:17:01.224864Z digest=sha256:d3573de4d5b8d20e82b07a41f52cf219a0e4714350ff45077d64697f412f6dfd

Observation 55a8c0fa-0d46-4c56-b91a-e99cbb69e10b · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.399083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:756b2dafd4a5bcb80ee2b46dc06d38e0eb395f680c8da08fd860b4bb782c2cdd

Observation e0850522-cc4c-4a6d-b3a9-e4588c5ca04c · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 86

Resolution
verified exact
arxiv_id, observed 2026-07-01T14:05:47.186844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T22:16:47.743387Z digest=sha256:3556ebad1a51334ac7714955d67700b177e8251decab464052d350c9a698bede

Observation c071e1c9-a7f0-4a95-bb91-8ad7a83c8038 · inbound

Temporal Preference Concepts and their Functions in a Large Language Model cites this paper.

Temporal Preference Concepts and their Functions in a Large Language Model AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-12T17:03:44.315006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:03:44.315006Z digest=sha256:fef07dced281a1734fc88e2a27bd52fa6cd02153e18006164d09c6dba8206042

Observation 8a3f1075-186a-449b-b0a2-55435a3b799c · inbound

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing cites this paper.

A Model of Multi-turn Human Persuadability Using Probabilistic Belief Tracing AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 98

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.557548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T06:17:01.173495Z digest=sha256:5894dd1ac0a9006002dac7af87a52042135f86101f613f63a1c46098021889bc

Observation 2a312bf8-81cf-49b0-acbf-840faee45865 · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.739073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:2e50e5013c27787c5a4ba2b78ed6a964320f0559fa29a4fd998e72b766c045c4

Observation f52ccfb1-c325-4ff3-a71f-49b148eab779 · inbound

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology cites this paper.

The Model Organism Lottery: Model Organism Interpretability Strongly Depends on Training Methodology AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:07:08.176437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-02T15:57:48.589980Z digest=sha256:fe77c85decb8fe3f33beeb06df8b259c18dda2f79265bbe0f33e6f408ac74a4f

Observation fdaa1d02-3c7a-49f4-83b0-308edfcea7f0 · inbound

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring cites this paper.

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring AI Deception: A Survey of Examples, Risks, and Potential Solutions

Reference 103

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T00:56:41.067496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-07-10T00:52:47.537142Z digest=sha256:6cd450ca8de2f542d9c7683f11734a69b046a344d070679cfd5ca09969d5dfaf