Pith. sign in

Paper Citation Record · LEDGER

Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2407.04694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:09:26.150070Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89b2d313-17d3-4a53-8ad1-5346b4b3a174 · inbound

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct cites this paper.

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:53:23.281474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T19:51:14.630756Z digest=sha256:0b369ebe984b3f7fc85dcc57843343dd0e0ee0341c60fb333a100f515142bf77

Observation 652de3e4-a831-4732-9b28-0e8e1c601c8f · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:22:01.659365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:a780f8923b1a036d9a211147f09f7996b9b46a2c886b17d613cb0188bf1acaeb

Observation a73ccfc0-6f4b-4647-a8a0-703489c7f742 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.181472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.181472Z digest=sha256:285078fd92f407c5b672370743fae7fdbdcf0d5b5e3656fd0ff33373d4c77b3b

Observation 2ee47776-1d8d-47e2-b4be-e7fa232e1549 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:26.150070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:26.150070Z digest=sha256:9bb8ab19f9ae33c437fa0dcd7f416e7c7ce8d82e0f44147d4dd0f79639bf319c

Observation 48250677-93fe-4649-9ae7-99e51a99c4c1 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:56.096143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:56.096143Z digest=sha256:d244259abc76bad4b5a376eab3f0a7e579b6ac9b435c1fc0efcd0efe3cb53af4

Observation b9940761-e6b3-4bbb-9d9b-261c20e169ab · inbound

Honeypot Protocol cites this paper.

Honeypot Protocol Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:34.140641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T14:35:16.230357Z digest=sha256:87fa6bd2b7c47a5a20a2789a07d498d370ee52408479cc18545d33cf2aeb6dc8

Observation cdf4da5b-7e9c-49d6-88af-84404529e2a1 · inbound

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity cites this paper.

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T22:04:18.014075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T10:23:02.697982Z digest=sha256:5ba87a777589264609f0adb7003bb1d12d6e3a7b77a592943eabf55e5856a5b1

Observation 2f90b689-13ce-4f6a-9573-0a5439e1beac · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.866648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:30:27.297971Z digest=sha256:a955857cda80d40003601893d9a10477a5041e77f9b71f2885731706f19fed94

Observation a2620d65-f901-4f6c-abc7-595152701dc7 · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.126158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T18:00:29.237972Z digest=sha256:fd1bc9e687e5f5f3a705a6050030fbed6e752d11920897f5f2136d90e4130a36

Observation b2bb0f61-eb8d-4fef-adc2-ec8a9f3ca6c3 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:54.636218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:0607191720ab15e806b78bf1e873640e20cceba2c9ff7168e999b84af3259497

Observation c4c25cb2-581a-48cf-a0da-eff404e991fe · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.569952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:46f62ce4f6d46a46e37cf1d180f127ddd533d983fcb777fae9bc3baf4a16a19d

Observation 2b5d35e9-af89-4a2f-8cd1-559e209cfb8a · inbound

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models cites this paper.

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.582606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T13:28:25.837654Z digest=sha256:9e37225084162d5879abf730ada8dda4bf3b84c729ef344f81a754a187c534e5

Observation 285b66a0-a140-4bb6-9465-c3ae3571166c · inbound

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models cites this paper.

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.867248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:23:20.122338Z digest=sha256:6b5a142d879375e3470922b6a0ea4ca7af12232360c72d022b4e7767d839b313

Observation d2ed037c-9493-4e88-8dfb-31217cdf5da6 · inbound

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models cites this paper.

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.516874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T09:04:28.487424Z digest=sha256:33f67d078c9f8f89ec4fb179cea76c90ab75ea753f7de2f588a0f0b91a83ffb5

Observation 53786e67-c5f6-4beb-8f51-a7ba59b149e5 · inbound

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation cites this paper.

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T23:29:51.616839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T23:29:51.616839Z digest=sha256:c47f67092cb0b3fe8763d4c53a29857e6997a53be34a6a75783e6450638f2fa2

Observation 53f27d21-25d7-4f24-a9c9-7fc440b71283 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:49.333243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:49.333243Z digest=sha256:ddde607a6cc437ab6aa03bf9a48c098ba79095c72718c358fd7527754209e466

Observation 0179c9e4-7368-4988-b995-fe5db4933da3 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.789026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.789026Z digest=sha256:c69dcc1b9372f0c5962bcbd3e81e99ebd750dd0f43d42faaadcc34b13d4b9053