Pith. sign in

Paper Citation Record · LEDGER

Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2407.04694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:30:37.638319Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89b2d313-17d3-4a53-8ad1-5346b4b3a174 · inbound

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct cites this paper.

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:53:23.281474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-23T19:51:14.630756Z digest=sha256:0b369ebe984b3f7fc85dcc57843343dd0e0ee0341c60fb333a100f515142bf77

Observation 652de3e4-a831-4732-9b28-0e8e1c601c8f · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:22:01.659365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:a780f8923b1a036d9a211147f09f7996b9b46a2c886b17d613cb0188bf1acaeb

Observation 23e32c6e-8ca0-4f2e-8a64-9c180c04444a · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.638319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.638319Z digest=sha256:e432878f278b721f12281d9654814bca078b2ebee8e09ac3881867ad3f7114a3

Observation a73ccfc0-6f4b-4647-a8a0-703489c7f742 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.181472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.181472Z digest=sha256:2303b613c8c00bfa51ef121f0b4ffad0d05ecfd7b76bac211badae0053c4a3d3

Observation 2ee47776-1d8d-47e2-b4be-e7fa232e1549 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:26.150070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:26.150070Z digest=sha256:9bb8ab19f9ae33c437fa0dcd7f416e7c7ce8d82e0f44147d4dd0f79639bf319c

Observation 48250677-93fe-4649-9ae7-99e51a99c4c1 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:56.096143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:56.096143Z digest=sha256:d244259abc76bad4b5a376eab3f0a7e579b6ac9b435c1fc0efcd0efe3cb53af4

Observation b9940761-e6b3-4bbb-9d9b-261c20e169ab · inbound

Honeypot Protocol cites this paper.

Honeypot Protocol Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:34.140641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T14:35:16.230357Z digest=sha256:87fa6bd2b7c47a5a20a2789a07d498d370ee52408479cc18545d33cf2aeb6dc8

Observation cdf4da5b-7e9c-49d6-88af-84404529e2a1 · inbound

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity cites this paper.

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T22:04:18.014075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T10:23:02.697982Z digest=sha256:5ba87a777589264609f0adb7003bb1d12d6e3a7b77a592943eabf55e5856a5b1

Observation 2f90b689-13ce-4f6a-9573-0a5439e1beac · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.866648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:30:27.297971Z digest=sha256:a955857cda80d40003601893d9a10477a5041e77f9b71f2885731706f19fed94

Observation a2620d65-f901-4f6c-abc7-595152701dc7 · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.126158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T18:00:29.237972Z digest=sha256:fd1bc9e687e5f5f3a705a6050030fbed6e752d11920897f5f2136d90e4130a36

Observation b2bb0f61-eb8d-4fef-adc2-ec8a9f3ca6c3 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:54.636218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:0607191720ab15e806b78bf1e873640e20cceba2c9ff7168e999b84af3259497

Observation c4c25cb2-581a-48cf-a0da-eff404e991fe · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.569952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:46f62ce4f6d46a46e37cf1d180f127ddd533d983fcb777fae9bc3baf4a16a19d

Observation 2b5d35e9-af89-4a2f-8cd1-559e209cfb8a · inbound

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models cites this paper.

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.582606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T13:28:25.837654Z digest=sha256:9e37225084162d5879abf730ada8dda4bf3b84c729ef344f81a754a187c534e5

Observation 285b66a0-a140-4bb6-9465-c3ae3571166c · inbound

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models cites this paper.

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.867248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:23:20.122338Z digest=sha256:6b5a142d879375e3470922b6a0ea4ca7af12232360c72d022b4e7767d839b313

Observation d2ed037c-9493-4e88-8dfb-31217cdf5da6 · inbound

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models cites this paper.

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.516874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T09:04:28.487424Z digest=sha256:33f67d078c9f8f89ec4fb179cea76c90ab75ea753f7de2f588a0f0b91a83ffb5

Observation 53786e67-c5f6-4beb-8f51-a7ba59b149e5 · inbound

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation cites this paper.

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T23:29:51.616839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T23:29:51.616839Z digest=sha256:c6705f25649d077746f87dc3d66c05721f09751653a61375adc43f5f444b5232

Observation 53f27d21-25d7-4f24-a9c9-7fc440b71283 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:49.333243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:49.333243Z digest=sha256:4ee7663d8601a49490975e5b7582d6511fe62acf2f669aeb4dccbee5a62fd8ec

Observation 0179c9e4-7368-4988-b995-fe5db4933da3 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.789026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.789026Z digest=sha256:c69dcc1b9372f0c5962bcbd3e81e99ebd750dd0f43d42faaadcc34b13d4b9053