Pith. sign in

Paper Citation Record · LEDGER

Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2407.04694.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04694 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:56:13.116641Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 89b2d313-17d3-4a53-8ad1-5346b4b3a174 · inbound

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct cites this paper.

Inspection and Control of Self-Generated-Text Recognition Ability in Llama3-8b-Instruct Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-23T19:53:23.281474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-23T19:51:14.630756Z digest=sha256:b2f55e9eb1e70059add7a75c203c11d439e32bb8a157f459b7e74aa3585a1d58

Observation 652de3e4-a831-4732-9b28-0e8e1c601c8f · inbound

Frontier Models are Capable of In-context Scheming cites this paper.

Frontier Models are Capable of In-context Scheming Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:22:01.659365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-16T14:22:01.616448Z digest=sha256:4f70d1e97acb49f1eb7dd17152bc33e529c0c36ef1acf0a0635088b40d3ae36f

Observation ab6e74f6-07cf-4d03-a955-7b83ab3b0192 · inbound

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches cites this paper.

Large language models for artificial general intelligence (AGI): A survey of foundational principles and approaches Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 274

Resolution
unresolved
no resolver link, observed 2026-08-10T21:56:13.116641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:56:13.116641Z digest=sha256:a322b448e91e841875ad88231dbe54e4f4826db8ba87d7e6ba3c2cf8c002e461

Observation 6a77147f-4dd4-44e1-862b-0e8eb73c48ed · inbound

Open Problems in Machine Unlearning for AI Safety cites this paper.

Open Problems in Machine Unlearning for AI Safety Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-10T21:24:10.314189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:24:10.314189Z digest=sha256:2e477f703b0bed362d405021344b9c395d3d7a37ef434d151efa3c5e82385085

Observation 8274109e-2066-46e9-9ecb-19047f44c99c · inbound

Compromising Honesty and Harmlessness in Language Models via Deception Attacks cites this paper.

Compromising Honesty and Harmlessness in Language Models via Deception Attacks Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T05:42:43.460514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T05:42:43.460514Z digest=sha256:364527b478aa78fe4ab9e2992afaa66ef0f2a8a73af1635e4b8e3c53868f1326

Observation 23e32c6e-8ca0-4f2e-8a64-9c180c04444a · inbound

Does It Make Sense to Speak of Introspection in Large Language Models? cites this paper.

Does It Make Sense to Speak of Introspection in Large Language Models? Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T10:30:37.638319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:30:37.638319Z digest=sha256:8eb49bc1814f4e548bcd7b2317492027d1488d42988726c9e31716d555a6cad5

Observation a73ccfc0-6f4b-4647-a8a0-703489c7f742 · inbound

Model Organisms for Emergent Misalignment cites this paper.

Model Organisms for Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:07:28.181472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:07:28.181472Z digest=sha256:6eddd1be66c3eca74a2d6d9b3eda7248b8b71191a4e2ce971b0f0f254c10e73e

Observation 2ee47776-1d8d-47e2-b4be-e7fa232e1549 · inbound

Convergent Linear Representations of Emergent Misalignment cites this paper.

Convergent Linear Representations of Emergent Misalignment Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:26.150070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:09:26.150070Z digest=sha256:68e68c82fc31fdaf422706f55489de403807474ae441417c48bc94784fa65d8a

Observation 48250677-93fe-4649-9ae7-99e51a99c4c1 · inbound

Safety Features for a Centralised AGI Project cites this paper.

Safety Features for a Centralised AGI Project Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T00:23:56.096143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:23:56.096143Z digest=sha256:021bfb860fb6d1a74120023a3a91191e935dcfb714951023a04123112da72c18

Observation b9940761-e6b3-4bbb-9d9b-261c20e169ab · inbound

Honeypot Protocol cites this paper.

Honeypot Protocol Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:35:34.140641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T14:35:16.230357Z digest=sha256:89b794bf47f8dbadbb38645f90ff20ce2a86832265abc44c5b60e760bf26ee68

Observation cdf4da5b-7e9c-49d6-88af-84404529e2a1 · inbound

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity cites this paper.

Measuring Evaluation-Context Divergence in Open-Weight LLMs: A Paired-Prompt Protocol with Pilot Evidence of Alignment-Pipeline-Specific Heterogeneity Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-08T22:04:18.014075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T10:23:02.697982Z digest=sha256:e301bd3455344a20c03df7e27a314f2e7d576fd187510c168ba3cc7beb78d42e

Observation 2f90b689-13ce-4f6a-9573-0a5439e1beac · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T07:34:02.866648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T07:30:27.297971Z digest=sha256:d3aa2eed87e65d5d6d4b338d9f641fe3830eacc70214e31dc71df124ed730003

Observation a2620d65-f901-4f6c-abc7-595152701dc7 · inbound

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs cites this paper.

Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:04:58.126158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T18:00:29.237972Z digest=sha256:48367172d071c4ea27a7d4bf377ed68a5c2317c36e944e0df1c9de38c90ba030

Observation b2bb0f61-eb8d-4fef-adc2-ec8a9f3ca6c3 · inbound

AI Integrity: Defending Against Backdoors and Secret Loyalties cites this paper.

AI Integrity: Defending Against Backdoors and Secret Loyalties Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:54.636218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-04T14:56:53.806480Z digest=sha256:a09951d90326d1c9937a64b047ed4c4ed834d9dc1029391eb5eebd2e358736bc

Observation c4c25cb2-581a-48cf-a0da-eff404e991fe · inbound

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization cites this paper.

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:37:30.569952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T16:26:34.918099Z digest=sha256:4ac0571dbb456e34d8d6f1acb9c46c6421ca5849219c5d788f8d6e1e8977608b

Observation 2b5d35e9-af89-4a2f-8cd1-559e209cfb8a · inbound

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models cites this paper.

When the Chain of Thought Knows Better: Failure Modes in Multi-Turn Reasoning Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:07:38.582606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T13:28:25.837654Z digest=sha256:0b9603c7b46e4d8e2d15773adc0e1027f609dd100c11b979f8d8b0abd02d51c7

Observation 285b66a0-a140-4bb6-9465-c3ae3571166c · inbound

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models cites this paper.

Evaluation Awareness Is Not One Capability: Evidence from Open Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:49:46.867248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T08:23:20.122338Z digest=sha256:3a8c57ef716fa466d2abe71a933ffc022bcdb3f1ace52bd9253a47f14a907afe

Observation d2ed037c-9493-4e88-8dfb-31217cdf5da6 · inbound

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models cites this paper.

Representational Depth of Evaluation Awareness Shifts With Scale in Open-Weight Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:04:32.516874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T09:04:28.487424Z digest=sha256:7700714008f599f2a1ae389e2e114cf71e5da159f08e15ac882b5707a9db7225

Observation 53786e67-c5f6-4beb-8f51-a7ba59b149e5 · inbound

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation cites this paper.

Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-01T23:29:51.616839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T23:29:51.616839Z digest=sha256:fed3bcfc87779226240ddd848342b1412ca7f0a19bc5d2d25b19cd348b1cb8f4

Observation 53f27d21-25d7-4f24-a9c9-7fc440b71283 · inbound

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models cites this paper.

Routing Subspaces: Auditing Evaluation-to-Deployment Mismatch in Fine-Tuned Language Models Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T14:26:49.333243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:26:49.333243Z digest=sha256:4ee7663d8601a49490975e5b7582d6511fe62acf2f669aeb4dccbee5a62fd8ec

Observation 0179c9e4-7368-4988-b995-fe5db4933da3 · inbound

Asymmetric Communication: Large Language Models and Language Games cites this paper.

Asymmetric Communication: Large Language Models and Language Games Me, Myself, and AI: The Situational Awareness Dataset (SAD) for LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-31T16:43:58.789026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T16:43:58.789026Z digest=sha256:1388f0af4076cfefe1a7ef559d9aabead3d5743ee9d6fceb5662609a3bb85916