Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:38:05.468538Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 7 inbound Pith citation observations for arXiv:2506.04909.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:38:05.468538Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:00:51.707342Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
30 of 30 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation ea1967ed-7fb3-4bfb-a259-2ff5ac84da73 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7f2c0be-31e9-46d6-b0b0-ecc9b43caf03 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Mitchell, T
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bc0d4022-cb2d-464d-b884-7197d6d32dcc · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Discovering latent knowledge in language models without supervision, March 2024
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8cea9d1-5ff4-40ad-ad5a-f045ea23bb2c · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Localizing lying in llama: Understanding instructed dishonesty on true-false questions through prompting, probing, and patching, November 2023
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9bd5f315-7111-4238-8abe-a69eb08545b4 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models A mathematical framework for transformer circuits
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2276cea3-94aa-4d69-a342-392a370dd943 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models R., and Hubinger, E
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7a6dbadb-b282-4992-87e3-ce9132f2a1b7 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Deception abilities emerged in large language models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ff6f362-bac0-4344-bdde-0a39ca7bbecf · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e43df03a-f2fd-4914-b20e-501298c205bc · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Y., Song, S., Hajishirzi, H., Kornblith, S., Farhadi, A., and Schmidt, L
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb049e20-912a-435a-9dc0-7d6644527c2b · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models ( LLMs ): Survey , technical frameworks, and future challenges
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7fae250-c8c1-4559-96df-6ab28f368dd5 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models TruthfulQA: Measuring How Models Mimic Human Falsehoods
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaec6f86-1555-4ac8-9a0c-172a61f328ca · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Cognitive dissonance: Why do language model outputs disagree with internal representations of truthfulness? In Bouamor, H., Pino, J., and Bali, K
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76fa1426-5849-493d-95a1-7de680df65f6 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Faithful chain-of-thought reasoning
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04ccd2a7-7794-45a0-8e7f-e4ce61697f42 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Frontier Models are Capable of In-context Scheming
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e39c8e0a-756d-447f-9f73-173382954953 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Locating and editing factual associations in gpt
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d1ed4703-bad2-484d-8e89-88bf35bfc607 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Interpreting gpt: The logit lens
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31df4916-8878-494f-8a3b-af405b58173a · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models S., Goldstein, S., O'Gara, A., Chen, M., and Hendrycks, D
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9f1ec61-4e4e-4481-ae94-4bd3ba97bcaa · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models can strategically deceive their users when put under pressure, July 2024
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 218aef99-cf33-4e90-9326-c7eadb5a5ed0 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 72b9e771-10e1-463a-913d-1715f16c41cb · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwq-32b: Embracing the power of reinforcement learning, March 2025
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 45882e9b-bb11-42a8-8948-87676ff49ba7 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models L., Sharma, A
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16202906-50d5-480e-8771-259ecba3b2d7 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models M., Thiergart, L., Leech, G., Udell, D., Vazquez, J
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b7709c1-7bc1-44a4-a501-4270a70d1004 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models N., Kaiser, ., and Polosukhin, I
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4a24f4c-66a9-4b16-83dc-98d75501138b · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5497388-7067-482c-a3f7-2173f61acfe6 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models V., Zhou, D., et al
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a8dbc01-e3a2-4e56-9687-20e072740ba3 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwen2.5 Technical Report
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edeeb8da-e7d3-422c-b360-568a11f9766b · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Buzsaki, G
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d6147ff-ac8b-4ef5-a946-2a7379e31e15 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Reasoning models better express their confidence, 2025
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2acd94a5-b29b-4172-b401-4f472583fc24 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Kolter, J
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce335cbb-4e4f-4164-ad63-f60edd888e08 · outbound
When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d94b862f-9d0a-46a8-8e48-25a4526a7381 · inbound
Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c106958-39d9-4aca-85ba-031dcddc8406 · inbound
Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d7330bb-6195-47f7-9d55-0d3ac29e083c · inbound
DECOR: Auditing LLM Deception via Information Manipulation Theory When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4811b80-86b1-48d7-beab-a368e862aeaa · inbound
RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88499384-a89c-4608-a2f1-9f5addf98a37 · inbound
What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 116
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54f538bb-6061-47d0-a0c1-f4f38945e484 · inbound
Transcoders for Investigating Deception in Language Models When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a8ada123-d371-4e1c-99e8-b656e1ef1880 · inbound
Risky Business: Measuring The Faithfulness-Safety Tension When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.