Pith. sign in

Paper Citation Record · LEDGER

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

As of 7 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 7 inbound Pith citation observations for arXiv:2506.04909.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.04909 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:38:05.468538Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:00:51.707342Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation ea1967ed-7fb3-4bfb-a259-2ff5ac84da73 · outbound

This paper cites write newline.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.027941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.027941Z digest=sha256:f56844032fc42e1f1ef1869e948e5cd6952d5a1c1cdfe7a87dd99eda5fc5ded5

Observation e7f2c0be-31e9-46d6-b0b0-ecc9b43caf03 · outbound

This paper cites and Mitchell, T.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Mitchell, T

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.795250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.079573Z digest=sha256:ae4c9601edeae0d5a0f7681c74fef3a57e4df3ce6563835b5cce16cfc290755b

Observation bc0d4022-cb2d-464d-b884-7197d6d32dcc · outbound

This paper cites Discovering latent knowledge in language models without supervision, March 2024.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Discovering latent knowledge in language models without supervision, March 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.698680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.247102Z digest=sha256:33f3112439164ae976016d2ccc4ccfa1c0544c3ccedbc273ca62563aa663b067

Observation f8cea9d1-5ff4-40ad-ad5a-f045ea23bb2c · outbound

This paper cites Localizing lying in llama: Understanding instructed dishonesty on true-false questions through prompting, probing, and patching, November 2023.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Localizing lying in llama: Understanding instructed dishonesty on true-false questions through prompting, probing, and patching, November 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.506436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.364755Z digest=sha256:071de80bcb37afe81aaa65c852a783d0aee6c5190cf836278a24fdab4a02ccd6

Observation 9bd5f315-7111-4238-8abe-a69eb08545b4 · outbound

This paper cites A mathematical framework for transformer circuits.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models A mathematical framework for transformer circuits

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.444405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.444405Z digest=sha256:d0b7b9ddaeb7a7d9a51dc38aa1217bd2601f52f450f6c35cb19ded22fea2b79f

Observation 2276cea3-94aa-4d69-a342-392a370dd943 · outbound

This paper cites R., and Hubinger, E.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models R., and Hubinger, E

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:08.264957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.535294Z digest=sha256:9b0fdec32bb9295d3e522cc0a56e1b76a5f5aec780322bdbec5a962924630559

Observation 7a6dbadb-b282-4992-87e3-ce9132f2a1b7 · outbound

This paper cites Deception abilities emerged in large language models.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Deception abilities emerged in large language models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:02.688962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:02.688962Z digest=sha256:bb0ef5e244268ada03f01a9a368f8ca4c5166038d74d4a457edda8a764ca9685

Observation 2ff6f362-bac0-4344-bdde-0a39ca7bbecf · outbound

This paper cites an unresolved cited work.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:38:08.038259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.814492Z digest=sha256:2d4c083d3dbd043b917cd1356c7b7a5edfae21cd85328923f3d12a95f76019dc

Observation e43df03a-f2fd-4914-b20e-501298c205bc · outbound

This paper cites Y., Song, S., Hajishirzi, H., Kornblith, S., Farhadi, A., and Schmidt, L.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Y., Song, S., Hajishirzi, H., Kornblith, S., Farhadi, A., and Schmidt, L

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.930957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:02.936643Z digest=sha256:048f1ed9b7586b7fbceb3de1cc2135a618546e09e914db7dc2efa3ad5ef39521

Observation eb049e20-912a-435a-9dc0-7d6644527c2b · outbound

This paper cites Large language models ( LLMs ): Survey , technical frameworks, and future challenges.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models ( LLMs ): Survey , technical frameworks, and future challenges

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.057101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.057101Z digest=sha256:f1e61f1ae78f24976a614f8eb364994033b76d27a39ed2191b02e9cc5c3d10da

Observation c7fae250-c8c1-4559-96df-6ab28f368dd5 · outbound

This paper cites TruthfulQA: Measuring How Models Mimic Human Falsehoods.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models TruthfulQA: Measuring How Models Mimic Human Falsehoods

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.263205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.263205Z digest=sha256:82435e4c97fd170511480e8579825c8e5bbf4d859e995c137209240544ce9bd7

Observation eaec6f86-1555-4ac8-9a0c-172a61f328ca · outbound

This paper cites Cognitive dissonance: Why do language model outputs disagree with internal representations of truthfulness? In Bouamor, H., Pino, J., and Bali, K.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Cognitive dissonance: Why do language model outputs disagree with internal representations of truthfulness? In Bouamor, H., Pino, J., and Bali, K

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.360826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.360826Z digest=sha256:2e60915eb159bfa7e3de2c01e6643df7a5b83f4f99f419ab7406e822310fd85d

Observation 76fa1426-5849-493d-95a1-7de680df65f6 · outbound

This paper cites Faithful chain-of-thought reasoning.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Faithful chain-of-thought reasoning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.798430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.478408Z digest=sha256:8e030e4efe5e665e42fdadd607853d1c4291dcfe84800fa0486fd5a8d0b4af6c

Observation 04ccd2a7-7794-45a0-8e7f-e4ce61697f42 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Frontier Models are Capable of In-context Scheming

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:03.666723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:03.666723Z digest=sha256:0867247c41f4239836b446b3fee982501d483fe82dfcabdc36a1601aeadf82b4

Observation e39c8e0a-756d-447f-9f73-173382954953 · outbound

This paper cites Locating and editing factual associations in gpt.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Locating and editing factual associations in gpt

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.670229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.799399Z digest=sha256:235051c64f275f238040549601eda60adac61f9453a9053bfc634c233a6515a5

Observation d1ed4703-bad2-484d-8e89-88bf35bfc607 · outbound

This paper cites Interpreting gpt: The logit lens.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Interpreting gpt: The logit lens

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.579925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:03.887928Z digest=sha256:36bf6d4a7811510e78d45f3c285b58e999f80b90779f81c27ee57abe6d3921ac

Observation 31df4916-8878-494f-8a3b-af405b58173a · outbound

This paper cites S., Goldstein, S., O'Gara, A., Chen, M., and Hendrycks, D.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models S., Goldstein, S., O'Gara, A., Chen, M., and Hendrycks, D

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.443129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.007213Z digest=sha256:1727ed20a0e29dfa8e23ace30936c48c484e2c0b08c194fb66eb87c1049c9ec9

Observation b9f1ec61-4e4e-4481-ae94-4bd3ba97bcaa · outbound

This paper cites Large language models can strategically deceive their users when put under pressure, July 2024.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Large language models can strategically deceive their users when put under pressure, July 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.324772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.146066Z digest=sha256:09cbf13f6da5fa63ae6634fe9f058eba5de8a266425e9bb708bb6be6199f8e34

Observation 218aef99-cf33-4e90-9326-c7eadb5a5ed0 · outbound

This paper cites an unresolved cited work.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:38:07.182106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.241493Z digest=sha256:5861bbcc76ab3166908990bf460c10442b0eb031feae8573f48f7bd61f5cfd69

Observation 72b9e771-10e1-463a-913d-1715f16c41cb · outbound

This paper cites Qwq-32b: Embracing the power of reinforcement learning, March 2025.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwq-32b: Embracing the power of reinforcement learning, March 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:07.052683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.365502Z digest=sha256:47c5b7d572ecd8b6164f17bbae7e86dfb74c567a3517b42ccd7058fb029d89eb

Observation 45882e9b-bb11-42a8-8948-87676ff49ba7 · outbound

This paper cites L., Sharma, A.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models L., Sharma, A

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.977885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.499667Z digest=sha256:5dca7569204e680d0924e1242215d439a2bf50158ab624c3d9ac62b981dfa6c6

Observation 16202906-50d5-480e-8771-259ecba3b2d7 · outbound

This paper cites M., Thiergart, L., Leech, G., Udell, D., Vazquez, J.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models M., Thiergart, L., Leech, G., Udell, D., Vazquez, J

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.634111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:04.595514Z digest=sha256:58878cacf50ce889d4e2ed7d4e05456f5316854367ca5ab7c9b51a0d0c1986fc

Observation 5b7709c1-7bc1-44a4-a501-4270a70d1004 · outbound

This paper cites N., Kaiser, ., and Polosukhin, I.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models N., Kaiser, ., and Polosukhin, I

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.707675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.707675Z digest=sha256:2ef773f630eb70ce58a40d55cb3b0a1fbc51d21a6d7c05acb8539766fd84487a

Observation a4a24f4c-66a9-4b16-83dc-98d75501138b · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.803196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.803196Z digest=sha256:e088465f55bd297079cc772ed6863b02b28d1c9ac88fe62a849ff8516837abb2

Observation c5497388-7067-482c-a3f7-2173f61acfe6 · outbound

This paper cites V., Zhou, D., et al.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models V., Zhou, D., et al

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:04.941743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:04.941743Z digest=sha256:ba301e44c636108d780e2a2b477716f7f10c31fcbc48b808489f74205fd61dd3

Observation 0a8dbc01-e3a2-4e56-9687-20e072740ba3 · outbound

This paper cites Qwen2.5 Technical Report.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Qwen2.5 Technical Report

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:05.057064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:05.057064Z digest=sha256:d3cb5b05a32f03da32241c0b154d2e4266950018d5b8731ade75919a3a4069d6

Observation edeeb8da-e7d3-422c-b360-568a11f9766b · outbound

This paper cites and Buzsaki, G.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models and Buzsaki, G

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.305864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.173034Z digest=sha256:7fc7845e73b4eab68e4ad9d4b6bd58e2dc1de4846cb2a6fae0c9a3f16b6b6f3c

Observation 9d6147ff-ac8b-4ef5-a946-2a7379e31e15 · outbound

This paper cites Reasoning models better express their confidence, 2025.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models Reasoning models better express their confidence, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:05.292046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:38:05.292046Z digest=sha256:8b283225cc84addae60618afc8e5e214b381b7acd4123e77bc235e7c5236cffb

Observation 2acd94a5-b29b-4172-b401-4f472583fc24 · outbound

This paper cites J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Kolter, J.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models J., Wang, Z., Mallen, A., Basart, S., Koyejo, S., Song, D., Fredrikson, M., Kolter, J

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:06.029672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.352408Z digest=sha256:27bd18d4c17e15812571a2e11192f01122e017d90951774dbaf7679d0c001b8d

Observation ce335cbb-4e4f-4164-ad63-f60edd888e08 · outbound

This paper cites write newline.

When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models write newline

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:05.813873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T10:38:05.468538Z digest=sha256:e303856fc8c724ee612cebe1137a748d49fdb94ced13cef6c5e85b786ce2082b

Pith citing papers

Observation d94b862f-9d0a-46a8-8e48-25a4526a7381 · inbound

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers cites this paper.

Adversarial Activation Patching: A Framework for Detecting and Mitigating Emergent Deception in Safety-Aligned Transformers When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:00:51.707342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:00:51.707342Z digest=sha256:a64ac28344c92509d6648fcbe058e6f1479b4051f570895a0b3454f04a9d4a71

Observation 0c106958-39d9-4aca-85ba-031dcddc8406 · inbound

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs cites this paper.

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:45.041136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:52:45.041136Z digest=sha256:06bf6babf15394d07044331aa32b99dbc30b8bfde0f2f120f38ba4c0c3d9e3f7

Observation 5d7330bb-6195-47f7-9d55-0d3ac29e083c · inbound

DECOR: Auditing LLM Deception via Information Manipulation Theory cites this paper.

DECOR: Auditing LLM Deception via Information Manipulation Theory When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:28:05.389621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T06:27:10.445757Z digest=sha256:9e81e61af4c0cd59f2ee8b27045153baed0e937a63f0838e99255d91de28dd85

Observation e4811b80-86b1-48d7-beab-a368e862aeaa · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.716480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:cc00ce2143a557c52ed42b3a35be9bd879ebcc863d29e58ab2ad9c4e96a0ec69

Observation 88499384-a89c-4608-a2f1-9f5addf98a37 · inbound

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates cites this paper.

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-07-03T13:08:07.647083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-03T13:02:46.485260Z digest=sha256:c1236c7c3b5cd830349c1177d55146207bf24718814df001786225790430f228

Observation 54f538bb-6061-47d0-a0c1-f4f38945e484 · inbound

Transcoders for Investigating Deception in Language Models cites this paper.

Transcoders for Investigating Deception in Language Models When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T01:06:13.525747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:06:13.525747Z digest=sha256:da3e84162a67705dbfc1edfff451f62faa63c801e8ff313a5bedf896017b7385

Observation a8ada123-d371-4e1c-99e8-b656e1ef1880 · inbound

Risky Business: Measuring The Faithfulness-Safety Tension cites this paper.

Risky Business: Measuring The Faithfulness-Safety Tension When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T13:40:57.774902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:40:57.774902Z digest=sha256:12ffc4d2973d5a5967aba5074254c350fee5979e9294065ca66f4abf98a1c27d