Pith. sign in

Paper Citation Record · LEDGER

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings

As of 7 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2608.04735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04735 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:51:52.279090Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact6
  • verified fuzzy17
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7847c1fd-240f-4c1b-8b72-b44e8ba17873 · outbound

This paper cites Inspect AI : Framework for large language model evaluations, May 2024.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Inspect AI : Framework for large language model evaluations, May 2024

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:57.025625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.828961Z digest=sha256:7f3cfaf286f30d0f1748bf6e761540544cef699e9f5d4b2319a9a5f80c82fe58

Observation d2be7f1d-5e18-467f-9912-e82813a20e15 · outbound

This paper cites Introducing Claude haiku 4.5, October 2025 a.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude haiku 4.5, October 2025 a

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.835810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.876010Z digest=sha256:6c95487ccfdb6eb0aad027e46bf231e44d5f1d1bfc4bbdb012173bd1923901a0

Observation e95b416b-de1e-4fd8-92ee-637b4e2811a0 · outbound

This paper cites Introducing Claude opus 4.5, November 2025 b.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude opus 4.5, November 2025 b

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.491211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.923038Z digest=sha256:6b55d67bbfe769c230ab62809948e965cc319e9eb81524c903a286619f4778ba

Observation 1a90898a-bc18-40f9-8b9a-f0032c091536 · outbound

This paper cites Introducing Claude sonnet 4.5, September 2025 c.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Introducing Claude sonnet 4.5, September 2025 c

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.238274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:50.968522Z digest=sha256:326b0204506e642b0ccc144435372314fcae10d017c9042d345cd0cc057d47c8

Observation 96f401f5-f452-466e-8c89-fc3e111287e3 · outbound

This paper cites Chain-of-Thought Reasoning In The Wild Is Not Always Faithful.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Chain-of-Thought Reasoning In The Wild Is Not Always Faithful

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.013543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.013543Z digest=sha256:a8419ba0cb69764adbe3609326ab9e621990b36d954d9a4d87ab077aaa27a7f4

Observation 366a74fc-0792-4b39-b5c3-c60280eb4461 · outbound

This paper cites Biases in the Blind Spot: Detecting What LLMs Fail to Mention.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Biases in the Blind Spot: Detecting What LLMs Fail to Mention

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.079079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.079079Z digest=sha256:9c3ee4c6caf80556b396ffcc6562325a8640776ea0cf6b083ad3461a418ca53b

Observation 3b791013-7f29-4d0b-888a-19e074ff7dea · outbound

This paper cites How does information access affect LLM monitors' ability to detect sabotage? arXiv preprint arXiv:2601.21112, January 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings How does information access affect LLM monitors' ability to detect sabotage? arXiv preprint arXiv:2601.21112, January 2026

Reference 7

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.773964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.133048Z digest=sha256:69de4d68682590ae22aa42aea11ca85b4414276f18b174eb550f1ab68c60bb74

Observation e411a152-ded3-44c0-bd5f-77beed122afc · outbound

This paper cites CoT red-handed: Stress testing chain-of-thought monitoring.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings CoT red-handed: Stress testing chain-of-thought monitoring

Reference 8

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.696529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.189440Z digest=sha256:4a13e8031d6b4283d042e6ab9549924f1be109cb765f6f23c64ff2358f48b406

Observation 0aa7a58c-ad9e-47a7-99cb-42408b2e00b0 · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.240956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.240956Z digest=sha256:3989deb9f269f69d30610a725fc3d1c29318fba69de477530fd11cc96a647708

Observation 791437e5-9cf6-41f3-beb7-14463f28461b · outbound

This paper cites Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Value Leakage: An LLM's Answers Are Silently Shaped by Its Own Values

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T17:51:53.503854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.301008Z digest=sha256:4059b3146ea8602a4668e6b7311b478fb9f3e947459a3425282f00936545061d

Observation ee052de4-fd50-4743-b7dc-1918bc7d93d5 · outbound

This paper cites Censored LLMs as a natural testbed for secret knowledge elicitation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Censored LLMs as a natural testbed for secret knowledge elicitation

Reference 11

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.622765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.364168Z digest=sha256:a440ecfdaefc4016547c7b6e31ce5afc2372ca288e8ed1c79b11da4c4ffd0eeb

Observation a1af007d-4c82-4494-8810-dfe623e17938 · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Reasoning Models Don't Always Say What They Think

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.482929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.482929Z digest=sha256:c91e581ca7323b0718d1b2323b22f7617b6de7423482408fa3cf8e0faea3116f

Observation 31bd8a52-def3-4980-b486-f4912b8a29b7 · outbound

This paper cites Lee, He He, Ian Kivlichan, Bowen Baker, Micah Carroll, and Tomek Korbak.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Lee, He He, Ian Kivlichan, Bowen Baker, Micah Carroll, and Tomek Korbak

Reference 14

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.542529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.556485Z digest=sha256:b6ca10d3535c01db9e2ed30e1b40be659fc5f6499c1801a81839ebb849be1770

Observation e5e97ee0-eae1-4eb5-b90e-08d1df8d9865 · outbound

This paper cites Are DeepSeek R1 And Other Reasoning Models More Faithful?.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Are DeepSeek R1 And Other Reasoning Models More Faithful?

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.627288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.627288Z digest=sha256:a1dc7d783d058dfaf445e62b945372fb4f4d7c2550551cba0821d2b3364a7f8b

Observation f3d8dc31-19d1-4cda-9ead-c45b620d98b6 · outbound

This paper cites When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:51.719517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:51.719517Z digest=sha256:93755b5269a80b9e670746f10cb6a70c08c84b9212b0c5eb6e3ea4cd66ec8d91

Observation e76d3e12-a314-4abe-94a4-1407cb4be026 · outbound

This paper cites (some) natural emergent misalignment from reward hacking in non-production rl, March 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings (some) natural emergent misalignment from reward hacking in non-production rl, March 2026

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:56.031994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.834128Z digest=sha256:6dd3b82b798b24d58420afddf03bba82ee4aa038c34f1d065391e099f3cc2ba4

Observation d3aea62b-7a06-4803-9dad-47dc30961199 · outbound

This paper cites Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Guan, Miles Wang, Micah Carroll, Zehao Dou, Annie Y

Reference 18

Resolution
verified exact
doi, observed 2026-08-06T17:51:52.467492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:51.946000Z digest=sha256:8cba25552002b66e9fe8dbc99866ccfe59ecf04c1d06fef19b149a7dd15f60df

Observation 8cb673e0-48bc-460e-b3d1-bed4380fa23f · outbound

This paper cites Noticing the watcher: LLM agents can infer CoT monitoring from blocking feedback.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Noticing the watcher: LLM agents can infer CoT monitoring from blocking feedback

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.058436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.058436Z digest=sha256:2dc879f83711bb5f1736e4a07ba7ba5c9a091c4f829a81d8e7c69702672cd7dc

Observation 3e760a5c-06ea-4874-9859-3c1560cfed5b · outbound

This paper cites Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.127591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.127591Z digest=sha256:b73f5f43326286c29ba24187f1d1e2e76c7bbc66039ebade6a5058e67a8b7690

Observation 22c44c1e-22f5-4d04-b6fd-fce58ce9fa4f · outbound

This paper cites SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.183344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.183344Z digest=sha256:45c6fe9799b5e9d6e65d48bead9923eb496a72d5a038f3dcd9bb32aa8a2998c3

Observation 1876190b-3bb6-4f5d-8e47-bda706cf4f6a · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.186621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.186621Z digest=sha256:18f17e6654b97cbc2db25026042ce07181f020582c95c00c371d0313a4279b10

Observation 1c2634c9-8748-435f-89fc-1556adf7a8e6 · outbound

This paper cites an unresolved cited work.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.188944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.188944Z digest=sha256:e8dd699504c11bd069f84113303f258864d00f034e1d86b71b17ab37e543c7f4

Observation 3651c42c-9b21-436c-ab27-1bab369e683f · outbound

This paper cites Kimi k2 thinking, January 2026.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Kimi k2 thinking, January 2026

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.762764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.191108Z digest=sha256:2fe7ec6de3c8baa1f66655026a85c4136c108ea78b1888c14e692a9dbf731780

Observation 4760181c-3f11-4e6a-8253-4d579d9e9421 · outbound

This paper cites gpt-oss-120b and gpt-oss-20b model card, August 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings gpt-oss-120b and gpt-oss-20b model card, August 2025

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.503751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.193355Z digest=sha256:27a50cdcc9c82bdae9c7b4e661c4278c389faa738d13629edf050d1bc2aaa155

Observation d17cf599-3150-4a5b-8cc7-14591bd740eb · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Llama 2 via Contrastive Activation Addition

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.195387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.195387Z digest=sha256:1c7c964658b1634b28f0d61173b4b24984f3aee60fa4ced0635ca08bd2946d07

Observation 4269579b-a6a1-498b-8a71-ec6210044950 · outbound

This paper cites Large language models can learn and generalize steganographic chain-of-thought under process supervision, 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Large language models can learn and generalize steganographic chain-of-thought under process supervision, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.200090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.200090Z digest=sha256:7bd2c630eeabccfd3b9ceb87301dd5c04048bb5655bda8731b941672cf55ad5f

Observation d19fe64f-252a-44d0-be19-488d429369ef · outbound

This paper cites Steering Language Models With Activation Engineering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Language Models With Activation Engineering

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.202208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.202208Z digest=sha256:d40e889fae62569c11dea6664970950e4d29bc8bd28ad4071240ba19b496445e

Observation 1e354193-0b20-4880-bdd3-940b439a681b · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.204689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.204689Z digest=sha256:5e488fd0b8632dbb327914e307a6751fb4449122c0ad3bc2a5adceb19732adca

Observation fc5cb71a-18b1-49eb-a2ec-5d8a375928a1 · outbound

This paper cites Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.207562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.207562Z digest=sha256:3d29a9a5ba2d5651b80797feed71b7f7bc329b49fece1197ad270edf850d7724

Observation d43bd6ce-84b7-4f2f-b8b1-aa28a045cfa8 · outbound

This paper cites MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings MMLU-Pro: A More Robust and Challenging Multi-Task Language Understanding Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.209806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.209806Z digest=sha256:e54042e9dac8e479be4406493f5b90d109a2dbcc3dcfb23e20860a6902018ad8

Observation e8b455d8-8d05-46f3-859c-f38727e2e722 · outbound

This paper cites Grok 3 beta --- the age of reasoning agents, February 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Grok 3 beta --- the age of reasoning agents, February 2025

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:55.227558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.212330Z digest=sha256:fcf8411abb2d92ba2c3a585a1fcf9094ec6dc9dbb3583ab4c57c0941f6491edb

Observation 54fed9a0-8407-4395-ae8a-33038bcb0cfd · outbound

This paper cites GLM -4.7: Advancing the coding capability, December 2025.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings GLM -4.7: Advancing the coding capability, December 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.965790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.214472Z digest=sha256:4e321200918b37e72b3edf5cfc88388a4dfe5fb6032e6ed9afbcfab71e8424bf

Observation 13fb16fc-05d5-414c-8ed0-4e4729999ed4 · outbound

This paper cites Can reasoning models obfuscate reasoning? stress-testing chain-of-thought monitorability.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Can reasoning models obfuscate reasoning? stress-testing chain-of-thought monitorability

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.216459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.216459Z digest=sha256:6b2dcaa7ab2495b0e77b860e19ea6bae8953546ff7dffe4d1d86c8f834fc55fa

Observation 2437502c-1a34-4113-ae1e-bba68ec7ea31 · outbound

This paper cites 2023 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2023 , eprint=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.218672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.218672Z digest=sha256:23ea41bac157d13f34af2cde95d2bf552a32ff0b263ea83ae50d125a07cf99aa

Observation 927958a4-176c-4630-b2be-86610271f78c · outbound

This paper cites 2023 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2023 , eprint=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.220659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.220659Z digest=sha256:602a8ab6c2283df12ad9cca9e33bfe4e541bc9e22d2d6a0d2a635b6b3bf97d16

Observation 98c666c8-bd05-4431-8064-c1bb5dc17370 · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.222703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.222703Z digest=sha256:d1a4af2043346971c0be74d8449f1726356520b38ea130effe17952f10f08a76

Observation 4f9459d3-121b-4223-a87d-8015784b7ea4 · outbound

This paper cites 2026 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , eprint=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.777138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.224969Z digest=sha256:a40afc262773250f6523425a224c138e760cbd91f0d739d888015059c34a7187

Observation 5178187b-cf63-4f2c-8935-92370c368a2d · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.665282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.226927Z digest=sha256:7a1a8b4ef45a6566488f4bf8dfd809b31ae9be707ed70cb05e8f5a1a359729ba

Observation c9b6f91c-e608-4553-900d-9d7171c2293c · outbound

This paper cites 2026 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.582474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.228836Z digest=sha256:cecaf273e44e875f47915a0b58b4a62c7de89f45023476c59969f6c762d15543

Observation 8a77ec12-81e6-421d-95ab-c5a11106bdca · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.230732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.230732Z digest=sha256:cb43c60097564d36bf650458234e697f443bf03006e4b8df74796779750b5e77

Observation 2216db55-6d18-447f-abcf-399d79cf8587 · outbound

This paper cites arXiv preprint arXiv:2512.18311 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2512.18311 , year =

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.232686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.232686Z digest=sha256:0ed77136be63f9a463f55deade3cd470d5a528bb5c5f75cefb797de83c8dd860

Observation f2441273-de61-4d74-98c2-3bd6b0405e43 · outbound

This paper cites Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.234780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.234780Z digest=sha256:94c5f84acb37d7d18b040d4dc2cd9bb350dc9fd4c8b77ee9f8f85b13a5c036e4

Observation 03053e1a-46ae-4826-8f62-66d4aa37aeb3 · outbound

This paper cites arXiv preprint arXiv:2603.05706 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2603.05706 , year =

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.236856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.236856Z digest=sha256:0ea9653b3bd32dfb0ced510cc0df5052cec37c0006f49420ae9b8dd62b8265ad

Observation 73fad7a8-57a9-467f-9bc2-e574c9639c7a · outbound

This paper cites arXiv preprint arXiv:2510.19851 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2510.19851 , year =

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.238875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.238875Z digest=sha256:448c5eb5dd78f2962dbe5b8c051aee5fcace479a562f93cde42c809e57a52ed5

Observation d596953b-7e25-450b-b47c-9a8a9b4bc214 · outbound

This paper cites arXiv preprint arXiv:2505.23575 , year =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings arXiv preprint arXiv:2505.23575 , year =

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.241151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.241151Z digest=sha256:e85b6af9b9b8042a7851610d9451d8e32bccfb86ee3607787bdee18e2279d916

Observation b9bf8812-08ae-4748-befb-961011b835f5 · outbound

This paper cites When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.243168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.243168Z digest=sha256:a8770efd526957616747fddabc6ae6127d7a1be2e3d191facd133815fc2e3004

Observation b8b8928b-266a-431f-8ecc-d9792f291cf1 · outbound

This paper cites Noticing the Watcher:.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Noticing the Watcher:

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.381037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.245734Z digest=sha256:278672374b6124ff8efa76a8383644856c8bc2e595a674d2a2312e362f98d5fb

Observation f77f12fa-2906-4938-afa5-a16bcb1a9463 · outbound

This paper cites SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.248062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.248062Z digest=sha256:2f028d3ae6bf1f4b2638d633b9ad8b525752a4540c67e3a8ccfc03ec7b9e74be

Observation 9aedb683-1bd7-40c5-815c-4ffddab805dd · outbound

This paper cites How does information access affect.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings How does information access affect

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.250176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.250176Z digest=sha256:c5c90b01ed9f0c2242709da2cc4a22dafcebf3793f8c170a1d5a50335443738a

Observation 9f4c68f4-055c-4a3b-a21e-2af33c9e81f9 · outbound

This paper cites Censored.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Censored

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.251993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.251993Z digest=sha256:4a5f95b96001316b60d21456be73dfeaab812a74fb880e5783b90257eb4a041e

Observation 8866e0b7-52c9-4da1-818d-d74a206bfd76 · outbound

This paper cites Steering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.249447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.253855Z digest=sha256:21180686e943bd9a3706588a36ca69972e406a5a23f4a03e9e8805a521f22f12

Observation 1bde4f1c-47e6-4640-b119-782c5d80421e · outbound

This paper cites Steering Language Models With Activation Engineering.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Steering Language Models With Activation Engineering

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.255930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.255930Z digest=sha256:d572358e65bb748d18215e6b0bd8357ad659ff977fc8f2bcd274be4a0610f8e6

Observation 47e18b2a-f703-4c80-9d3f-be93e2d71ed6 · outbound

This paper cites Are DeepSeek R1 And Other Reasoning Models More Faithful?.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Are DeepSeek R1 And Other Reasoning Models More Faithful?

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.258144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.258144Z digest=sha256:00b68c4a1435e263dbbb171161a629cf2e840eb809cf28e812670b2049787a02

Observation 27e34088-3de7-4aec-a62c-7553b5bb1345 · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Reasoning Models Don't Always Say What They Think

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.260226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.260226Z digest=sha256:b4481f2b061ab36553a5d299a60f729d17011551e3f3e74a3e30e4ab9112bebe

Observation 8daac443-4250-474f-8592-87fb3f5eef50 · outbound

This paper cites Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Teaching Models to Verbalize Reward Hacking in Chain-of-Thought Reasoning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.262506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.262506Z digest=sha256:ec7a64115e52ceefcdaf74f979534309d30575d53b86f0e5c74e8f07db122288

Observation 10dc86da-4f31-49dd-89d6-1d84367a99da · outbound

This paper cites 2026 , month=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2026 , month=

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:54.034479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.264644Z digest=sha256:054206da7aebc6f5d111b8bc9a2de5af5d2aa1d867f230f7d2308748e76f758d

Observation 0370dbc7-3f1a-4713-a0a5-382c4d7f45b6 · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:53.801715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.266564Z digest=sha256:6361ac5786a4b64d2b34a4bf82c88dc10a5582c1a17b1b1cbb5b2c74a8251a24

Observation 10862a44-87c1-49bd-a152-b497a20cabab · outbound

This paper cites 2025 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , eprint=

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:51:53.612616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T17:51:52.268522Z digest=sha256:f9bbe5c51e51c5ceaef094ff77b71bb89b0a24095ee32fdd2a8805e8f3bc83ae

Observation ff7822dc-7acd-4f27-a4e3-83cd1b74624f · outbound

This paper cites Humanity's Last Exam.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Humanity's Last Exam

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.270575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.270575Z digest=sha256:0030e4e17785361ffc7fca32700497b1605d406f70af1dee8e79f5420bcd1c63

Observation f68615f0-7475-41bb-8659-4c087ba4a378 · outbound

This paper cites 2025 , month =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2025 , month =

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.272644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.272644Z digest=sha256:a6fbd48934805375c4e8f28de5090d21b62697b2948f3fb9d86fbfaee4b629e7

Observation 8e3d9c8b-8daf-4730-91e2-6918e9a77e7c · outbound

This paper cites 2024 , month =.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2024 , month =

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.274605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.274605Z digest=sha256:901e5a556b159e505f49f45dd4d51f14e88e9c23117248f1e24f84b0147d4750

Observation 5a939a46-89a1-43c6-a786-78eb546c2959 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.276967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.276967Z digest=sha256:73c4f414db8facdb34d81103dafbb0f36e81e439746c327e90a5f6bc204a0e76

Observation 7975a77d-4bac-4655-b354-d47530af65b9 · outbound

This paper cites 2024 , eprint=.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings 2024 , eprint=

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.279090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.279090Z digest=sha256:4997259002268abb3eff445ac416d1ffc1bf104c42e2e59c8b0af8e97939afec

Pith citing papers

No inbound Pith citation observations are available.