Pith. sign in

Paper Citation Record · LEDGER

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

As of 10 August 2026, this Paper Citation Record lists 100 of 115 outbound references and 78 inbound Pith citation observations for arXiv:2507.11473.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.11473 v2

Coverage vector

measured 100 of 115 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-20T14:19:44.695462Z

measured 178 of 178 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 78 of 78 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:51:52.127591Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

100 of 115 outbound references displayed

  • verified exact11
  • verified fuzzy78
  • unresolved3
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch6

External citation measurements

2
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 2aef35ed-b3ff-4c9d-b1be-a76fbcb70bc4 · outbound

This paper cites AI safety via debate.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety AI safety via debate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:19:44.804131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:05be777f1a47431a92c00a9b7d120eb1fee63125352f9e757cc2ebf59495664b

Observation 9b047c66-e8e2-47bc-b8b6-0be905068fce · outbound

This paper cites Fine-Tuning Language Models from Human Preferences.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Fine-Tuning Language Models from Human Preferences

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:19:44.783757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:7ed337e25efb953faae419555c9e274b1878d05f99d6bfa0cf77a370199f32a3

Observation e6636576-0880-48bc-b5be-d61325ee222a · outbound

This paper cites 2024 , howpublished=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , howpublished=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.988055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:eefad0c261916b50d701913f7bf64e51de4840168320caa8672cb696b489c819

Observation 3b158cbf-28f4-4ac2-bf11-16c27e367374 · outbound

This paper cites The Checklist: What Succeeding at.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety The Checklist: What Succeeding at

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.989804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:aa69d1fc20e68ec274ea8f3d92355d80665005de8b4a87a26dad36f471e43e8e

Observation 4b9e95c8-ae1f-4e1e-94b4-af488562dd9c · outbound

This paper cites Scheming AIs: Will AIs fake alignment during training in order to get power?.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.780078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:a6ec01debc69c99e94868b94c8373101b811f3b5c49d9ab9889db8a5a176a3a6

Observation e0c65d65-7d6e-4db3-931d-12fa2e3007d9 · outbound

This paper cites and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , journal=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , journal=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.991922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:5594e92dc6262ef6e6ccb2b9e1e0c3a4d6e53361d35650742d6d6213da7386d3

Observation f3743c47-da26-417b-b5d8-d29dd48572d7 · outbound

This paper cites an unresolved cited work.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-20T14:19:44.993921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:ef2c2949721da9a009df07d96e7c526ae8b3d4e4dd40c2f2b618c02355d05e3c

Observation 9efc77cf-7478-4b84-810b-75157149f0e5 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.995965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:271e1e691a8eaab29ff8e35f78be47ba5b6653ef96a38b8bb05977fcffa231e6

Observation 825d166f-3522-47b3-818c-c6ec0fc90406 · outbound

This paper cites Safety cases for frontier AI.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety cases for frontier AI

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.788052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6f1553ba4ecaa9c2ed3363ccec703c7326e30d8abf0e0eb564e3437a7bd923fd

Observation 568afe59-59a8-4d94-b118-19056f047797 · outbound

This paper cites Safety Cases: How to Justify the Safety of Advanced AI Systems.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety Cases: How to Justify the Safety of Advanced AI Systems

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.738859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:1b80bdd8b0204788a863b3caade6b4f913c85ac4ff613fcfe54a89e3fa6afd15

Observation 50cb30c1-7088-4f5b-ace4-a84181c42154 · outbound

This paper cites Safety cases at.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety cases at

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.997695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:d04e91e95a29e069dbca0fa863537eef13bf32b2c41753be4d0730ae14e0d3f3

Observation efff4b1f-3d17-4b9c-9bf1-457596ed7d9b · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.999449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:ac295d9e62fc2a4c42cd32a66760379f93c2e4ea87fd540c354d2e2c7d8cada9

Observation 6268b255-31c4-458d-bb0a-a812173d80d6 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.001359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:313a83f328f408b1290191bee4c7d80de6ceb910062a810820f34ce7a69b2f9f

Observation ec62a315-c042-449e-9333-a533529262ee · outbound

This paper cites and Lucas, Caleb and Guest, Ella , year=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Lucas, Caleb and Guest, Ella , year=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.003856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:18342c3e39b947451c9d59d8e9313fab4d51723e0c0f7a100e9d11b25898531d

Observation 3e975ea1-46f7-4dd6-b3cf-7ad09b292068 · outbound

This paper cites 2024 , howpublished=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , howpublished=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.006090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:0a5aa6a8f6d9eb86205cff71b0382e9f70dd482286686a3cd9f2e9ce08353550

Observation 7c1b270f-5ff4-4c15-921c-5ec82e65f3ff · outbound

This paper cites an unresolved cited work.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-20T14:19:45.007928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:e448cd7afbed2e3f24d15a8be8c40530113eb0c8461e88c4554eb259a9bbdcd4

Observation ee9fc45e-13c5-40df-9051-bd978a2f9fbd · outbound

This paper cites 2007 , institution=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2007 , institution=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.009660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:7b488f4309166cd4a3c192c54e02d902f8fe2b6530d7477d85ef49c3310c40d4

Observation 731272a7-e51c-4346-a0d6-9b9f4d872278 · outbound

This paper cites Managing extreme.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Managing extreme

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.011718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:9e85d1757583f1b609247aba2913ed8cb7a59a4cadd64a98aa11c45f4ea6511e

Observation 37f5c1e9-5c7c-49e8-8fec-c1c26b191dad · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.013409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:83bed2665ac7f463f60f0d2e5efd3788b92e0c9fe908684552a4c1e2a084224b

Observation 8f092750-c822-4864-b6c9-1925435fd126 · outbound

This paper cites Safety and Reliability , volume=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety and Reliability , volume=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.015153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f8219f2ba8f3baf3ed826b2a3c6798fae039b3664cc340cda6bf802470def137

Observation 488b22dc-180f-4ad0-922c-132093072637 · outbound

This paper cites SPE Asia Pacific Oil and Gas Conference and Exhibition , year=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety SPE Asia Pacific Oil and Gas Conference and Exhibition , year=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.017112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:12bf6b2ecc0dffb4cf1a51796248f29867ff5d97700037d0ba85c84f3b9ad441

Observation c5cc610a-52d3-43ab-b84a-76af7a08b6e6 · outbound

This paper cites Safety Science , volume=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety Science , volume=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:45.018894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:427e6a76fc6d4e314b6145bb3fdfeaf20acb13b2460c8483b0d187a967fdfaa1

Observation 6351afb8-c122-4ead-adc4-394d993631a1 · outbound

This paper cites Safety Science , volume=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety Science , volume=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.806657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8a99ec515e6c2646044c4a6a69c14b1eb476ec539b2676fb60ff9d2e3ea380c3

Observation f7aeffe1-2e71-4fac-8bae-a45f34e4f3a8 · outbound

This paper cites Foundations of Computer Software , pages=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Foundations of Computer Software , pages=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.809216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8a393ad56b46cf5adee7c01faa4d2ec90190ff314037dbf68dae79b8f2babff8

Observation 7d127e1d-0a79-4c4f-9c35-270233d74633 · outbound

This paper cites Safety Science , volume=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety Science , volume=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.811558Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:4f925af3c110300588ddaf9f94f7741b408f2e3cec9dd500e420ab2e0676edda

Observation 34f5d9af-0338-403c-8696-fde5b488e3e6 · outbound

This paper cites Safety case template for frontier AI: A cyber inability argument.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Safety case template for frontier AI: A cyber inability argument

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.792729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:9899858a11c7cf628ce25751f866a7a35519d566d6278d2c24670f048458437a

Observation 2196496e-8316-41b8-959c-d7264084d7d4 · outbound

This paper cites Towards evaluations-based safety cases for AI scheming.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Towards evaluations-based safety cases for AI scheming

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.800220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:e85dee5cd55060a4e44330c1e8c172d87ac7421ec4ba56ed3bb947a961fa21e0

Observation c02eb5d5-ec65-4923-8ca2-e87387523acb · outbound

This paper cites booktitle=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety booktitle=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.813756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:965472c466831f69b7d20a89bbfbeb7ccf203663d27fa195b1c2f804778c409c

Observation 7eca4a44-634e-4288-828e-ad2fcd68d077 · outbound

This paper cites RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety RE-Bench: Evaluating frontier AI R&D capabilities of language model agents against human experts

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:19:44.776446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6641913e1e402f3869052968600f574ee27f72c2d3da13012d99218ccc06cbd1

Observation d9c04503-1bf6-4a8b-8ca8-efe930d06bd4 · outbound

This paper cites Shell Games: Control Protocols for Adversarial.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Shell Games: Control Protocols for Adversarial

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.816032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6e7b2669a45cd79c235000bd1e56e4fcf2e92fc694d26524f20bc9abf7c75639

Observation 0e5567ae-1564-44ca-ab85-54cab6c5f213 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.818066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:df9fb93507cadada9702e06e37d153a99291fde37066a35425cafcd7a214d4d9

Observation f97d2b4b-521b-43a4-b985-6cc7b72d6e90 · outbound

This paper cites 2023 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , eprint=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.820445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:a0bce064e0e15f71342a1d47527cad633f1b3f1f92c7f804b0f6909bea51fa6a

Observation 64a08022-20a5-4d68-9b28-4567c931f117 · outbound

This paper cites 2023 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , eprint=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.822800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:4cd1ffd0c96383e7fedff754c30d928769c1c8cb67303dae3d0ca62c27c1280a

Observation a6f60f66-1770-4004-92d4-9139ea957be1 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.825016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:098d3dc7607fbea5c4fba1251faeb383a2ebc50395985a711e39876156e045a9

Observation 1b60cc92-d048-47d8-a052-341217d6a80d · outbound

This paper cites 2018 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2018 , eprint=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.827273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f563036622ddbbc95b960558f662bc819b2201435164b3ee3a28e2d4da0f2b60

Observation 6ff7940f-0a24-4028-99df-701ba6085d86 · outbound

This paper cites and Phang, Jason and Bowman, Samuel R.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Phang, Jason and Bowman, Samuel R

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.829537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:b964fce567eb2ba89ace1a714e4d26e5ec9377a8241e1f14e85878d957ed1790

Observation 88f66590-9c8d-4b7d-8c27-e0fa8229c420 · outbound

This paper cites 2020 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2020 , eprint=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.831791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:94b8bc9c9cb68da0dc2c3df2a45b840037ab249b9ab9b3c49947d5878f220a00

Observation c042567e-c29b-4bc9-86be-498a8ac4b15b · outbound

This paper cites 2022 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2022 , eprint=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.834280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:da5fdf70ab97559a48985eceee96a97fec44674f2487c5b98470bdae5c9e5188

Observation ff5ed61d-215a-4d42-aa28-6372369c8de6 · outbound

This paper cites an unresolved cited work.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Unresolved cited work

Reference 39

Resolution
parse uncertain
raw_fallback, observed 2026-05-20T14:19:44.836641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:bc614e2914d2c7a18d76426eb999ca80da98bea51058554003c3ad914b28884e

Observation c749b60a-efe5-402b-bf40-8ef3c84b06a2 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.838685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6ae5acf7338b7fb2850772f227a0e7169d9f4086350470a42cbf0dfd85d477d5

Observation e5fdb9a6-96ee-4b6c-a491-f98581034778 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.840703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:4c3087982628d03ce221e6f992f66783e8567c513f1625da8ce6788f2a99e0e6

Observation 16eb96b6-2ddd-4c88-9dc0-e88542b15e8c · outbound

This paper cites 2024 , note=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , note=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.842808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:a6e76e7022a50646f6bc8ab5b07716f9bc1a7d8065b1167d6db4b6e4cc84c16c

Observation 615e9e43-3fdf-447d-ae62-bbb75929e08a · outbound

This paper cites On the Choice of Loss Function in Learning-based Optimal Power Flow.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety On the Choice of Loss Function in Learning-based Optimal Power Flow

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.742672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:2d106b1220536a923ab98e2eec86948570178489ac5436ad520943b83876083a

Observation d82d28eb-19e1-4866-b18f-23d8360e1bd6 · outbound

This paper cites LLM Critics Help Catch LLM Bugs.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety LLM Critics Help Catch LLM Bugs

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:19:44.746613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:01abf63ccffdc74d54fc689e9f70f2fd9c0f558aa0d516e6f6b578801ecea3b0

Observation 1bf448f2-6dcd-4830-905a-704b7d8ece0f · outbound

This paper cites Preventing Language Models From Hiding Their Reasoning.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Preventing Language Models From Hiding Their Reasoning

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.761298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f453548b1d0f6edf0ca3581ef6fca93a8e856e2e49ac7a2b7f4ebfd1fd98be47

Observation 152d93e8-2efa-42b0-bd22-5bf803dc541f · outbound

This paper cites Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Games for AI Control: Models of Safety Evaluations of AI Deployment Protocols

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-05-20T14:19:44.765406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:613c1d48aa78542ff3e784b79069cb3297c989e7b634760e7c9865049124b037

Observation 843fbaf1-6657-4a52-aa54-6c50da9bbb74 · outbound

This paper cites Risk thresholds for frontier AI.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Risk thresholds for frontier AI

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.772240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6d98ed397e5cd93352835032edce9cac03d9ad7a1a48f78fc15697ec9877cda7

Observation ee3c1419-8552-484a-889c-4175ae6720e2 · outbound

This paper cites A basic systems architecture for.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety A basic systems architecture for

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.844717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:513c6c73bd3789f53d1e062962c7ea451d5ebc67d1c4f2df94f58f89d651020b

Observation 32a09101-1490-40d9-a4e3-92bec89ca4f8 · outbound

This paper cites Three Sketches of.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Three Sketches of

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.846608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:4e655b76bed51f5d2de5accb7f41e1d9f935b69ac712e3bb68041683c0de79d3

Observation 9ccf0494-3037-42b2-b7c5-dc92560dfbc0 · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.848747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:178827d3e25b051f772e0f0f55f46370b9ee5d9e453323a44e6e26545bcb3b0a

Observation 2d69419d-2e8e-49e0-a240-240ca7e18664 · outbound

This paper cites and Thomas, John P.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Thomas, John P

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.850556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:91aad6b59fb88fe85ee6c6fd209f7c5193c5aa82606bf0ee3f6ff0f973768e4e

Observation b224a236-2fca-498c-a067-20668f432e53 · outbound

This paper cites Failure Mode and Effects Analysis (.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Failure Mode and Effects Analysis (

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.853294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:7bc5b6d56abf60484a1fd142e6423a0037b158db47bc520398ad2f5036471c7f

Observation b5ae26a7-c9db-42c2-ad9f-c81fbf261b9c · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.855452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f265f46cf521da7eb1c0d005d15d0f91803a11b9b9616432f857345eca4630a7

Observation 7f9f0fe9-2767-4f8e-a04b-f6c57b10193f · outbound

This paper cites 2025 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , month=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.857705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8871891e47f752e3e664841568e0f5db9bab0900a083741ea6367e2b36ca3e50

Observation ff161714-8109-488e-83ce-718457c7b30e · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.859825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:c24515a2ee0c8918b8a8929d86f3566614d087b6444f92f2b680f31b804a14e4

Observation fe62e661-360b-4116-ab92-3a0208eba227 · outbound

This paper cites Building Blocks for Assurance Cases , year=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Building Blocks for Assurance Cases , year=

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.862536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:ae70e872adcea7910cdcfb0b1898d25b5ba71e8266297bb52fcc96839aae81c4

Observation 73b49422-0a3f-4e68-9ad6-84527748bd9b · outbound

This paper cites Thoughts on the conservative assumptions in.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Thoughts on the conservative assumptions in

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.864684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f67f600782ada4f0fce0a443144557cc5f81a64f52e8873bf950538c34c45e44

Observation 623c5053-6f99-45c4-89ed-81e743051059 · outbound

This paper cites 2023 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , eprint=

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.866896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:2731d23e9494cd2a551d8ea404c516c72b0a424d58049c632c4520ffb69682d4

Observation f1aa6499-97b0-4b0a-af9b-ec8edbaff01e · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.869029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:cb24a0726fd39ed89e5553dd5ccd7ac71f020dd12784a8ee5e27505261d7abcc

Observation b39906e4-ced2-4f77-bb60-2862e6b91dff · outbound

This paper cites 2023 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , month=

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.872273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f7483f38fb5546f7a3e90cf45d264424e28df24f153906225880b6835c9042ca

Observation 6107eaf7-753d-4ff6-801b-e4dce07f1d84 · outbound

This paper cites Me, Myself, and.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Me, Myself, and

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.874443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:46c17f600a755f8888a1e0cdf31236982be0c801ce828af848a338a4a6ae2121

Observation 48285e7b-bd2b-4617-886f-846adcb68f20 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.876786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8423e01b366ddb3382f11657c28016f003fe50624cc2d4fb0fce13779598574e

Observation 3f829c75-f151-4d60-9cae-5fc3f29481a8 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.878845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8ea21dfc274907c8d9b99621e0561c144d19694f0d76be0337c76249da29c561

Observation bdc9b807-0b7a-44d4-9c08-aae334363185 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.881300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:d4528202a3bfd45391bbb46dc94b005a5d4df517fa43a1acc207c823003f0c78

Observation 4980c193-45f4-43f7-8e94-584de216fc53 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.883579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:507a69adb7e2920b29376d084fc95bef34d14f6cd7382f50d38ff63ea98be2f3

Observation 86689f15-1f61-4b6b-b69c-1b81933b5801 · outbound

This paper cites The AI Agent Index.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety The AI Agent Index

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.732817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:5f60bdaa8360c89f1b5525130cb6f4f64cf63ad42110cdb9b1630d3dddb8f5b0

Observation 6756eed0-be02-4a1a-9159-f00ed8b3799f · outbound

This paper cites 2024 , eprint =.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint =

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.885465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:06cc92e196b1cd8018d61f051746dec6203c7da201e60653f5631312afe877c7

Observation df00cfca-3f4c-405b-a1c4-0ead86431c57 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.887411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:3a2949ceaf15767955a9dedc7de03115336b50817214a746ac4241aab8fccaf3

Observation b895fe83-923b-4400-854c-079c47c17e40 · outbound

This paper cites 2021 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2021 , month=

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.890070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:5593097ff6cb6489d59a3c9b3e85b6b00b008cd4cc7127111516edde63792f93

Observation 487f6696-aff3-47ac-88c6-9d9503abc0a2 · outbound

This paper cites Recursively Summarizing Books with Human Feedback.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Recursively Summarizing Books with Human Feedback

Reference 70

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:19:44.750729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f42c099647ce993f653f52fa297cab97bd95bfbec165504d9a9a5a2380b39d0a

Observation 61ba18f2-0447-455e-a6ff-50664e8f0a02 · outbound

This paper cites Self-critiquing models for assisting human evaluators.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Self-critiquing models for assisting human evaluators

Reference 71

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T14:19:44.756323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:8a46ce1430c0cf7067795c6d320ad437ad59665d0917f1f5581f3297c0b9d4c5

Observation 423fa106-fc10-4c3f-a54f-eebcdbf11d74 · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.891977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f08d0e39ce697cdcbd7c868920671228cc42503c6ba191037b1534e77fc3e9d7

Observation 1f4ed779-0a44-4586-9309-7d92fed9bb5f · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.893792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:e3a82b7b51b3c4b3418597ddcbb5b31149ab163dc59aa8a206fce7d797858b78

Observation 54f92d21-0e0e-4fa9-86dd-d37ccbf120c2 · outbound

This paper cites Addressing Intersectionality, Explainability, and Ethics in AI-Driven Diagnostics: A Rebuttal and Call for Transdiciplinary Action.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Addressing Intersectionality, Explainability, and Ethics in AI-Driven Diagnostics: A Rebuttal and Call for Transdiciplinary Action

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:44.769076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:6c120ae61b1081c189ff38316b4454643046f8d1f695ff1ca933146646dbeae3

Observation 023cb5db-ea80-4949-8960-5785c631f2de · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.895600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:d8c30f4ae7718c613f886c9ac73c306a0ed81939cc7250f9e3aaa749a4790234

Observation e5687d1e-9838-461b-900b-557f6a2fec31 · outbound

This paper cites an unresolved cited work.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-05-20T14:19:44.897581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f324ef783d1bc43681fc188f603b3e7a1bae13c05a88b60914f56dfbce241333

Observation b67597c1-38a9-4258-8a93-94026ae6741d · outbound

This paper cites 2025 , url=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , url=

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.899841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:df8e7343e094e0fb7f45d9f897e7012284d62d2e05c341201e328592cf328198

Observation 6c4ae0ef-39a5-44d9-8d0a-f92ff6b43301 · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.901601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:95f62978920fcff070dec0c4e741ca1d2b4aa3633e8a059693f14333da10b85f

Observation a3fc8fe2-b381-435b-934f-c5e10389cfb7 · outbound

This paper cites 2022 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2022 , eprint=

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.903852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:2dfe79bf5962b7acaacdb9857c79061ef53527b3aee1588a7926c635702acb0c

Observation 5ea2be2e-aa31-48dd-aadc-d09ee026e439 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.905837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:93d8dcf7c241013b2156d509a7f4f146dfb0e2ca42258787108637793894fbe2

Observation 58bd8690-7d8f-4611-a994-fa334720aa68 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.916693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:be43d49bdb03d560543467133fee2d32634f8168f709e455700b7fee661cb307

Observation 3cf9ad6b-1eff-4357-87b6-3ee948a39a27 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.920221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:f922e5c02340d6d8eb723ffc0b8e0bb836f6a34166d93f57b85196e572e88199

Observation 8206dfdc-6dc5-4fbd-8801-6a61e59b0a14 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.922089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:4a18395b429610e9a6038eab6ccd9d009e1304d2bd42abaeeaf11d0ff5e228d6

Observation b72d09b3-8246-4acf-a79e-ea6d3a9350af · outbound

This paper cites 2024 , month=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , month=

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.924069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:810afc5d8205401399428d943c8322ca90fa685c4c38ada017abb665df0c8525

Observation 48188bf6-2ad8-4be5-b97e-e3653d1e8b95 · outbound

This paper cites 2021 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2021 , eprint=

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.926105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:28bc6ca6f77cd9d8ee6aeb8b09a5cc4c38f837f18aa6468b0f1bac3f62a82cdd

Observation fc300d9d-a94b-4e6f-ae8f-1dda55e9ef17 · outbound

This paper cites Proceedings of the 36th International Conference on Neural Information Processing Systems , articleno =.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Proceedings of the 36th International Conference on Neural Information Processing Systems , articleno =

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.928585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:729297759a1016744ea9da62b35549395000b5940f046a9f35e4194bd64cddac

Observation 8134b1c4-b6a4-478d-9e30-ab8f2d27cc42 · outbound

This paper cites and Le, Quoc V.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Le, Quoc V

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.930843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:c0cfe9a548c30609b6dd1478188e61af80a9f8eb649f9373e3c2b08c9d389f3d

Observation d0be5f31-bced-47d5-897e-3b11e907b8b5 · outbound

This paper cites 2024 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2024 , eprint=

Reference 88

Resolution
parse uncertain
raw_fallback, observed 2026-05-20T14:19:44.932612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:74a9413ff023bba116b40e8a1414a99f116b6ef658a62cfcca4f1e56190c4091

Observation 6f5df163-a52c-49fd-9a37-8523eef62f0b · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.934415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:9256a776586edde92a076e3857758e8f27d8fea2fbb183951fb6f8ffdf47ad05

Observation 2fd4b8c1-f163-4a44-a2a6-f8db66c4edf9 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.936652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:88d34d1e9116e74a49ccc1a0b4a2eaf89a73eff40be1623fd7f7497f7e7c87bd

Observation b351967e-fd15-438a-8b08-e513e244ef66 · outbound

This paper cites 2023 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , eprint=

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.939246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:d5e7806d50b2f7ed3ca913b21d20260a80b86b2e2b483de131ebfb39941e8d34

Observation 30e1b8bc-505a-4061-944e-50b00b067e22 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.941389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:300e9b5696cae4c1abef6c6f27121894ab284f018822a6674e84d0348bb0515a

Observation ba526406-09d2-42a3-b9b3-2e4737cf43dd · outbound

This paper cites 2020 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2020 , eprint=

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.943184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:b7b49fa6e905f6ed0860152997f667992559de6d989292e9e2c4232460262852

Observation 7de6f035-0da8-46a8-ad46-8621a0989731 · outbound

This paper cites On reinforcement learning and distribution matching for fine-tuning language models with no catastrophic forgetting , year =.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety On reinforcement learning and distribution matching for fine-tuning language models with no catastrophic forgetting , year =

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.945543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:3ebddeccf1b73600d9510f968d1b38d0461e96135b68b1f2958a87d2d238b7c6

Observation b1921201-fc9f-463a-ae88-13e326e4900d · outbound

This paper cites and Kaiser.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety and Kaiser

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.947887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:05b5e22814e22cea680dbdc6ff2e3e4dd03865157f5b4e28aa6db592d315530b

Observation f70f8a8c-74de-4f37-a9ee-4c0c7f4693c8 · outbound

This paper cites The Twelfth International Conference on Learning Representations , year=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety The Twelfth International Conference on Learning Representations , year=

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.950471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:575af1a619c69b02b5eb3cb5feec6567930f986a8923de157d0ca556c628f120

Observation 9b1e8f52-0c73-4f13-a76a-665fe920ca87 · outbound

This paper cites 2025 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2025 , eprint=

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.952251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:1891a5382266552183e011174d449ea946fbfa782b75570a1cdb9cb5f9904967

Observation 8adbb5f1-9fd1-47e4-a853-b4040b903b28 · outbound

This paper cites Transactions on Machine Learning Research , issn=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety Transactions on Machine Learning Research , issn=

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.954585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:d7b0b3325cdb6525ca0cf06a6e5c793662598cd5691b81f53d117e3da6a57fa2

Observation 0a41f282-d53e-4a10-80d7-7e7cf3ec791b · outbound

This paper cites 2022 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2022 , eprint=

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.956959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:a7c93ad0f7c1d82f9acc98a9b51a1fd496b8096c878020a7ff5573442d9710d6

Observation 700e32f5-8d98-4544-bc4e-18fd5bb49cf9 · outbound

This paper cites 2023 , eprint=.

Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety 2023 , eprint=

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-05-20T14:19:44.959006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:19:44.695462Z digest=sha256:7f9c021698b583a2f8db38e9fbf85c5a10324113e2f9df6b8fe709a6d1e936f2

Pith citing papers

Observation 542fb94c-7c11-436d-acc8-29d26a4fea0b · inbound

Reliable Weak-to-Strong Monitoring of LLM Agents cites this paper.

Reliable Weak-to-Strong Monitoring of LLM Agents Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-05T15:53:51.319268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:53:51.319268Z digest=sha256:2b1de7304316e1b3fccd2f43dd9735b6c4c5bac6a6acf2e4eb884c16574702e8

Observation 7df2a5da-d825-428d-a547-9e820323b24d · inbound

Reasoning Language Model for Personalized Lung Cancer Screening cites this paper.

Reasoning Language Model for Personalized Lung Cancer Screening Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T00:06:41.352727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:06:41.352727Z digest=sha256:fe0c74362c8f2dc633c6cef853cf590ff68aab210a2a7d55fe56b1c096034da8

Observation 26c31bba-7c20-4d2b-b948-809a6de4f633 · inbound

\texttt{R$^\textbf{2}$AI}: Towards Resistant and Resilient AI in an Evolving World cites this paper.

\texttt{R$^\textbf{2}$AI}: Towards Resistant and Resilient AI in an Evolving World Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T23:10:46.316066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:10:46.316066Z digest=sha256:516ab031d836a9b93eb9e5176a40445a4403b646f8e73d4441e905d2ce98bf27

Observation 6270b95f-84e8-4a4b-bf6d-e438de84b155 · inbound

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data cites this paper.

Talking Trees: Reasoning-Assisted Induction of Decision Trees for Tabular Data Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-21T22:05:41.485914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T22:05:11.146329Z digest=sha256:3aa493ef7ae86e03108f1c14c8eac5e037da44f36bdd307cb53572c1b575e159

Observation 8345d0bc-ee95-4be4-ad50-38bd4b3d3d22 · inbound

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents cites this paper.

SafeSearch: Automated Red-Teaming of LLM-Based Search Agents Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:50.532615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T14:43:50.532615Z digest=sha256:f12cb75218e67376de16b2b5822f550d10eb35a7606be3baee5564551cf45529

Observation f9fe0967-cfc4-4a39-805b-4b68c40615c7 · inbound

RECON: Reasoning with Condensation for Efficient Retrieval-Augmented Generation cites this paper.

RECON: Reasoning with Condensation for Efficient Retrieval-Augmented Generation Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T10:20:20.487943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:20:20.487943Z digest=sha256:b8553d7a41d929d845ab2a0ed0b22fe7abde9d087374d39e43bee6114b929d1a

Observation 01c74dc6-7c76-4cf3-a8cd-350684d2cd0a · inbound

Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thought cites this paper.

Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thought Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T02:44:48.729794Z digest=sha256:5f6418f1be9388c1b3467aa008ea9499bca844e0da104110acf68e7bbbe8ffe8

Observation c5c07aa8-1ca4-470b-8bbd-dbedad43250f · inbound

Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thought cites this paper.

Can Aha Moments Be Fake? Towards Quantifying Decorative and True Thinking in Chain-of-Thought Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T07:43:10.083400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:43:10.083400Z digest=sha256:486c368d7b35cec9406d763181baa159302a49312b7317f178b775e07a163f18

Observation 9f32fefd-b237-4492-b6a0-3bed6e38926b · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:43:11.620446Z digest=sha256:4d0d49203bcd503e3dd4d6dbf21bfde28cfb6df9ba1f5cb0b97a6eac9efd96b4

Observation 63be8710-ddea-48c0-bb7a-fcb08eb1ec48 · inbound

Scaling Latent Reasoning via Looped Language Models cites this paper.

Scaling Latent Reasoning via Looped Language Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T07:31:49.234854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:31:49.234854Z digest=sha256:4a765e6bc7c69868f20fad76f11b24615112b9aefabdcda1aa9ddf960b579be4

Observation 25498b6d-158d-469e-9e1e-a938e42aead3 · inbound

OpenAI GPT-5 System Card cites this paper.

OpenAI GPT-5 System Card Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T21:06:36.680344Z digest=sha256:bcad858284391882fb22c0ad526ed1b898d4572f809b82d1f9a022d746c22bee

Observation 8e235981-f841-4242-9ad2-48d22f97dc1e · inbound

Legal Alignment for Safe and Ethical AI cites this paper.

Legal Alignment for Safe and Ethical AI Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 2027

Resolution
unresolved
no resolver link, observed 2026-08-03T12:10:01.515778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:10:01.515778Z digest=sha256:092bd0028f9194e0169bcfcf727b89f2d8bb94f7f707024ac5285553fd129642

Observation 89d4ddf9-bc13-4cc1-a565-b1934746b98a · inbound

Diagnosing Pathological Chain-of-Thought in Reasoning Models cites this paper.

Diagnosing Pathological Chain-of-Thought in Reasoning Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T23:26:30.279570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:26:30.279570Z digest=sha256:1272aca63a72490e979ad8c2e9b52f53ba84ef82f4b8c055dd376a63309c435f

Observation 024bf749-cdf2-417c-b718-5b33e39aa676 · inbound

NEST: Nascent Encoded Steganographic Thoughts cites this paper.

NEST: Nascent Encoded Steganographic Thoughts Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T23:21:33.388146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:21:33.388146Z digest=sha256:a02ad7f02a16d21df38b1d032182b89aad32de686b66567fe29bcc4b6b858df4

Observation 501d69b7-4613-4601-a539-90ba11857e13 · inbound

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring cites this paper.

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T19:01:08.076475Z digest=sha256:5087189c4f86c61688823cde5fc787884070c4cc28e0aa2b68f1c7c35c8c0409

Observation be144ec8-1e94-4630-aed8-4f7b8f324801 · inbound

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook cites this paper.

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 95

Resolution
unresolved
no resolver link, observed 2026-07-13T14:03:01.974171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:03:01.974171Z digest=sha256:98f4d6ff9d23ad2c3ba1b64de535a44218a04543fa95223e38b3030ed64dd25d

Observation eed55445-795f-48ad-b3d8-581e86d0b76f · inbound

Are Latent Reasoning Models Easily Interpretable? cites this paper.

Are Latent Reasoning Models Easily Interpretable? Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T19:45:06.229399Z digest=sha256:4b9e7d0ec4266a3a94771164b36dce6f0119b3f557f00ff5636ebcfd02e2c616

Observation c7263dc6-26d3-40dc-9e8d-ea22def1584d · inbound

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning cites this paper.

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T19:55:50.232871Z digest=sha256:082b305ed71d7795a1d8eaa7d532209b1ad0e587be1a02fd21041ec5e4a757b6

Observation 6a740188-0dcb-4af5-970b-47487c67f185 · inbound

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models cites this paper.

SafetyALFRED: Evaluating Safety-Conscious Planning of Multimodal Large Language Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-10T02:21:29.463149Z digest=sha256:5aa6944be4c1730411925163632f435211583bdfc2dec02ffffba6b4379f7864

Observation b8696292-6404-4fc7-afe1-1e8abb7361bd · inbound

Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought cites this paper.

Thinking Without Words: Efficient Latent Reasoning with Abstract Chain-of-Thought Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T11:52:00.969949Z digest=sha256:1c56319084bbc1983b177554c7d728fe02dc51e3630ec4999d9cae2981006bbf

Observation f95ebe2e-a79f-4368-ac1d-f5ae071ba116 · inbound

Architecture Determines Observability of Transformers cites this paper.

Architecture Determines Observability of Transformers Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T04:42:16.480831Z digest=sha256:c25cdfd23665e536814dc252a506f2142e5262084a7524e59389b45dd95deeda

Observation 8cede5f7-b7d7-4710-87e2-d88626198d4a · inbound

Architecture Determines Observability of Transformers cites this paper.

Architecture Determines Observability of Transformers Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T07:32:09.488225Z digest=sha256:a1b753cbb265b1aec89c321a1dd795431edb67e22d59332a8ecdf9c7be5c279d

Observation 63d30539-b1c1-402f-bfb6-29b542e6691c · inbound

Compared to What? Baselines and Metrics for Counterfactual Prompting cites this paper.

Compared to What? Baselines and Metrics for Counterfactual Prompting Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-09T19:02:46.991897Z digest=sha256:a438dcd5ba8eec838cc503042b07a60d8815792b4640c54ff0bc5500cba9bb0b

Observation d820622e-aa7b-4b3a-8e4f-9666deca004b · inbound

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure cites this paper.

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T19:09:30.284613Z digest=sha256:977a3f08cd45f07562b82b45b42edbc7d31bd648e4a93abce672ac6cc0c68ed6

Observation 2490a3cc-d123-489f-bec7-92cad2b0bfcf · inbound

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure cites this paper.

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T06:35:31.424885Z digest=sha256:c357ede4e8750ac315ea8968039acccf09b33e12a47ea8408489e902f6bedce5

Observation 00545680-df47-4709-8e81-3799ff98586f · inbound

Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering cites this paper.

Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T11:49:47.994456Z digest=sha256:d1d98d2f2e35761494b2d351713c87bd93d325dbd5d5cefcd74e37f445ad97e2

Observation 0134727d-24c8-424a-b670-97131fac9c86 · inbound

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight cites this paper.

LLM Wardens: Mitigating Adversarial Persuasion with Third-Party Conversational Oversight Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-12T00:51:17.434889Z digest=sha256:988664e7c9cf99d68e968b9c73427fa18f17d86d2e3df236010dd24a9b57231c

Observation e1c5ed0f-2d8c-46c5-9d09-eea15040da4b · inbound

When Reasoning Traces Become Performative: Step-Level Evidence that Chain-of-Thought Is an Imperfect Oversight Channel cites this paper.

When Reasoning Traces Become Performative: Step-Level Evidence that Chain-of-Thought Is an Imperfect Oversight Channel Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T06:30:12.558660Z digest=sha256:836373f91738fa66bea435ef40424bdb0ec9541c344e4e8fce0b3de012132b99

Observation bc9f9a25-82dc-4095-b9c5-322d5d6ad964 · inbound

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack cites this paper.

Do Androids Dream of Breaking the Game? Systematically Auditing AI Agent Benchmarks with BenchJack Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T20:31:50.043920Z digest=sha256:76b17256587dcdb7aedc215d34ce441521b7cff7322efdd6bd03560026d2dd5c

Observation e13d0afc-8695-4b05-8cdf-d0fb47c0359b · inbound

LiSA: Lifelong Safety Adaptation via Conservative Policy Induction cites this paper.

LiSA: Lifelong Safety Adaptation via Conservative Policy Induction Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T01:53:46.630161Z digest=sha256:fb69ae2579e5251fd5c78a572523b69a2785e062cd086d3c2d9566fdc8cb6fe4

Observation 272007fa-eb74-46fd-8cdd-a94a49ce4c62 · inbound

From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement cites this paper.

From Sycophantic Consensus to Pluralistic Repair: Why AI Alignment Must Surface Disagreement Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-30T20:35:03.004285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T20:26:13.783556Z digest=sha256:953893c54bdba02dcad702dfbb99abea3f5dc5967e8927a0f612f787573294dd

Observation c24e72bb-0c14-4a8e-a3c0-b7411ae3e6ab · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-05-20T20:13:43.316688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:a29c45650a56af37662688ec249401e6d5f043462482b9218de8c06b3934bb9d

Observation 7ad951b6-2b0b-4be3-b4c1-4fbc3a54de08 · inbound

Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems cites this paper.

Formal Methods Meet LLMs: Auditing, Monitoring, and Intervention for Compliance of Advanced AI Systems Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 60

Resolution
verified exact
local_arxiv, observed 2026-05-20T17:28:47.996777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T17:26:02.346222Z digest=sha256:579343c2081b16132c5ac530427882a7afa0252692b6851afd1d69c46c628110

Observation 313333b7-bf8a-4a27-ae83-577dc5a48e56 · inbound

PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media cites this paper.

PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T14:05:14.737146Z digest=sha256:5b0f9c86f965be0c2249020b3782d5e0f21261a93b340c23d9afb9dc9ec449a0

Observation 4541e96e-4c36-4095-a05e-9b7643342e95 · inbound

Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models cites this paper.

Stop When Reasoning Converges: Semantic-Preserving Early Exit for Reasoning Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T12:29:30.505761Z digest=sha256:949da8c86901bbe7e91e0c7e882dec61233aaf63520a2607b0945546d08b627a

Observation 42343de3-3760-45aa-8e91-63138567d29e · inbound

Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics cites this paper.

Monitoring the Internal Monologue: Probe Trajectories Reveal Reasoning Dynamics Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:54:20.285841Z digest=sha256:e1dc9c921522b14b9cc8d2d38293708a2b1a5e6e0aeb54060a9196ca0ea6365e

Observation 56c523a7-1443-4857-9bf8-53c9e8d34f88 · inbound

Counterfactual Likelihood Tests for Indirect Influence in Private Reasoning Channels cites this paper.

Counterfactual Likelihood Tests for Indirect Influence in Private Reasoning Channels Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:19:45.019656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T13:05:41.969600Z digest=sha256:4a63b3986ee319d11099f8b487354c0f670a37eac50e374359042184404813da

Observation 42a5fc82-e545-48fc-a3be-77c90db11f7b · inbound

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective cites this paper.

On the Cost and Benefit of Chain of Thought: A Learning-Theoretic Perspective Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:13:58.661889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T05:09:37.588841Z digest=sha256:143162bf9fc10141d53067a648ce4724d835a690c83852fe67106503ca9ca560

Observation 2fefb117-0c87-4a23-8681-b86e8b170323 · inbound

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models cites this paper.

The Readout Shortcut: Positional Number Copying Dominates Arithmetic CoT Readout in Small Language Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-25T06:30:25.605167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-25T06:27:42.024094Z digest=sha256:94eceda0f6f716d25489df1a3d02554332b6979789a2129b5e7f40f5130e231d

Observation 0c422f65-466f-4de3-ab00-a792993675ab · inbound

No Certificate, No Execution: Certified Traces as a Foundation for Trustworthy AI Agents cites this paper.

No Certificate, No Execution: Certified Traces as a Foundation for Trustworthy AI Agents Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:34:38.971438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T12:27:09.893183Z digest=sha256:f59e877a7bd89fba17d9190bb025c2420794887c0cc73b97edfcaed11d700726

Observation 0d47f1bb-2106-4b33-ba7f-b634204078eb · inbound

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth cites this paper.

Faithfulness Metrics Don't Measure Faithfulness: A Meta-Evaluation with Ground Truth Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-06-30T12:04:39.017442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:56:53.355299Z digest=sha256:4ee401c36a53faf01104f1cbe30eba9d0833c10f8d3b2fded3e3e53b3aa277cd

Observation 5f3c2b89-1ea8-4631-b873-17443aa63a70 · inbound

CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning cites this paper.

CORE: Contrastive Reflection Enables Rapid Improvements in Reasoning Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-06-29T11:53:23.838406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T11:48:18.855264Z digest=sha256:f4426a31b8b24c1ebe07f2e684cdee30c29f329266349960aa63f8098dae4c3a

Observation 9533fabc-59a2-494c-8636-d57885408967 · inbound

ReasonOps: Operator Segmentation for LLM Reasoning Traces cites this paper.

ReasonOps: Operator Segmentation for LLM Reasoning Traces Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:13:15.914008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:04:44.270536Z digest=sha256:82189bd25fe0ed1d80aa2af2ad934ccf4c1aa1c9f2b0483cded81c704d14ad7f

Observation a8d0b19b-5204-441b-ba74-3a9b740f2dfb · inbound

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models cites this paper.

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:07:13.017238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T22:10:01.701850Z digest=sha256:210ac7247b283c63d7553b4f553cfc6ddd57e0406eb6c073ab0ede2468e8eb8d

Observation dc8a9399-7603-4b36-b853-7ac4a847fbcf · inbound

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models cites this paper.

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T05:53:08.786506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T05:50:40.480874Z digest=sha256:f3af50bcc4e59d3c9711f5f75c7ad517ef5088773794066eb88d4d65fe5c2b7e

Observation d9b1ef5e-4f70-4b76-8495-a0d0b40be888 · inbound

Sycophancy Towards Researchers Drives Performative Misalignment cites this paper.

Sycophancy Towards Researchers Drives Performative Misalignment Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 62

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T22:27:26.211572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T18:52:29.375027Z digest=sha256:65eeaf4dc7abbb1ad466f9de3b4cc02e0a1d32bbfd6180678c21b3ed58acd7b6

Observation 48ec688a-a5cc-42f0-a504-05e311b68ee3 · inbound

A Note on the Strategic Confinement Problem cites this paper.

A Note on the Strategic Confinement Problem Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-27T17:31:06.808576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T17:30:12.165074Z digest=sha256:6833eed4a3ba647a6697238fcb3ae8c4b811f8815b70885fa0c58b7aef4f2483

Observation 1c22a08a-e98b-46fa-90c8-abb472bd2ad5 · inbound

The Distributed Detectability Band Against Marginal-Preserving Attacks cites this paper.

The Distributed Detectability Band Against Marginal-Preserving Attacks Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-03T05:57:41.866414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T12:58:22.056355Z digest=sha256:70d28bf7440a51ec443245bfa499b4073cf6d947c7f8ed80cb86a332ec5be1b6

Observation 38c0b967-0b4c-4ea7-bdb0-ab921b64b788 · inbound

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment cites this paper.

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-06-27T13:30:56.614999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:26:14.457195Z digest=sha256:f082cb0355cff050d7510df22905439ad55eeef03f34387b2f6d3fea98f47be4

Observation 3eecd365-6f83-4a9b-9d61-2854cab0eb6a · inbound

Observable Patterns Are Not Explanations: A Causal-Geometric Analysis of Latent Reasoning Models cites this paper.

Observable Patterns Are Not Explanations: A Causal-Geometric Analysis of Latent Reasoning Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 46

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T11:18:03.821668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T09:36:05.700067Z digest=sha256:8fe71e6253dda4450cad0c3cc41a3d614f25a0dd7ad7f27ad385604aecbedce7

Observation 56f6f921-80cb-4205-9a06-c20b67647b62 · inbound

Decoding Hidden Deception in Reasoning LLMs: Activation Explainers for Deception Auditing cites this paper.

Decoding Hidden Deception in Reasoning LLMs: Activation Explainers for Deception Auditing Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-03T20:18:56.656841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T01:28:38.810889Z digest=sha256:1bde5763d8b97c15eef0b4abf3d08e78aa415b68c049b1affb92425c7655c706

Observation c822910c-de5b-4e92-80a4-93e3150ad721 · inbound

How Transparent is DiffusionGemma? cites this paper.

How Transparent is DiffusionGemma? Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T03:09:29.706294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T18:23:41.734117Z digest=sha256:b9ad4a310fa204d5cd66858ba7726658413a76d6c0a3041f02f215e03aa55c37

Observation 5bc94bd1-c13a-4a51-9dda-e6243a7d5e45 · inbound

Reinforcement Learning Towards Broadly and Persistently Beneficial Models cites this paper.

Reinforcement Learning Towards Broadly and Persistently Beneficial Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T11:39:46.771704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T07:51:13.283619Z digest=sha256:06af08406de92bf1867e386a22c7e13797aa0fccce9f6685ecd65b43a8086d2e

Observation 419cd2fb-efd0-444c-8e98-3f32e0810863 · inbound

Radical AI Interpretability cites this paper.

Radical AI Interpretability Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T13:09:51.070994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T05:26:47.854890Z digest=sha256:f0a30e4fa15ddb2d87c2b57c0f85242acd20cc28483a3fd6d0bbf94f4aadd2b3

Observation 2165b35c-e23b-4005-acf9-c3f861bbce40 · inbound

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems cites this paper.

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-01T16:15:50.102892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-30T00:42:20.220668Z digest=sha256:92dffae6a64fe3c8cb52d97d51e0b70f4e9d1c6a6db117b148da543f6758484b

Observation 5027a41d-c9c5-4cd7-a106-81d19138bc75 · inbound

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression cites this paper.

Why Struggle with Continuous Latents? Interpretable Discrete Latent Reasoning via Rendered Compression Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:44:18.726440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T06:43:20.893142Z digest=sha256:cf96fc6dcbb180df31688345a94435320df8390e0645d8cc5428866095e33ea9

Observation 58f7a6e6-5852-4328-9c08-2e6ea1b24c7d · inbound

Introspective Coupling: Self-Explanation Training Tracks Behavioral Change Despite Fixed Supervision cites this paper.

Introspective Coupling: Self-Explanation Training Tracks Behavioral Change Despite Fixed Supervision Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-01T10:35:42.791677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T05:18:30.190239Z digest=sha256:ed4dfdfb86605403ee986bf6ce25e79f3302606aacf7d8bb7dbb1f465fd8a038

Observation 0b1f7055-17d5-4225-b04d-ed5ec15c0ebc · inbound

Conversable Complexity: Agentic LLM Collectives as Interpretable Substrates cites this paper.

Conversable Complexity: Agentic LLM Collectives as Interpretable Substrates Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:46:56.198325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-02T12:46:40.700728Z digest=sha256:b23782ba7275a162bbf06cdb04c21530331f2c6af029d785b198bb88a7c7607d

Observation e2e6535b-bda1-4760-91a8-2638e341c5b3 · inbound

Geometric Signatures of Reasoning: A Spectral Perspective on Task Hardness cites this paper.

Geometric Signatures of Reasoning: A Spectral Perspective on Task Hardness Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 40

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T17:38:43.202586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-03T17:36:09.322804Z digest=sha256:4befc6a4dfec307aff500ea206762c8f54490b9490f56eda9a68e69896c4da3c

Observation d86fa7b3-9659-458f-b319-16261d1d91b1 · inbound

Online Safety Monitoring for LLMs cites this paper.

Online Safety Monitoring for LLMs Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-03T12:58:07.490637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T12:53:49.681641Z digest=sha256:bc6e105ffa81db939b1a03ae2624347a18dbc341820b1b5f1109b7484f131700

Observation a45c86ca-3278-4a1a-8fff-70ec459ee952 · inbound

Distributed Attacks in Persistent-State AI Control cites this paper.

Distributed Attacks in Persistent-State AI Control Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-03T11:48:04.376898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T11:47:05.585463Z digest=sha256:7b9bc5f2b0e2b60f6fe59713f43a1410a9c5157aa04949db02ea54e6149ad511

Observation 0985f9cf-8ed0-4e46-9831-2174196571ee · inbound

Distributed Attacks in Persistent-State AI Control cites this paper.

Distributed Attacks in Persistent-State AI Control Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T07:59:05.507834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T07:59:05.507834Z digest=sha256:6972debc06807b87a11fc314c0d45022bd021bfafbbd5109ff055e3ef174cf6b

Observation 07902bed-b5ca-49f9-96d4-8d36a6f4b10b · inbound

Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety cites this paper.

Position: Preventing AI-Generated CSAM Necessitates New Approaches to AI Safety Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 97

Resolution
unresolved
no resolver link, observed 2026-07-12T14:28:50.627444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T14:28:50.627444Z digest=sha256:dc511e4c793cb14c225ada7598e2818ccaa74ab3717f9e79db602648c5402c68

Observation 3a8df77f-a393-4a11-a01a-5f512aa34551 · inbound

Predicting LLM Safety Before Release by Simulating Deployment cites this paper.

Predicting LLM Safety Before Release by Simulating Deployment Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-09T18:16:25.820530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-09T18:07:42.556329Z digest=sha256:faabd926a5e728cbb7fd499590e1851d2adc6dbed4866a84498814248ecca03b

Observation 454b0589-1219-414e-8397-1debf5570977 · inbound

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring cites this paper.

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 92

Resolution
verified exact
local_arxiv, observed 2026-07-10T00:56:40.987209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-10T00:52:47.537142Z digest=sha256:8e1778df8ce87535b9c8bd4407882eb98a4ed12d0bb583bcc00a3a204f9d3ca9

Observation 2ac22c00-237c-4925-a7b5-e6cf1fcbf8d5 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-14T15:45:54.532529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T15:45:54.532529Z digest=sha256:de3f66300cb30cd144e571963ae6b2d9a66733eeff511d77fa81eaba57fffbdf

Observation 097f8438-1a9e-4ba5-8146-b1c1f020dd13 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-02T08:06:10.804751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T08:06:10.804751Z digest=sha256:5b38c94619d6fa1e2a7cb311382c404452408c2e1635cb5ed965f470b5993b32

Observation 4cd93958-929e-4098-bddf-200e0f8b80c0 · inbound

Length Penalties Make Chain-of-Thought Less Monitorable cites this paper.

Length Penalties Make Chain-of-Thought Less Monitorable Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T04:30:28.364330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:30:28.364330Z digest=sha256:5035a2df7f970a17e42442cd81306a22e7bdf9bb6b2abfdb15ccd4ca80158a8a

Observation c956e04d-15f3-4430-b18a-18ab741f967f · inbound

GDM AI Control Roadmap cites this paper.

GDM AI Control Roadmap Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:53:52.430294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:53:52.430294Z digest=sha256:dfa2a10e95f630f3db1c568b58e88af391db674bde9cd24d7b3b8fe2522aaa26

Observation 7f733765-f9be-43b0-bd9c-a2579c0b9889 · inbound

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary cites this paper.

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T15:08:37.570947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T15:08:37.570947Z digest=sha256:86d2a1e76f643a4c36a38ade06284ce1e731a52241b6eb3ab3392ace62fa9d86

Observation 4dc50c46-c9be-4e21-bded-b442bc85ff3c · inbound

Not All LLM Reasoning is Visible in the Chain-of-Thought cites this paper.

Not All LLM Reasoning is Visible in the Chain-of-Thought Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T04:14:44.668324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T04:14:44.668324Z digest=sha256:e7eb581e9c7d4e41caaa589c3a027f90f0da2140dd8d3ecf392edf1272703359

Observation f5b88481-bcb1-49ac-956c-4e793b1c6f0b · inbound

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness cites this paper.

On the Generalization of Steering Vectors for Chain-of-Thought Faithfulness Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T14:26:28.259868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T14:26:28.259868Z digest=sha256:4a14bc5b4c63f51e7bdd969ab66b8279fd789ad655eaf171bdb59e0e86a21611

Observation 4ee04524-1bec-470f-9a20-77c41943bf1c · inbound

A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense cites this paper.

A False Average: Chain-of-Thought Monitors Collapse Where They Are the Only Defense Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T00:40:38.474169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T00:40:38.474169Z digest=sha256:d32c9bd0315da9c6b863e3e8f6933eb062436cf4ebcf26b08617bb3042ea3608

Observation a373dd73-b609-4f09-97fd-80d1fbab6cd3 · inbound

How Much Does a Reasoning Summary Reveal? An Observability Ladder for Large Language Models cites this paper.

How Much Does a Reasoning Summary Reveal? An Observability Ladder for Large Language Models Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T15:22:54.216027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:22:54.216027Z digest=sha256:019e97ff2e176b9ba1fc962f9ca5efc918ab714ca8f57782753852bcd2c7525b

Observation b6e315fa-63f3-40de-96aa-13372c40000b · inbound

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics cites this paper.

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:36:45.164802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:36:45.164802Z digest=sha256:e8f118c754bc155731cb58144049a10991525d50daaf5dca52da84a121b6f136

Observation b7df419f-7609-427d-9ec9-fb2750533583 · inbound

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics cites this paper.

The Tell-Tale Trace: Detecting Reasoning Failures in LLMs Using Chain-of-Thought Dynamics Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T21:36:45.527818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:36:45.527818Z digest=sha256:e405b557f87c226e93e490d80273117ed3d6058b1e9833d04d7a402636af830d

Observation 9eddd5c5-71cd-4642-8852-daecc35eba2f · inbound

Risky Business: Measuring The Faithfulness-Safety Tension cites this paper.

Risky Business: Measuring The Faithfulness-Safety Tension Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:40:55.680701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:40:55.680701Z digest=sha256:ec66e2fd7befb4c47f51d416e29d68a8114a015d61198d828e47117ddbba3ec9

Observation 3e760a5c-06ea-4874-9859-3c1560cfed5b · inbound

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings cites this paper.

Chain-of-Thought Monitoring Can Be Unreliable in Implicit-Influence Settings Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:52.127591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:51:52.127591Z digest=sha256:f84b3c61af7d181f0b98721feae6550d714b341008fff548db2014b31390664d