Pith. sign in

Paper Citation Record · LEDGER

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage?

As of 11 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 1 inbound Pith citation observation for arXiv:2606.05647.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.05647 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T01:48:56.367899Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:50:00.612066Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch16

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c5cf12ac-2454-4d40-bc2c-7f7c10b1fa0f · outbound

This paper cites Information and Software Technology , pages=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Information and Software Technology , pages=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:a2a35c89733e6a95ab358f7e1736278a35cd9557d2e41f4b4edf9fefeeeefa3c

Observation a75ae00e-f95c-4d12-bc1d-c850dd9a1104 · outbound

This paper cites and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , booktitle =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? and Yang, John and Wettig, Alexander and Yao, Shunyu and Pei, Kexin and Press, Ofir and Narasimhan, Karthik , booktitle =

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:7d9b70014374b0e3135d3e59fd469e46e0b34468af340741cb44bd75d5b3fec1

Observation 9cab7743-e59f-4e8a-ae1e-ee35058de401 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Advances in Neural Information Processing Systems , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:8f5c55323e5ee9e4c1a52ac5e9856564d5d84c6dd8c6953bc924cf5cb58710b2

Observation 99045116-2597-43ba-aded-dc792c07ca4e · outbound

This paper cites ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models , booktitle =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? ResearchAgent: Iterative Research Idea Generation over Scientific Literature with Large Language Models , booktitle =

Reference 4

Resolution
verified exact
doi, observed 2026-06-28T01:51:28.890593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:5ee488656ffd25288811b52abd6baba9496af95d68ca1ae272e16eb49da20090

Observation 36eb4624-5023-45bd-a2f4-158dd4183ff8 · outbound

This paper cites Retrieved from https://arxiv.org/abs/2512.14012.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Retrieved from https://arxiv.org/abs/2512.14012

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:46:57.264600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:3b171bbf3b078b9fa30ad2ccbbc81ae59ec10222087766c56aa71d66b19a5920

Observation 71e53e0f-4dbb-4f5e-b36f-7a8e38b73c40 · outbound

This paper cites Horikawa, H.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Horikawa, H

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.281030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:40c680ce824bfd80a75b484c830bd389f389c4b8be5ac3d271c61f8af33123c8

Observation 9b58cfe1-7e13-4a3b-a5d3-7ac32affe3c3 · outbound

This paper cites Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.283555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:2329a5fd0801a41a4477d69fd1a4154452f365df593f8d5bc82af59fe4e9f3ae

Observation 3849e3d0-15e6-4617-8d9b-78cc4f217974 · outbound

This paper cites Agents of Chaos.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Agents of Chaos

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.288595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:705a7b65a3a544668b4ec231fe2a16cebbf0922c4ad42bada1bd529ac71a6956

Observation 04ee2710-624d-4d15-a8ed-22a4121c84ff · outbound

This paper cites ICLR 2025 Workshop on Building Trust in Language Models and Applications , year=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? ICLR 2025 Workshop on Building Trust in Language Models and Applications , year=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:99e4c137a653e0c6266493f82a54e0179aee21ff237be330419840a447a75743

Observation 724378d9-82a6-4909-9019-3b82872b7f95 · outbound

This paper cites Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems , pages=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems , pages=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:4acc8aa74c7c368a2609cefdbdeddc752fc6f303858aa56f67418454db8aea2c

Observation f27775fe-bfa1-46bc-b86e-09ebca92eac6 · outbound

This paper cites How Does Information Access Affect LLM Monitors’ Ability to Detect Sabotage?.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? How Does Information Access Affect LLM Monitors’ Ability to Detect Sabotage?

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.278635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:d97f4aff7a8266f07ead745bcd65923b2ceebaa91bb2691aca67f443b5018148

Observation 548eac33-5415-4f56-bcf3-012132e23035 · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:46:57.277907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:0081bbcd749667d29c1c18a7db9c22a5a6a6bc5a956ccd19216667f0848f1427

Observation 2e3f7154-9cf2-4915-a185-ecdd0f8137fc · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:76c110355d946ccb4166c742f50f36e1afbf77ae6d0b000d1a3cd5e3ea148fd4

Observation e2308a4c-3a33-43ff-b475-5079dedeaeb2 · outbound

This paper cites A Survey on Trustworthy.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? A Survey on Trustworthy

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:9bca934a43850c88443915a7def581bd3e6ef01c8f5e8470cd85ccea7f032265

Observation a89f16ad-0224-4d49-8727-66d5f4a3f87e · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:9b8d872100ac8d9704e1397acd5bd5cc9e31312ed1bcd2286fd0f2c969bbfb25

Observation 5212ba0f-3037-41c2-86cc-a141e13a1a44 · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:5a4fcf436aeede8e4e387f65e43698236505d71ea154fc51bfc973fc51c019ed

Observation 88c25f8e-ebc2-4222-bd8b-e495a781d2ea · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:04b22c36a3172187e6f572d0567d6cce69d9fd42db398006f93ff3b5fad3366a

Observation 158293f9-b2be-4a73-a15b-3f0466147936 · outbound

This paper cites 2025 , journal =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? 2025 , journal =

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.286024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:97f44e1e334b4284f84a4325a5124759f790e43c9febdc98f28c3cc2900819fb

Observation c5a38714-17ac-4b11-80ed-2ca7e4a0e4cb · outbound

This paper cites Async Control: Stress-Testing Asynchronous Control Measures for.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Async Control: Stress-Testing Asynchronous Control Measures for

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:68c6a9064d1911dfd0f4a813fa37aa35990317c237f70edd0e158647e3eebe24

Observation 7da168a9-a7f1-4ba9-97e3-4cedd4cbf827 · outbound

This paper cites Reliable Weak-to-Strong Monitoring of LLM Agents.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Reliable Weak-to-Strong Monitoring of LLM Agents

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:46:57.291196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:bf0c0e119331f07bed1134ddd9cd204ddf27fa0135c09a5846bb6486432e069a

Observation aec8c6f0-c3d2-4711-994f-4b029290bde2 · outbound

This paper cites 2025 , journal =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? 2025 , journal =

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.264776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:6ff6a9da4dff676150b18d145939adaab901a06d2c8378d620e99e6f99e1c338

Observation b61a4d78-074f-48f1-8ec3-94d54cc51402 · outbound

This paper cites Alignment faking in large language models.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Alignment faking in large language models

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.262186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:9a257971e91a4d9893fe8f2702ac4531f5df8058c39a34c1b3e61eee7728e2e5

Observation 6e3e06bf-5acd-4d39-b3e0-d68820c2c69b · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:e606c8d47ef2a63fac4cd4bb8190ed22a5cbd697e4578e8313b4ebd17df40056

Observation 0715d350-1acc-43b0-a23c-ce60b0cebc9a · outbound

This paper cites Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Agents

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.267236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:7d05fad8687b667b51fa47c2eed67f6a88b477a65d533ef7dd72804f1de5e68d

Observation 0fcbb8f8-b50b-4ca2-8579-d9b1692feb33 · outbound

This paper cites Assessing.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Assessing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:22b506c1a5255d644b337b3db92a933721e027ffba2c2d8e90362636c7e4e8b8

Observation f9e179ad-05c4-47b9-b70a-252c895875c4 · outbound

This paper cites Ritchie, Soren Mindermann, Evan Hubinger, Ethan Perez, and Kevin Troy.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Ritchie, Soren Mindermann, Evan Hubinger, Ethan Perez, and Kevin Troy

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.269807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:a0558b4d86b68af7af33e98913222c5ec58bed098d9cc6c923b90506e9262353

Observation 4371d167-04c8-4802-bea1-f30afd6ec100 · outbound

This paper cites Natural Emergent Misalignment from Reward Hacking in Production.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Natural Emergent Misalignment from Reward Hacking in Production

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:547adbaa3dc13c06b32387e67fff65ffb6e1f196cf04c0b5855465ecb617c560

Observation 06f27233-de57-462a-bd99-04576628c8c3 · outbound

This paper cites Stress Testing Deliberative Alignment for Anti-Scheming Training , url =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Stress Testing Deliberative Alignment for Anti-Scheming Training , url =

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.259807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:ebf5ccc480b587de0e26b23ad76e8a07deda7579f471b21ee67d5b5576ec7389

Observation 73d7b82c-4499-4dc7-8ff0-f4c9ac8f398a · outbound

This paper cites , journal =.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? , journal =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:1b55ad4c4e916c3d16308b4e8a372d478d621fef2d48abde10038fb24c2c23e2

Observation ce3459a6-c672-4bbe-a28f-7c0f62a61807 · outbound

This paper cites Cot red-handed: Stress testing chain- of-thought monitoring.ArXiv, abs/2505.23575.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Cot red-handed: Stress testing chain- of-thought monitoring.ArXiv, abs/2505.23575

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.253840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:7200b9afd779790f738a58a7b0a5c3a7bfa82f53062a76c282ae9b9d59ef5d09

Observation 5a8ca333-d64e-4da0-8a1b-3c54ac3a2859 · outbound

This paper cites When Developer Aid Becomes Security Debt: A Systematic Analysis of Insecure Behaviors in.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? When Developer Aid Becomes Security Debt: A Systematic Analysis of Insecure Behaviors in

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:0e9f3108f291b63fd6fda700479cb2000b1b79ed9785c7b8b7ad45e66ebbfce7

Observation f9adf27e-512e-4567-acbb-1d32dd378503 · outbound

This paper cites Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:46:57.248117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:88b6e2d959f25139fd827c3640b74b3e9723561a5cac35848c46080895d64775

Observation 562f5cc0-a16f-41c5-b069-98ae9fa55e2c · outbound

This paper cites Is Vibe Coding Safe?.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Is Vibe Coding Safe?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:5631cc82a25cf20d264d41c95fefc28ab94caa3c0817298247acd567d7177934

Observation d336efa8-bfa0-4201-95aa-e1fe95390763 · outbound

This paper cites Maloyan and D.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Maloyan and D

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.293745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:b13ff47a508cca19c998dc9acb0867ae5b6e99c3363fd60fb0ee07394751b7e2

Observation e7d06ee2-51e1-4aed-b1ab-d7a163826097 · outbound

This paper cites Evolution of Programmers' Trust in Generative.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Evolution of Programmers' Trust in Generative

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:8221b05d982a261070457d495b871934669c39b9fd49397ed6e6f08b147c529b

Observation 73db284a-fb6e-4af1-9535-ac8873cb59d4 · outbound

This paper cites Privacy Leakage Overshadowed by Views of.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Privacy Leakage Overshadowed by Views of

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:8332a728a9882d3af534f905b2e390ce0409626da184f87c212d6664edbe2577

Observation 49fa82c2-df53-4912-87c2-0ec106798590 · outbound

This paper cites Not What You've Signed Up For: Compromising Real-World.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Not What You've Signed Up For: Compromising Real-World

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:d8378b5c1c9f99e064343804fc04d0cac4bbe4e4dce3e34b93c92a0f8fa4840d

Observation 7a6e55f6-a435-42f7-9429-14e3cbb5b190 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Frontier Models are Capable of In-context Scheming

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.251299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:43bf8e499489441995a8c744375a1edbf146db752e4f873d174c27a98e08d48a

Observation fb42a312-2899-4dcb-b526-a34d427c02b4 · outbound

This paper cites Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T12:46:57.256904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:9637a3644c9c9f6fa1a86b29f7c06cf5c6669b47f211c2c22e3b33bc8f0178be

Observation c970b384-e342-48b3-9543-cc33e94b987e · outbound

This paper cites Asleep at the Keyboard?.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Asleep at the Keyboard?

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:07092b6878230677526e7fde3ed54c9d3f3c56623bb864334a899228e7a3ec20

Observation 3bf517bd-ef18-433c-bd76-7d65558826ff · outbound

This paper cites Do Users Write More Insecure Code with.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Do Users Write More Insecure Code with

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:af208d7b19a5020f833eed4c7ec107c8cd8509b0f9e26c4c6a745a6ebc149ac4

Observation 718e5478-0239-4c92-a815-0fedbc601566 · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:a369ce1896ace20da072a53ed9767d387b4b443887a4c98ac35f02c447f6d150

Observation b61d1e8d-b559-4a7e-ad9c-761dd17e43f1 · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:a96f0fa24974c5925479efe940473d379d5bbee9cbe15537f3fff9c37a9a8794

Observation 0e5fd8cd-7c50-489c-8cc6-e171ff683a6e · outbound

This paper cites Red-Teaming.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Red-Teaming

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:867448763f8567b14cc0bd9f7f013afcd49f00ce5efe97a3ba36744b19da8936

Observation 87090d60-508a-47b2-9f36-0551a4e07d2d · outbound

This paper cites an unresolved cited work.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:bc98156b53fa786db1e718f26aa839bd909aaa8f5acdcd6f930cb7378612a1f0

Observation e11dc3fa-c38d-40ae-a0d0-364508b3e5ec · outbound

This paper cites SafeArena: Evaluating the Safety of Autonomous Web Agents.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? SafeArena: Evaluating the Safety of Autonomous Web Agents

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.272709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:84678d5fe8e6fe06b3af5aa9fce937a456fc8113ed99c3450fbdbb8101d1677e

Observation 23b73752-0f46-4c3a-93ef-24e996d857f1 · outbound

This paper cites Usability evaluation in industry , volume=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Usability evaluation in industry , volume=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:6354e8e83a9a7b77effc6e939af4bbe00c1350802633e2e39837ea024c907994

Observation 28eb45e3-cc5b-416f-afe9-da1d99c19e17 · outbound

This paper cites Sabotage Evaluations for Frontier Models.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Sabotage Evaluations for Frontier Models

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:46:57.283112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:9dc3d2d429a7cf8dda2ca009e56ae91d3f8873160107d49ee49ae9b40227276d

Observation 92df3bc5-d35e-492c-a519-d8c4b080cffe · outbound

This paper cites Authorea Preprints , year=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? Authorea Preprints , year=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:038a6198d947984dad864773f8428b96828609811273bf1bfdcf6598b55a30ec

Observation bb7f1493-42fe-4fb4-8f06-7c2896a09bd2 · outbound

This paper cites IEEE Access , volume=.

Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage? IEEE Access , volume=

Reference 50

Resolution
unresolved
no resolver link, observed 2026-06-28T01:48:56.367899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-28T01:48:56.367899Z digest=sha256:40c4658a91d712e90b96563cf44e2581eecdab8ea6122b93e687fefbb0b68a8c

Pith citing papers

Observation d684a475-b17e-4e4b-977a-6aac26fb3534 · inbound

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D cites this paper.

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D Coding with "Enemy": Can Human Developers Detect AI Agent Sabotage?

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-01T12:50:00.612066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:50:00.612066Z digest=sha256:934ac26eeb9dc7663067411cd99048c78b4fa5377ebc01a55e8b1d279faf3dda