Pith. sign in

Paper Citation Record · LEDGER

Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2505.04806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.04806 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:25:42.450138Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04941d1a-7427-48ec-a832-4c17fd59af40 · inbound

Exposing Hidden Backdoors in NFT Smart Contracts: A Static Security Analysis of Rug Pull Patterns cites this paper.

Exposing Hidden Backdoors in NFT Smart Contracts: A Static Security Analysis of Rug Pull Patterns Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T05:25:42.450138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:25:42.450138Z digest=sha256:98739c61ac31b09b76042c8847304ec696880817836092c8b9ad5ade0aac18e7

Observation b4282a96-6bb9-4961-b94c-0cb55c1f85d4 · inbound

A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents cites this paper.

A Survey on Autonomy-Induced Security Risks in Large Model-Based Agents Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:45.015831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:45.015831Z digest=sha256:99338bc1b0d1c5d3087ca65d8ca8af52db4fa9af1fdedbdc1727b441110e5c26

Observation f06f63dd-10ab-4e04-a6d4-df2ba5aebb24 · inbound

Mass-Scale Analysis of In-the-Wild Conversations Reveals Complexity Bounds on LLM Jailbreaking cites this paper.

Mass-Scale Analysis of In-the-Wild Conversations Reveals Complexity Bounds on LLM Jailbreaking Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:05.267094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:56:05.267094Z digest=sha256:bbd7b6f6311a8f982e899b6bd82e2850eaf231ef765a66693a2b632825a6decd

Observation f9984e62-8605-4ef4-8134-5a73f9b6d88a · inbound

ANNIE: Be Careful of Your Robots cites this paper.

ANNIE: Be Careful of Your Robots Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T11:00:38.221707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:00:38.221707Z digest=sha256:bc967adf8087f7b1b6dbf61588e7018d06503594191308e25a0147e0edab8b9d

Observation 4915318a-e3b8-4ea4-8341-f944fe4a7c5b · inbound

Beyond Context: Large Language Models' Failure to Grasp Users' Intent cites this paper.

Beyond Context: Large Language Models' Failure to Grasp Users' Intent Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:11:13.575849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T20:09:25.827452Z digest=sha256:bc3ea7128c95367bce81f19bd81f8aa5c40da0643944b9487933c05ae9fb4c56

Observation 267ac4c7-1726-48ad-87cc-fe52b623d097 · inbound

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection cites this paper.

Whispers of Wealth: Red-Teaming Google's Agent Payments Protocol via Prompt Injection Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T15:20:17.397162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T15:18:58.274360Z digest=sha256:84ce2d4fc1f95a98876a6b41dca34e95bbfbae1e8ccc5239d6da8591da1bb8c7

Observation acb0caf6-b910-471d-bc89-14b628f3f354 · inbound

Evaluation and Hardening of LLM System Instructions Against Extraction via Encoding Attacks cites this paper.

Evaluation and Hardening of LLM System Instructions Against Extraction via Encoding Attacks Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-13T14:38:58.879673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T14:38:58.879673Z digest=sha256:c4dc1fd5922116c144e3a29760247a1de3817e1df8351954717dcb4caefb803e

Observation 8658f980-8186-4615-b11b-930412589c82 · inbound

An AI Agent Execution Environment to Safeguard User Data cites this paper.

An AI Agent Execution Environment to Safeguard User Data Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:11:05.891404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T02:14:40.639143Z digest=sha256:9a66b61a7c8140a46d9ed025634d69fd6ae35d77aebf4a75ad863bb3cdffa177

Observation d42c1592-e1ca-42eb-8b34-26f2a22ef125 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:46:28.112736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T01:57:22.554881Z digest=sha256:a28ee0477725d588a85c82cea274b465eda4789d39d7e6b40567d3e1c8803253

Observation f3b68ddf-f963-4730-90a3-b625018c5aa8 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:45:45.437788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T22:52:54.684185Z digest=sha256:16498304ffec31f5442877d846dded70aba1dc2e8fcbbbd51f950cefe3525c6f

Observation ec7cb2de-065f-4dac-865f-71bd54ed6d78 · inbound

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents cites this paper.

OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T05:17:31.221017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:17:31.221017Z digest=sha256:513f27dd40ea6de3726f3977715859c02708fc4784d74d46daa2a1239a01bde2

Observation fde27398-675a-4de3-a358-7073ae7f9a9e · inbound

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection cites this paper.

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:05:20.411461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-25T04:05:15.708438Z digest=sha256:606ab291d90fe232f09042b2d5decbd7e33512e6df89f7213f8343f611616b48

Observation c6f664a5-4dfe-4664-a750-3302e263781a · inbound

Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals cites this paper.

Prompt Injection Detection is Regime-Dependent: A Deployment-Aware Evaluation with Interpretable Structural Signals Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:50.228134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T18:41:03.566306Z digest=sha256:d2ec381e4626c2d82fc7109571d59715e95b9b7d832546ad4e790fc06bcd8b67

Observation a81e95ba-693a-486a-ac6d-7d77d24d1de0 · inbound

How Reliable Are AI Attackers Against a Fixed Vulnerable Target? A 400-Run Empirical Study of LLM Penetration Testing Consistency cites this paper.

How Reliable Are AI Attackers Against a Fixed Vulnerable Target? A 400-Run Empirical Study of LLM Penetration Testing Consistency Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T14:13:30.554033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T06:53:48.841077Z digest=sha256:5ae686d92d4e10b57a4f7f2bcd97580b4dbda9a21ce9e060c2fadddfde848e11

Observation d21b7dbd-c8a7-487e-ae3a-4a75b3def877 · inbound

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP cites this paper.

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:19:54.253542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T04:07:14.506108Z digest=sha256:57ac34e96107413abaee12bfd42459c50cb48bf957bc158f5a0c36346ad94089

Observation c137ff90-863f-4cee-a111-6f3369f719e1 · inbound

kNNGuard: Turning LLM Hidden Activations into a Training-Free Configurable Guardrail cites this paper.

kNNGuard: Turning LLM Hidden Activations into a Training-Free Configurable Guardrail Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:08:42.782288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-03T17:03:06.656045Z digest=sha256:13fe4526223e9f1595bbbc5a595df98110405017f26ec08acd11b45a54afc622

Observation 73024db3-e914-4cc7-90d6-c919535614fb · inbound

kNNGuard: Turning LLM Hidden Activations into a Training-Free Configurable Guardrail cites this paper.

kNNGuard: Turning LLM Hidden Activations into a Training-Free Configurable Guardrail Red Teaming the Mind of the Machine: A Systematic Evaluation of Prompt Injection and Jailbreak Vulnerabilities in LLMs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T08:27:42.250470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T08:27:42.250470Z digest=sha256:59897491043507feed50b0361e79dcc19a10b64bd73b3bc88b8a471d059b64fa