Pith. sign in

Paper Citation Record · LEDGER

Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2504.11168.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.11168 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:03.097933Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f296bbb6-81ee-408c-a4db-8080aae50f7d · inbound

A Byzantine Fault Tolerance Approach towards AI Safety cites this paper.

A Byzantine Fault Tolerance Approach towards AI Safety Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:16:58.783644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:15:09.861706Z digest=sha256:9c27c5a54c5410741cc02b8dc50b1cd781e516a999cb60d483d081a18c3c5685

Observation 656ad9a6-f126-46b6-be17-5d45e6ad8bb2 · inbound

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection cites this paper.

EVADE-Bench: Multimodal Benchmark for Evaluating and Enhancing Evasive Content Detection Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:03.097933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:03.097933Z digest=sha256:17b85fb82a96e57dda7b4d132f697be4ea27870f1b8d8bfe2f0934cad9d2d167

Observation a94fe550-8ff2-4ed4-b795-cb290d71e09d · inbound

When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review cites this paper.

When Your Reviewer is an LLM: Biases, Divergence, and Prompt Injection Risks in Peer Review Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T18:34:49.857445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T18:34:49.857445Z digest=sha256:ce997cdacdd5a737acc88699249b2c975923e9726490d85c3c55347eac83ca0a

Observation 76bcc266-7f5b-4bf9-ae9a-6f331c26adf1 · inbound

Exploiting Web Search Tools of AI Agents for Data Exfiltration cites this paper.

Exploiting Web Search Tools of AI Agents for Data Exfiltration Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:36:07.217273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T08:36:04.262528Z digest=sha256:42077f389e27681045656c8ab19e62d35119f30de8c4926680025b9f15bb9915

Observation cbe39f0e-1efa-492f-9af0-78502964a57d · inbound

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection cites this paper.

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T15:56:48.257415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T15:56:48.257415Z digest=sha256:fda8d0a015a3b04c8ed04c99c4c373f013c67e3a91a27d33cf03575727c6907a

Observation 5b6e66dc-3014-4f87-bc5f-81cb25b6775c · inbound

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations cites this paper.

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-02T03:36:29.442477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T09:55:56.227335Z digest=sha256:be71ffb4433d6a8508b4a8c2a40e4847613f9b30c8e1985c34728a8b4e46bcec

Observation 9c7a0633-192b-4021-b033-0c2f1c7bfaa4 · inbound

Short paper: Models in the dark -- Rectification and erasure under GDPR in ML supply chains cites this paper.

Short paper: Models in the dark -- Rectification and erasure under GDPR in ML supply chains Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-28T03:31:30.366296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T03:24:45.955589Z digest=sha256:37c81a26ae3e97a354c8937a14f401d9eb405fc69f89f01c7d8dea9baa40a3f6

Observation adde285b-0aec-4e56-bb24-13ed97d70d5a · inbound

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails cites this paper.

From Shield to Target: Denial-of-Service Attacks on LLM-Based Agent Guardrails Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:58:43.511470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T04:42:05.886984Z digest=sha256:cf63fed8b027a227d80009cac4a7276c00e2803bcff76f1bbee1824d2d959023

Observation c42ea2ac-1377-4141-8751-65258d51ca10 · inbound

Investigating The Security of Modern AI and Cloud Infrastructure cites this paper.

Investigating The Security of Modern AI and Cloud Infrastructure Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:42.302084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T11:31:39.910784Z digest=sha256:4623608b48c5cff471ba14ce2f116b3774a6ff6cdaaf89c9998cb2ada8ecdd0a

Observation 3d867835-2020-452d-a5c9-1100f787cedb · inbound

Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines cites this paper.

Toward Self-Evolution-Ready Workflow Harnesses: A Reversible Migration Path and Convertibility Taxonomy for Expert LLM Pipelines Bypassing LLM Guardrails: An Empirical Analysis of Evasion Attacks against Prompt Injection and Jailbreak Detection Systems

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:48:46.404918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T03:39:24.426862Z digest=sha256:78d54bc9af3b2fa708a9e9dd52884911d8eabf26b620ba78d116460cbf47aad7