Pith. sign in

Paper Citation Record · LEDGER

JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2312.10766.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.10766 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:28:09.757042Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T18:57:16.788391Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 84898307-2d6a-460b-9b0e-844b61e48c63 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.478385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:da01c6a04363416f6b0466edd41b96fe92b0d9011f9ea7f7587512bf08f22329

Observation 16d0d785-3c51-4686-beff-511e41aa0e5f · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 285

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.065394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:751c3a2f04dc9bd7e5f3d63102da00880a671f23067f90e45557ad91879027c3

Observation d2c92e86-d9b9-42da-a6d9-68ff65260dcd · inbound

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion cites this paper.

RedDiffuser: Auditing Multimodal Safety Failures in Vision-Language Models via Reinforced Diffusion JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-23T00:15:14.890497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T00:13:08.603115Z digest=sha256:51a0f52f8f9c7d968e66232c37be710c9f702ba26e0054d1cd3967330ef403c8

Observation 6d76b6df-9bab-4d18-b895-567e9462008f · inbound

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking cites this paper.

PRISM: Programmatic Reasoning with Image Sequence Manipulation for LVLM Jailbreaking JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-19T03:37:01.100658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-19T03:36:24.013477Z digest=sha256:5edce7a8d22065ba5eff60b7a4897fa5e9df70e7f446c24b03e85838921e76cc

Observation 5958d44c-33f8-4679-820f-2af6eaebb139 · inbound

Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors cites this paper.

Eyes-on-Me: Scalable RAG Poisoning through Transferable Attention-Steering Attractors JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T13:28:09.757042Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:28:09.757042Z digest=sha256:1f1b294ab768111dc4f26a702f92f505955422193288d690179e57c1cada1557

Observation 82d3e7b4-c6d6-4e96-8b15-bd50fd26f621 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 232

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.307952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.307952Z digest=sha256:854c067fd31c0d8b1249ed3d1d631fddd62aa1e241e4532bf5cb237c64c4556f

Observation b29b5132-8deb-4996-9b96-274fe6c9a157 · inbound

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute cites this paper.

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:13:43.357656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-20T20:12:03.715605Z digest=sha256:b364cc3f2982df646696adba6fe53e2de35d3ff452311b007729d451ef603a1a

Observation ef969e13-3f7d-458f-b67f-603b00687965 · inbound

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models cites this paper.

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 76

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T18:57:16.789828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-27T21:47:22.295896Z digest=sha256:44ce411c8955e4e016951d532b7d1be1332a895e04031fa3fe6730c1efe6e4af

Observation 97814ace-f5d1-403f-b93e-e31e5bdf6161 · inbound

Safe responses matter: Output-aware safety guardrail mitigate over-refusal in MLLMs cites this paper.

Safe responses matter: Output-aware safety guardrail mitigate over-refusal in MLLMs JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-14T17:33:49.041193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T17:33:49.041193Z digest=sha256:009f9f6416dd1c359fc57d76c55d2088ff0c37884eb32ce3121ea6fee8176316