Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:25:45.961893Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 100 of 255 outbound references and 3 inbound Pith citation observations for arXiv:2510.15476.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:25:45.961893Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:32:02.668360Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T20:00:07.826531Z
100 of 255 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9b9a8f4d-0a44-4fef-974e-f8169200dfc4 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses glaive-code-assistant, 2023
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baf3582e-15f8-4574-8754-8951849bde68 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Detecting Language Model Attacks with Perplexity
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b077a8a-59d7-4555-b6f9-9457827ace9d · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Are PPO-ed Language Models Hackable?
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a359a567-a5ea-48b8-bab4-853163c268df · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Jailbreaking Leading Safety-Aligned LLMs with Simple Adaptive Attacks
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed84495c-ce25-4a75-8c76-f0d6ba08505b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Jailbreaking leading safety-aligned llms with simple adaptive attacks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fea0ec6-61d4-4d00-aa42-eef9b3cf69bf · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Does Refusal Training in LLMs Generalize to the Past Tense?
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec056e70-a0bf-40d1-87a5-ffe022dfc284 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Does refusal training in llms generalize to the past tense? InThe Thirteenth International Conference on Learning Representations, ICLR 2025, Singapore, April 24-28, 2025
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a7de3c8-ddd0-42b6-9c48-785eba686cbe · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Bowman, Ethan Perez, Roger B
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f92f486-5722-4298-9119-79da1760dfc2 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Gemini: A Family of Highly Capable Multimodal Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3aba117-b1ff-40dd-921b-6140cac603f1 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Surro- gateprompt: Bypassing the safety filter of text-to-image models via substitution
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7739b3f9-3994-46b0-826b-960823cce346 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Qwen Technical Report
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b346aeeb-be63-4fe1-9478-41fe118ee20b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a633e9ce-f650-46b3-bd21-9df0549dcbef · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses How (un)ethical are instruction-centric responses of LLMs? Unveiling the vulnerabilities of safety guardrails to harmful queries
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e41f6a8a-e9f9-4042-9641-5bcc6b2df7a1 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3018bdf-9640-480a-8863-347328d5ded5 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses CLARE: Cognitive Load Assessment in REaltime with Multimodal Data
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6386057-4902-4bd3-8347-511b75b75692 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9714b2ec-69f2-40db-9f43-a930fd320da8 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34530978-c609-4bd9-a876-bac5191efeda · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Camelai domain expert datasets (physics, math, chemistry & biology), 2023
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0596d4c6-4e43-4580-9e75-31ef44c45463 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Abu-Ghazaleh, M
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049cd096-529f-41f3-a266-a0ea331e1ad1 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Synneure: Intelligent human-machine teamwork in virtual space
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12a379a3-78ae-46e8-95f9-842f97684580 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Pappas, Florian Tram` er, Hamed Hassani, and Eric Wong
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77568c13-4dc4-4203-bf10-059230c62010 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30779e11-7784-4282-91b0-1954946d0ce4 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Pappas, and Eric Wong
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32e7f717-16b9-4033-a048-8ed85123dc67 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses StruQ: Defending Against Prompt Injection with Structured Queries
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 431632c8-01bb-403c-b47c-a9451766b5ef · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses When LLM meets DRL: advancing jailbreaking efficiency via drl-guided search
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83850c27-a627-4be6-86ff-bd3abf98e28d · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses RL-JACK: Reinforcement Learning-powered Black-box Jailbreaking Attack against LLMs
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50577956-3357-40d7-a052-48ca48250d2a · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Leveraging the Context through Multi-Round Interactions for Jailbreaking Attacks
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b6fa656-4f8b-4db8-b93a-60076a35138b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Prompt Injection Attacks on Large Language Models in Oncology
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bedcacc9-ef63-49fd-9418-127f1168ef4c · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Cogstack, 2023
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93f9ec51-6210-42a8-a429-997c188f1592 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses A Jailbroken GenAI Model Can Cause Substantial Harm: GenAI-powered Applications are Vulnerable to PromptWares
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a0e7d8-929c-45da-9d63-741195854150 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses MoJE: Mixture of Jailbreak Experts, Naive Tabular Classifiers as Guard for Prompt Attacks
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c45bcb92-58bf-43b9-a57b-83d2ccbcb919 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Moje: Mixture of jailbreak experts, naive tabular classifiers as guard for prompt attacks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab74301b-214b-4d37-b953-ee73a08f750b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses alpaca-gpt4-cot, 2024
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f38120c-10fc-4679-954b-6e5b77f129bb · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Dataset and lessons learned from the 2024 satml LLM capture-the-flag competition
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a0ab53c-274f-4eea-9840-214d7c8b968b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses AgentDojo: A Dynamic Environment to Evaluate Prompt Injection Attacks and Defenses for LLM Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00863392-c6b6-418d-a1df-f21c64a6756a · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Attack prompt generation for red teaming and defending large language models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9817fec3-0c68-4669-a3f1-b040e5fcea88 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses MasterKey: Automated Jailbreak Across Multiple Large Language Model Chatbots
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c0115b3-6007-4d3b-a37c-5413d25effe7 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Pandora: Jailbreak GPTs by Retrieval Augmented Generation Poisoning
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe063ec7-405b-46c2-bccd-e3c714248afd · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Multilingual jailbreak challenges in large language models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e4a384a-071d-4ea2-9755-6c5ee37db158 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses A wolf in sheep’s clothing: Generalized nested jailbreak prompts can fool large language models easily
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03e90c74-720b-4f37-b830-b4c646a49900 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7191af2-caa0-40af-ba24-745f0d852e4f · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Harnessing Task Overload for Scalable Jailbreak Attacks on Large Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7645fa9-ea0d-47a0-a6de-365ab87a2eb1 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12484ebc-aed0-47c4-b236-83de6b3fcc95 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses h4rm3l: A language for Composable Jailbreak Attack Synthesis
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f7eb929-631e-42f1-a104-96d669327d06 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Multi-turn jailbreaking large language models via attention shifting
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9594e70f-3e35-4f86-aa52-c6c430098b00 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses VLMGuard: Bootstrapping Malicious Prompt Detectors from Unlabeled Vision-Language Prompts in the Wild
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02103ce5-3b21-4a7d-aa4c-47433d021e50 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Atoxia: Red-teaming Large Language Models with Target Toxic Answers
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e408f8ef-b6a3-48a5-bccf-d2bfce19d96e · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Vr-hand-in-hand: Using virtual reality (VR) hand tracking for hand-object data annotation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddb40eed-ee9a-48ad-af23-4dd5eaa62854 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses A comprehensive survey of attack techniques, implementation, and mitigation strategies in large language models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 621c4223-71ea-4e8f-8f65-626e4fd54b5f · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Unbridled icarus: A survey of the potential perils of image inputs in multimodal large language model security
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff93b40b-3006-4fca-bfc7-8ceef6c257e9 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses JailbreakLens: Visual Analysis of Jailbreak Attacks Against Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41d41791-47f3-478b-8b93-8b1a43fcb9fe · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Safety alignment in NLP tasks: Weakly aligned summarization as an in-context attack
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc437b38-c7f0-44d9-9a3f-f019cf882af7 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Cross-Task Defense: Instruction-Tuning LLMs for Content Safety
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0104abb8-b93e-4a6d-ba73-08658c0f765b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Merging Improves Self-Critique Against Jailbreak Attacks
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db601228-2c5c-4aa0-865b-c89d4baef68e · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses MART: improving LLM safety with multi-round automatic red-teaming
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4c1acff-d1c7-4c9a-8f25-cb5f0c6c7363 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Coercing LLMs to do and reveal (almost) anything
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbb98816-ca6a-40f5-96a0-89c93ed6d197 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Attacking Large Language Models with Projected Gradient Descent
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2d13794-9a2b-4399-9469-ee2f79f5efd0 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses AEGIS: Online Adaptive AI Content Safety Moderation with Ensemble of LLM Experts
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea59c527-9285-4805-8af7-78b1f2e8efef · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Emerging Vulnerabilities in Frontier Models: Multi-Turn Jailbreak Attacks
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcd95c5d-ffb9-484a-b7fc-9d929a2e496a · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Breach By A Thousand Leaks: Unsafe Information Leakage in `Safe' AI Responses
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f41054c-378c-4f36-a00b-8c506c3c32b4 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Gradient-based adversarial attacks against text transformers
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ad2add-d635-41d4-859f-6f32aa0172ad · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Cold-attack: Jailbreaking llms with stealthiness and controllability
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d461019a-a9cc-49b4-bdad-6cf6e54bf68f · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses From chatgpt to threatgpt: Impact of generative AI in cybersecurity and privacy.IEEE Access, 11:80218–80245, 2023
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3301909c-d3bb-4f92-aac1-796af4abbf4f · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Unresolved cited work
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42915daa-e987-4308-a05b-cfa38d400ce0 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Gajera, and Chitta Baral
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6791c844-f679-405d-9538-5d0b15001a93 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Wildguard: Open one-stop moderation tools for safety risks, jailbreaks, and refusals of llms
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d9b2e1d-9a45-4dac-b6bc-f9f8ed73cb27 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Privacy of federated QR decomposition using additive secure multiparty compu- tation.IEEE Trans
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43b66e3c-e880-4716-a81e-adb2da2b930e · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Unresolved cited work
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fe12d56-4d6b-427d-8c8b-16655549bc71 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Springer Briefs in Computer Science
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cc54660-f3a2-4a99-918c-f20b7b49f2c7 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Query-based adversarial prompt generation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0735b77b-8282-401f-8bfe-7b518102fb1a · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Gradient cuff: Detecting jailbreak attacks on large language models by exploring refusal loss landscapes
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 455d9866-9f50-4a02-8f64-8bf5c5e63266 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Defending against indirect prompt injection attacks with spotlighting
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c51da07-703f-4b26-94b6-c2c89dfdf775 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Unresolved cited work
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16c3402c-6f54-42f2-840b-f2182031128a · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses SafeAligner: Safety Alignment against Jailbreak Attacks via Response Disparity Guidance
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96ae2779-3cc2-4ded-9764-a56c00db3b65 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Summon a Demon and Bind it: A Grounded Theory of LLM Red Teaming
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 592b08cf-b8fc-47db-84f2-61193b21fecc · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Catastrophic jailbreak of open-source llms via exploiting generation
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4caaa46-b992-4e71-9ba7-e9340233975d · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Beavertails: Towards improved safety alignment of LLM via a human-preference dataset
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd32e0e4-2fb8-4cb4-882b-c3330f9b24c6 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a4f3330-ff82-440f-b9a5-8e93d4572fc1 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Automated progressive red teaming
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a66055a-f2f6-4ede-9f09-0dac4dcfb177 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Improved techniques for optimization-based jailbreaking on large language models
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b67286f1-d679-48a7-8b1a-f8d062120594 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses How Well Do LLMs Handle Cantonese? Benchmarking Cantonese Capabilities of Large Language Models
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 993f03bd-2797-4073-b236-6c4772f2f376 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Artprompt: ASCII art-based jailbreak attacks against aligned llms
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc745b0d-06b8-4b6d-aa5d-f51d89c7c6e0 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses An Optimizable Suffix Is Worth A Thousand Templates: Efficient Black-box Jailbreaking without Affirmative Phrases via LLM as Optimizer
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b48ce965-8368-4aac-839f-57502c835e31 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58b33b15-decc-4f83-94c5-397ff037b586 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses GUARD: role-playing to generate natural-language jailbreakings to test guideline adherence of large language models.CoRR, abs/2402.03299, 2024
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cb0e55-3ef8-44af-b1be-3c2f0f1f0849 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac69714-b70b-42f7-9c0a-9d2a565ae173 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses airoboros-2.2, 2023
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41e4b29d-9027-4200-8c48-8ec8fb5729b6 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Jailbreakzoo: Survey, landscapes, and horizons in jailbreaking large language and vision-language models.CoRR, abs/2407.01599, 2024
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61c77a38-f94c-43af-8797-e095fed36918 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Adversaries Can Misuse Combinations of Safe Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4b97758-ce27-4c7d-9aab-6405ad414d27 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Dragan, Aditi Raghunathan, and Jacob Steinhardt
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 335c8800-3be9-4431-add4-ab1500a44bb0 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Prompt Injection Attacks in Defended Systems
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c932a8fe-9906-402a-b4f3-6e6a69f40eab · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Exploiting programmatic behavior of llms: Dual-use through standard security attacks
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d03ab2ff-9879-4317-98e8-9dd6494ca598 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Testing the Limits of Jailbreaking Defenses with the Purple Problem
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b80ef8-a9d3-4ade-9607-bfd534c63dad · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Break the Breakout: Reinventing LM Defense Against Jailbreak Attacks with Self-Refinement
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 510fa07b-1d80-43a3-9c68-e3914854dc9c · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Certifying LLM Safety against Adversarial Prompting
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9abbf7f-1f7f-47bf-9ac7-84ff298caac9 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Empirical Analysis of Large Vision-Language Models against Goal Hijacking via Visual Prompt Injection
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d877c2-d706-4788-9dcf-4d42d5b19472 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Platypus: Quick, Cheap, and Powerful Refinement of LLMs
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ab876c1-6575-4d55-b711-3fb3217f109b · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Open Sesame! Universal Black Box Jailbreaking of Large Language Models
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16f4b4f1-0897-4ce5-b30d-3bc42dd98d7d · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses A Cross-Language Investigation into Jailbreak Attacks in Large Language Models
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c28a4b1-0b8e-4d39-9786-d5530b3e1275 · outbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Multi-step jailbreaking privacy attacks on chatgpt
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6354cbb-4b4a-47b8-b2c4-482deab1990e · inbound
One Goal, Many Commands: Characterizing Denylist Fragility in AI Agents SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dcc69231-6c3f-4338-9e8e-94f2e02cdc18 · inbound
Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f9cbcb8c-5570-4e46-a76b-b0d6fa736dd8 · inbound
SoK: Intent-Oriented Systematization of Multi-Turn LLM Jailbreaks SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.