Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:03.780401Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 7 inbound Pith citation observations for arXiv:2505.18889.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:27:03.780401Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:12:22.885999Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-17T00:31:24.539214Z
84 of 84 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1b66ac07-a466-4c8b-9424-707bdcc9a53e · outbound
Security Concerns for Large Language Models: A Survey Concrete Problems in AI Safety
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40fc1189-3a08-465a-b301-751bdf4b694b · outbound
Security Concerns for Large Language Models: A Survey System Card: Claude Opus 4 & Claude Sonnet 4
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a406cd1d-1821-4e37-b4ca-a6701e8d1913 · outbound
Security Concerns for Large Language Models: A Survey Theclaude3modelfamily:Opus,sonnet,haiku
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 317e49e5-374c-40d5-8786-b7729eb72d50 · outbound
Security Concerns for Large Language Models: A Survey Embedding-based classifiers can detect prompt injection attacks
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97ce77fa-9edb-42ac-a154-3bd1a17c4167 · outbound
Security Concerns for Large Language Models: A Survey Monitoring Reasoning Models for Misbehavior and the Risks of Promoting Obfuscation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 466f5ac9-9c5a-48fe-80b1-f32da3f1006f · outbound
Security Concerns for Large Language Models: A Survey Deception in LLMs: Self-Preservation and Autonomous Goals in Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08abe4a7-9e46-4070-a905-96ca93301b08 · outbound
Security Concerns for Large Language Models: A Survey LIAR: Leveraging Inference Time Alignment (Best-of-N) to Jailbreak LLMs in Seconds
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 554e00f1-c718-43d1-882c-12be9106f299 · outbound
Security Concerns for Large Language Models: A Survey Superintelligent Agents Pose Catastrophic Risks: Can Scientist AI Offer a Safer Path?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf7d8278-25e5-430d-8210-761a9ce1c0b0 · outbound
Security Concerns for Large Language Models: A Survey Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3852f90b-0c1a-4b47-bc46-85e3acf1666e · outbound
Security Concerns for Large Language Models: A Survey Language models are few-shot learners
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ec921e-878c-44c2-9900-bfff5d05ad88 · outbound
Security Concerns for Large Language Models: A Survey Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48eb96e2-9bbc-4668-9c72-6313023d7d9b · outbound
Security Concerns for Large Language Models: A Survey Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95ad58b2-1d5a-4c5a-8193-bf587a9eb541 · outbound
Security Concerns for Large Language Models: A Survey Here Comes The AI Worm: Unleashing Zero-click Worms that Target GenAI-Powered Applications
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50b95a10-6167-4fa7-89c3-058e2b64e073 · outbound
Security Concerns for Large Language Models: A Survey Formally Specifying the High-Level Behavior of LLM-Based Agents
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1762adc4-8824-456b-b369-5d8da4c9c261 · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e49c603-d220-482d-b4be-a746810aba26 · outbound
Security Concerns for Large Language Models: A Survey Security and privacy chal- lengesoflargelanguagemodels:Asurvey
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4b6a04d7-0ff7-43f8-9b06-3d0a76b46036 · outbound
Security Concerns for Large Language Models: A Survey Emerging Security Challenges of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a31531c-5d8d-4bc5-8768-c8c5491ee37c · outbound
Security Concerns for Large Language Models: A Survey The Philosopher's Stone: Trojaning Plugins of Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb2fdb16-1cda-447a-87f1-5527d52d15b9 · outbound
Security Concerns for Large Language Models: A Survey StruPhantom: Evolutionary Injection Attacks on Black-Box Tabular Agents Powered by Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d82fce64-904b-421d-9092-11a1acc7c9e9 · outbound
Security Concerns for Large Language Models: A Survey Adversarial Tokenization
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20143416-9609-4ad7-9888-37115c98b2cd · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 428190b1-bbfa-4d5d-bc45-1d9bbdbd8c59 · outbound
Security Concerns for Large Language Models: A Survey Deliberative Alignment: Reasoning Enables Safer Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62e05885-63d8-47b6-b4a4-1f4c924b55e6 · outbound
Security Concerns for Large Language Models: A Survey System prompt poisoning: Persistent attacks on large language models beyond user injection
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c592c1ef-f1f2-4a3f-a571-497ca4cfcd5f · outbound
Security Concerns for Large Language Models: A Survey Red-Teaming LLM Multi-Agent Systems via Communication Attacks
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f70f9be-cdb5-4ff8-8fe7-82d885618a5c · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d5d53c0-54d1-46bf-a773-794ecbf44098 · outbound
Security Concerns for Large Language Models: A Survey Pleak: Promptleakingattacksagainstlargelanguagemodelapplications,in: Proceedings of the 2024 on ACM SIGSAC Conference on Computer and Communications Security, pp
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c19edb68-7eff-4127-a61e-c8bac5e6f212 · outbound
Security Concerns for Large Language Models: A Survey Poisongpt:Howwehidalobotomized llmonhuggingfacetospreadfakenews
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a94f58ae-8610-4a7f-b05b-0dd5f644d0a1 · outbound
Security Concerns for Large Language Models: A Survey Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98d8d85c-1ef6-47f4-a1d3-809687dc6d90 · outbound
Security Concerns for Large Language Models: A Survey Llmsecurity101:Defendingagainstprompthacks
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f1994eb4-52c4-481a-beca-94f65dd64760 · outbound
Security Concerns for Large Language Models: A Survey Baseline Defenses for Adversarial Attacks Against Aligned Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f5ae23f-3b17-494b-8494-438332375b7b · outbound
Security Concerns for Large Language Models: A Survey A watermark for large language models, in: International Conference on Machine Learning, PMLR
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30fba0b6-0662-48c2-9a97-08c7f135c1c9 · outbound
Security Concerns for Large Language Models: A Survey Fun-tuning: Characterizing the Vulnerability of Proprietary LLMs to Optimization-based Prompt Injection Attacks via the Fine-Tuning Interface
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d70dfd76-74a7-481c-9dd7-3d4d073c52c1 · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c6d1c47-9a87-431d-b297-a701ad27e765 · outbound
Security Concerns for Large Language Models: A Survey Prefill-level Jailbreak: A Black-Box Risk Analysis of Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72ec2fdd-26f8-470c-944b-1b448705f2fb · outbound
Security Concerns for Large Language Models: A Survey Backdoorllm: A comprehensive benchmark for backdoor attacks on large language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8abbee08-07ad-450d-a4eb-89465f60a515 · outbound
Security Concerns for Large Language Models: A Survey Measuring Faithfulness in Chain-of-Thought Reasoning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 663ca6f8-5b6e-4dd9-92f1-0f8e03ee9663 · outbound
Security Concerns for Large Language Models: A Survey Autohijacker: Automatic indirect prompt injection against black-box llm agents, in: Submitted to ICLR 2025.https://openreview.net/forum?id=11629
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2224e001-e7fb-46b9-9f51-819fc26db4d9 · outbound
Security Concerns for Large Language Models: A Survey FlipAttack: Jailbreak LLMs via Flipping
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 804323e0-a0cf-48dc-8476-d5e75e7d5b45 · outbound
Security Concerns for Large Language Models: A Survey Nature Machine Intelligence , 1–14
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61c3d0a4-f14d-4110-a060-48a4c97eaa41 · outbound
Security Concerns for Large Language Models: A Survey Agentic misalign- ment: How llms could be an insider threat
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e8d594a7-3a99-4da9-bf24-488c8974f8f0 · outbound
Security Concerns for Large Language Models: A Survey Tree of attacks: Jailbreaking black- box llms automatically
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 594998f1-24d2-4b97-9986-48c367269cce · outbound
Security Concerns for Large Language Models: A Survey Formalizing and benchmarking prompt injection attacks and defenses, in: 33rd USENIX Security Symposium (USENIX Security 24), pp
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 76516eb3-79fb-4df2-8745-0a325c84f7ea · outbound
Security Concerns for Large Language Models: A Survey Fully autonomous ai agents should not be developed
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 944c7bf4-9622-499a-8461-ec9b37e178a4 · outbound
Security Concerns for Large Language Models: A Survey AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af322d9b-6b9f-42ed-b73c-75edf804a0ce · outbound
Security Concerns for Large Language Models: A Survey Frontier Models are Capable of In-context Scheming
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f40d4b36-c785-4193-8798-510bb0213b15 · outbound
Security Concerns for Large Language Models: A Survey Eliciting and Analyzing Emergent Misalignment in State-of-the-Art Large Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6fa50bd-3b22-4322-b72c-ce82f7a1333f · outbound
Security Concerns for Large Language Models: A Survey Neural Exec: Learning (and Learning from) Execution Triggers for Prompt Injection Attacks
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77b8372a-7ad2-4b67-8a78-a7e57022c71f · outbound
Security Concerns for Large Language Models: A Survey GPT-4 Technical Report
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce91f6c3-7ae0-4a5f-a822-9edf9615ec9e · outbound
Security Concerns for Large Language Models: A Survey Hijacking Large Language Models via Adversarial In-Context Learning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26580501-d266-4a08-9291-d8ff75199f53 · outbound
Security Concerns for Large Language Models: A Survey From chatbotstophishbots?:Phishingscamgenerationincommerciallarge languagemodels,in:2024IEEESymposiumonSecurityandPrivacy (SP), IEEE
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a12f2e67-d3ee-4924-85bc-4a8b38807754 · outbound
Security Concerns for Large Language Models: A Survey Ignore Previous Prompt: Attack Techniques For Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24375860-c72e-4ad5-b9c4-f90a96155f81 · outbound
Security Concerns for Large Language Models: A Survey BadGPT: Exploring Security Vulnerabilities of ChatGPT via Backdoor Attacks to InstructGPT
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12b4b3aa-060e-4106-80c0-cf97e3f9d79d · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3a5fd50e-8463-4378-b555-0104ecddeb33 · outbound
Security Concerns for Large Language Models: A Survey Survey of Vulnerabilities in Large Language Models Revealed by Adversarial Attacks
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda06f9a-a7ab-4138-8483-610e2173d028 · outbound
Security Concerns for Large Language Models: A Survey Gemini: A Family of Highly Capable Multimodal Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 345b4c22-c783-4260-9166-3d5cb358380e · outbound
Security Concerns for Large Language Models: A Survey Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fff1a74-7046-469b-a5d8-c0b24f7ea928 · outbound
Security Concerns for Large Language Models: A Survey AdvancesinNeural Information Processing Systems 36, 61836–61856
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2440f6fb-d53c-47cf-a777-1a03490ad880 · outbound
Security Concerns for Large Language Models: A Survey Wormgpt and fraudgpt – the rise of malicious llms
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a4ee6f0e-964d-4f5d-8547-a982f9c59d9a · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3c72e84-9a5e-479f-9a66-84669cef5015 · outbound
Security Concerns for Large Language Models: A Survey Poisoning Language Models During Instruction Tuning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc0ea891-fa4a-49c0-a147-d83ddf0fe92d · outbound
Security Concerns for Large Language Models: A Survey DAN is my new friend.https://old.reddit.c om/r/ChatGPT/comments/zlcyr9/dan_is_my_new_friend/
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f43a025c-03cf-44e5-83f9-236dd5fbf0aa · outbound
Security Concerns for Large Language Models: A Survey Universal adversarial triggers for attacking and analyzing nlp, in: EMNLP
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc1e993c-71e8-4590-b811-d4843a5ebd9d · outbound
Security Concerns for Large Language Models: A Survey When Thinking LLMs Lie: Unveiling the Strategic Deception in Representations of Reasoning Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c7786f9-4dbd-4f3b-aa2b-3d99c574d136 · outbound
Security Concerns for Large Language Models: A Survey The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68defda7-725f-4125-a8fd-b54e3b8798ad · outbound
Security Concerns for Large Language Models: A Survey Jailbroken: How does llm safety training fail? Advances in Neural Information Processing Systems 36, 80079–80110
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 63ce99cb-7ea7-45b3-b288-d427f9a6e538 · outbound
Security Concerns for Large Language Models: A Survey AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15125eaa-6240-4c15-bf31-a7767070f8d6 · outbound
Security Concerns for Large Language Models: A Survey Trojan Activation Attack: Red-Teaming Large Language Models using Activation Steering for Safety-Alignment
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d99dff0-7482-4de9-a3f6-cb55c68aaff5 · outbound
Security Concerns for Large Language Models: A Survey Nuclear Deployed: Analyzing Catastrophic Risks in Decision-making of Autonomous LLM Agents
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d4cf9b3-0620-496e-ac16-24d848e360e7 · outbound
Security Concerns for Large Language Models: A Survey Persona features control emergent misalignment, 2025
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b8871b8-d32b-4189-af08-9dbd0804a0c4 · outbound
Security Concerns for Large Language Models: A Survey Asurvey on large language model (llm) security and privacy: The good, the bad, and the ugly
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64e5fd1a-3be0-4c55-99b3-8395982cc5e0 · outbound
Security Concerns for Large Language Models: A Survey RedAgent: Red Teaming Large Language Models with Context-aware Autonomous Language Agent
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 702d2ad4-02d6-4f69-82be-ff06d645237e · outbound
Security Concerns for Large Language Models: A Survey Rejection Improves Reliability: Training LLMs to Refuse Unknown Questions Using RL from Knowledge Feedback
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1a8b09f-38b2-4f95-b3b9-29b4a8fdcfa5 · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8b9d043f-428c-4a8d-adb8-edbe1f111dcf · outbound
Security Concerns for Large Language Models: A Survey Backdooring Instruction-Tuned Large Language Models with Virtual Prompt Injection
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a515249-b67b-43be-8da3-9b87c3a00f40 · outbound
Security Concerns for Large Language Models: A Survey AutoRedTeamer: Autonomous Red Teaming with Lifelong Attack Integration
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb96e0a4-8b38-43f2-ace2-cc8a809bae58 · outbound
Security Concerns for Large Language Models: A Survey LLM-Virus: Evolutionary Jailbreak Attack on Large Language Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2499618d-1ed2-418d-b7ba-c5cbdc64ff78 · outbound
Security Concerns for Large Language Models: A Survey A Closer Look at Machine Unlearning for Large Language Models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57120ddb-8970-4d96-b258-a53bc8fb3204 · outbound
Security Concerns for Large Language Models: A Survey Evaluation of LLM Vulnerabilities to Being Misused for Personalized Disinformation Generation
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 82d11667-9a15-4cae-a3fd-e0a0b67e667d · outbound
Security Concerns for Large Language Models: A Survey Unresolved cited work
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eb482b29-547d-4846-bd6b-42e9cf4b000f · outbound
Security Concerns for Large Language Models: A Survey Weak-to-Strong Jailbreaking on Large Language Models
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab4d47a7-5cae-41a2-9f9f-014b5ea4b4b0 · outbound
Security Concerns for Large Language Models: A Survey Ad- vprefix: An objective for nuanced llm jailbreaks
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6956a007-7c79-41f6-9f67-bbae36cce66c · outbound
Security Concerns for Large Language Models: A Survey Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 099fde57-8128-44e5-a0b5-d7ee8ff11def · outbound
Security Concerns for Large Language Models: A Survey Safe RLHF: Safe Reinforcement Learning from Human Feedback
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f03154d5-a4b3-4d51-a0de-e83b582db2b2 · outbound
Security Concerns for Large Language Models: A Survey Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ee4e6f2-7cd7-497b-b623-98bf6ffa8b58 · inbound
Bridging AI and Software Security: A Comparative Vulnerability Assessment of LLM Agent Deployment Paradigms Security Concerns for Large Language Models: A Survey
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b093e5c6-e241-448f-98f4-13db15d84283 · inbound
LLM Harms: A Taxonomy and Discussion Security Concerns for Large Language Models: A Survey
Reference 130
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a210001a-a16e-4ba2-9984-e67cd692f373 · inbound
LLM Harms: A Taxonomy and Discussion Security Concerns for Large Language Models: A Survey
Reference 130
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7698991f-040b-4cc0-affd-9512d14552e9 · inbound
Zero-Shot Embedding Drift Detection: A Lightweight Defense Against Prompt Injections in LLMs Security Concerns for Large Language Models: A Survey
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89b8a45c-01e6-42eb-8b6f-309f69ae204e · inbound
Gecko: A Simulation Environment with Stateful Feedback for Refining Agent Tool Calls Security Concerns for Large Language Models: A Survey
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4ecefb4-f0f3-44a9-a41d-79f573a1a009 · inbound
LLM-as-Judge Framework for Evaluating Tone-Induced Hallucination in Vision-Language Models Security Concerns for Large Language Models: A Survey
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f11db4e2-a816-4d47-aad9-ec665cf4ea53 · inbound
Adversarial Prompting Framework for AI Safety Assessment Security Concerns for Large Language Models: A Survey
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.