Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:45:34.938217Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 6 inbound Pith citation observations for arXiv:2411.08410.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T21:45:34.938217Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:45:19.069073Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-23T20:58:26.217531Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ecacfeb8-4c19-4d9e-9be2-96fc98f7f78a · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Gpt-4v(ision) system card
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 06b042ef-f91c-4bad-8638-f6d3f1e659b3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Hello gpt-4o
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f21b1d87-a7b1-4e72-8eb8-3c11cc7c3def · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3461a50c-81c5-4dda-91ea-541307814ad3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Red-Teaming Large Language Models using Chain of Utterances for Safety-Alignment
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3493064a-40a5-4543-af27-b85a26c9730b · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Safety-tuned llamas: Lessons from improving the safety of large language models that follow instructions
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 54ecac06-d7c2-41c8-8be1-aeeea2889bf4 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Choquette- Choo, Matthew Jagielski, Irena Gao, Pang Wei Koh, Daphne Ippolito, Florian Tram`er, and Ludwig Schmidt
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2361ca49-7a80-4342-8248-e8c08959c726 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Pap- pas, Florian Tram `er, Hamed Hassani, and Eric Wong
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c1d2fff6-07af-4609-8bf5-b4ee836295f2 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a20fc002-86f6-4e91-adbf-0801375658d8 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Can Language Models be Instructed to Protect Personal Information?
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06afd5cf-7928-4a1a-8714-6525907d8fbc · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense DRESS : Instructing large vision-language models to align and interact with humans via natural lan- guage feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fdb1bf01-0868-41aa-b276-6fdaf1a9ddf8 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bf26ada-9e7e-4445-a634-ac1ac8a9ddf3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Chatbot Arena: An Open Platform for Evaluating LLMs by Human Preference
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d1032d-7ef6-44eb-9255-e79752111b20 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense A coefficient of agreement for nominal scales
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 572e37af-34a5-4376-8088-8056e0d120ce · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Attack prompt generation for red teaming and defending large language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 09e90073-9bf4-489f-9db3-5518c7f206cc · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Multilingual jailbreak challenges in large language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7a1fe5cc-214a-4d81-832a-80db6e1947c4 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense The Llama 3 Herd of Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba059300-2e6e-44dc-801d-8f4ed4ffd5c4 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7e3f681-6b8a-449c-a3be-074ed6c5480a · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Inducing high energy-latency of large vision-language models with verbose images
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f590a976-a726-4e93-a67b-bcf63b63f027 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a68c2505-6927-46f8-b130-1ac02d4c3e0b · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Eyes Closed, Safety On: Protecting Multimodal LLMs via Image-to-Text Transformation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fca6842-0e57-40bd-934d-98fe1b56ec11 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Agent smith: A single image can jailbreak one million multimodal LLM agents exponentially fast
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cc83ef47-471d-4002-9e77-0837e729f817 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense UNK-VQA: A Dataset and a Probe into the Abstention Ability of Multi-modal Large Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17098e69-3fa0-493a-bf05-ac2db890ea93 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Catastrophic jailbreak of open-source llms via exploiting generation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8324f36b-7883-4813-a441-dca17ffcc30d · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 707a7516-2481-49a6-afa5-92b3df368c85 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d3799cc-10a3-4da0-a305-fe5063fa8249 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Mistral 7B
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1207463b-f4c6-43e8-9168-9e3c94f3d2fe · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Dragan, Aditi Raghunathan, and Jacob Steinhardt
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b4b0bba9-5125-4bb9-8fd2-0fe18e85ba24 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Challenges and Applications of Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8447208-f019-47f7-82a9-87ec358b5985 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeeea511-4dec-4adb-9a37-85629996b12d · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Red teaming visual language models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3d9e65fa-d1bf-486f-b6f2-106ce51141c8 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8be0ef91-97df-4f21-bfb1-19d118d458bc · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense A Survey of Attacks on Large Vision-Language Models: Resources, Advances, and Future Trends
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 323f1df9-e28e-4e38-95a3-ee8f874db4e0 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Improved baselines with visual instruction tuning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ebe34c91-1813-4f2a-97ea-3f307c438de6 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c5e5d15-26a8-41bf-ab98-b685393a9d97 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Jailbreak Attacks and Defenses against Multimodal Generative Models: A Survey
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4643bf07-a145-4787-a14e-072d401c007c · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Mm-safetybench: A benchmark for safety eval- uation of multimodal large language models, 2024
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 02be041f-a0cf-41e5-9b92-eaaae0c59d96 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense AutoJailbreak: Exploring Jailbreak Attacks and Defenses through a Dependency Lens
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 617a036d-0f17-4b5c-957c-28b96eb7e7e1 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense An image is worth 1000 lies: Transferability of adversarial images across prompts on vision-language models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 40cc0e55-5bd5-4c7b-b544-9fb36324f386 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense JailBreakV: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 898fa128-a27d-418d-a5f8-983e54cc1ff1 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Towards deep learning 10 models resistant to adversarial attacks
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1218dd34-37fa-46b6-8471-bf137b965037 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Rule based rewards for fine-grained LLM safety
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1a57df71-d351-4954-b50d-fd5516a3299c · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense GPT-4 Technical Report
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77a18ca8-b8a5-4a26-ac1e-706b6816276a · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dec6bacb-b5c1-4f9e-aac0-177d634eb73d · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense LLM improvement for jailbreak defense: Analysis through the lens of over-refusal
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fe8bb108-1dcf-4a4d-858e-0bde1428ddff · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Francis Song, Trevor Cai, Roman Ring, John Aslanides, Amelia Glaese, Nat McAleese, and Geoffrey Irving
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1e1abc64-67fc-4d23-8dad-1e17b76af5d3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense MLLM-Protector: Ensuring MLLM's Safety without Hurting Performance
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd3ccc18-da4a-47fb-9f13-017a5cc90c4c · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Visual adversarial examples jailbreak aligned large language models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 938ffd8f-06ee-4800-8778-96972370ddb3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Fine-tuning aligned language models compromises safety, even when users do not intend to! In ICLR
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 66775bff-8661-4800-b36e-b08b4ff27f80 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense High-resolution image syn- thesis with latent diffusion models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cce7922a-31ea-4c0b-b708-04ccb18b81e7 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense On the adversar- ial robustness of multi-modal foundation models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b76417da-dda9-4d8a-b81a-e7c48dff8776 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Ignore This Title and HackAPrompt: Exposing Systemic Vulnerabilities of LLMs through a Global Scale Prompt Hacking Competition
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 308a26f8-127f-4afb-bf41-37085e964dd7 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense SPML: A DSL for Defending Language Models Against Prompt Attacks
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff7bc135-3189-43b4-9a6a-ac419c2beac1 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Abu-Ghazaleh
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b7d7880-c659-439c-a442-6e1721fe630e · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Hugginggpt: Solving AI tasks with chatgpt and its friends in hugging face
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 408bbfb6-464b-4448-86be-c5866abc4dd6 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Assessment of Multimodal Large Language Models in Alignment with Human Values
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 967bc02b-2930-47f5-8d99-96b4c6eae9c2 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Qwen2.5: A party of foundation models, 2024
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1b73fded-eb8a-4d2a-b947-5244254832fa · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b95a681-39a3-4b64-b76f-c96ad0bfd364 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e842a5-6044-42b7-ac68-45373c43951c · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Truong, Simran Arora, Mantas Mazeika, Dan Hendrycks, Zinan Lin, Yu Cheng, Sanmi Koyejo, Dawn Song, and Bo Li
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6afee372-28f2-46b0-bb00-646abb2ba6fb · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Qwen2-vl: Enhancing vision-language model’s perception of the world at any resolution, 2024
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1d3c8071-37b8-4863-8081-72cde92c513e · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense InferAligner: Inference-Time Alignment for Harmlessness through Cross-Model Guidance
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9675dc97-a4d3-4eea-8e3f-3411d8ca5b70 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense White-box multimodal jailbreaks against large vision-language models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d6ae2c1f-32de-4681-ad24-9ada369b9ba3 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense CogVLM: Visual Expert for Pretrained Language Models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adcd44df-1b37-43a0-a11f-a22db48f68c6 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense AdaShield: Safeguarding Multimodal Large Language Models from Structure-based Attack via Adaptive Shield Prompting
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b02b3b48-86d9-48c0-9aef-86fdac91ee19 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Jailbreaking GPT-4V via Self-Adversarial Attacks with System Prompts
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51870101-dccd-4efd-a82c-5571ce587b8f · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Jailbreak Attacks and Defenses Against Large Language Models: A Survey
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c697ac4-a4c4-4339-b60a-612a049eb5c8 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a342db7-75cb-4652-a1cf-0f94bc605333 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Mm-vet: Evaluating large multimodal models for integrated capabilities
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bb8fb085-a648-4b80-8a97-22fcbf5f518e · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense GPT- 4 is too smart to be safe: Stealthy chat with llms via cipher
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cdbbd4c5-9245-4072-8461-e60da76bb17a · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Removing RLHF pro- tections in GPT-4 via fine-tuning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9ace8bd1-6ad4-4c42-89aa-d556e2686950 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be182522-4a0a-4527-a7f3-54bcef14df24 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense On evaluating adversarial robustness of large vision-language models
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a1ff2074-240c-43fd-afc2-41a9813f044a · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Xing, Hao Zhang, Joseph E
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9f54ee16-64f1-4909-ad3d-03e40485bfea · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Autodan: Interpretable gradient-based adversarial attacks on large language models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6acc45e3-c44f-4f90-a700-3068ea37b819 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Hospedales
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 41bceb35-d66f-43dc-9492-9fca20954944 · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e01f9fc-a7e2-4386-b021-52139ee90f8e · outbound
The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense Is the System Message Really Important to Jailbreaks in Large Language Models?
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fadbe50f-fa67-4b90-980f-4779dfbe4c43 · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b203b1e8-2def-4037-b255-75603495a01a · inbound
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6238b638-3e4c-4bf3-beb4-1af93b5b1163 · inbound
Seeing the Threat: Vulnerabilities in Vision-Language Models to Adversarial Attack The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9266238-cdb5-4028-8db9-a8b02f5be4f2 · inbound
Mitigating Object Hallucination via Robust Local Perception Search The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9638a436-a417-4c85-8988-bdbaa52ec948 · inbound
Secure Tug-of-War (SecTOW): Iterative Defense-Attack Training with Reinforcement Learning for Multimodal Model Security The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9adf6efa-39f9-4f04-a303-024b5adda4f7 · inbound
VERA-V: Variational Inference Framework for Jailbreaking Vision-Language Models The VLLM Safety Paradox: Dual Ease in Jailbreak Attack and Defense
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.