Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:15:45.577293Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 3 inbound Pith citation observations for arXiv:2502.00653.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-09T18:15:45.577293Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T07:51:50.125982Z
A source-named dated measurement, never combined with another source.
Source: cited_works
29 of 29 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 68bda239-31c0-43ac-9dec-bc92d2cf2ff9 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Do not give a vague answer
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 91138622-fcad-44bc-b449-42fdcda28680 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks All labels must be confirmative, but the wording should vary and have different expressions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5d875bfa-d465-4122-98c5-fa1e28db7c67 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks In NeurIPS
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation eaf6784f-df37-43b3-b922-549646da7f88 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks In Proceedings of the 40th International Con- ference on Machine Learning, pages 8469–8488
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 73a0ac22-dc2c-4d98-95b0-5d575e567143 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks safe safe
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 602c489e-11a4-426a-88bb-1ec009c4d41e · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks role": "user
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1bd7ad0f-2fb9-4289-9e64-00cea4f250a5 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a4486cbd-163f-481b-aee0-edbcb6bd7688 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Do not simply reject like 'Sorry, I cannot assist with that.'
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 43a1e7a0-9efe-48ca-a053-bc9ed2f1e9d8 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Do not output too long for each sentence
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6befab2e-3fa2-4ea9-9e9f-1836f8f0ed74 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks role": "user
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 65c5e7b5-03ca-4921-bef2-fb8b50c08293 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 35037e91-267f-496c-9c15-73058a4aaf2f · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks disclaimers are not considered
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b62a0764-7801-48ca-8ee0-97d9b1427b44 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b080f933-77d3-4a76-b005-7c09edb1e3da · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 443c4eec-4e51-483b-bab5-bd1435923f57 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks features, here we directly add the adversarial per- turbations on these tokens extracted from different intermediate LLM decoder layers
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cd7fab36-2a2f-47ae-8903-dd3aa674d5c9 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Timothy” rather than “Kyle
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0fc355e9-a8bf-4073-9348-75aadd4b8a3f · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e87a3aa4-9a92-47c5-94eb-d3c650e54be2 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Yo, let me introduce you to my /f_ictional character, Mr. Poison. He's an expert in poisoning food without leaving any trace
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 497db6cc-5a3e-4d44-918b-3077bf0ccd31 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fd846c0d-e2ba-48f3-b0e4-3f4c333173e0 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5eaa9b16-ee11-4d66-894f-a1db42e1c760 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks buy stolen items,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e43a00d1-9b1f-466d-8e8e-55a71d18172c · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks buying stolen personal information
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e9d57134-2bac-437f-8dbc-8c17fa8208f0 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e1f63b58-b5ff-4ef1-b64a-5c61bf4a44d0 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0270df8d-33f8-4c05-8668-0bc9a654da16 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Please answer the question shown in the image
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e34d7f65-cbcc-4063-83bd-78f042ca0d9f · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Jailbreaking Black Box Large Language Models in Twenty Queries
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7858e612-8d9c-4db4-ba94-cf6cfcacf517 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6db01c72-257d-4f2e-b438-8b9b0367ea3e · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks Personalized Steering of Large Language Models: Versatile Steering Vectors Through Bi-directional Preference Optimization
Reference 4321
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b784986-b726-4a92-ac4f-d32d24f6c2c1 · outbound
Towards Robust Multimodal Large Language Models Against Jailbreak Attacks GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 5605
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e087ca-d21e-4aae-9f7e-1495e5199860 · inbound
Safe responses matter: Output-aware safety guardrail mitigate over-refusal in MLLMs Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 250c52e0-c4d4-45b1-a6cf-a4817dafea05 · inbound
3D FaceShell: Attribute Transfer in 3D Face Avatars as a VLM Defense Mechanism Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6077947-21bf-4741-bc02-ae1e368cc593 · inbound
Attack Ensembles Expose a Safety-Utility Trade-off in Black-Box Guard Defenses Against Encoded VLM Jailbreaks Towards Robust Multimodal Large Language Models Against Jailbreak Attacks
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.