Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:27.712204Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2507.04673.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:46:27.712204Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cbe1ae7c-e1cd-4754-b783-12ba81f019ae · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68b78950-31ea-4ae7-a11a-c446900e07a2 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02c6bfce-8387-4b88-961f-51cdbc6e695a · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Fuzz-Testing Meets LLM-Based Agents: An Automated and Efficient Framework for Jailbreaking Text-To-Image Generation Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42e7b7f6-a4ff-454f-98b2-b7b0f6a236fd · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message & Lowe, R
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 39ef1738-b7ea-4156-ba56-a47d600122a4 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Reasoning-Augmented Conversation for Multi-Turn Jailbreak Attacks on Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a68bd5e6-34fd-4093-8a56-43f078dbbba9 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Red Teaming Language Models with Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6433a5ab-e4ec-4c88-93f2-acf3754ec031 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d0799c7-1f3f-4ff7-b50c-591d077226c6 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fd45238-5c2f-435e-842e-0f93b792de74 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c72c4f5c-94a3-4041-b26b-747ee4884a48 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 129cc2a2-71a6-4f6e-9f65-da89169c3590 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Multimodal Pragmatic Jailbreak on Text-to-image Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86029344-f94f-4cc8-958d-392e967e577e · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b2c57e0-11c6-4d76-903a-19dec5ccff73 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message F., Leike, J., Brown, T., Martic, M., Legg, S., & Amodei, D
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0d2a6696-a20a-4316-9611-5342157ccb23 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Jailbroken: How Does LLM Safety Training Fail?
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6418170-4cf4-49d2-b7f2-4d1e2b9b094d · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message ArtPrompt: ASCII Art-based Jailbreak Attacks against Aligned LLMs
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6982bd-78ac-4b73-aa55-315d166a41e4 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Chain-of-Jailbreak Attack for Image Generation Models via Editing Step by Step
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e4dedf08-ce85-4eb0-92da-562fef96fbe0 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5128f266-0187-4ed2-974e-3b94f5a72697 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message GenBreak: Red Teaming Text-to-Image Generators Using Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a5cbd56-88a6-4c88-8df7-b6e574ac8c03 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Jailbreaking Multimodal Large Language Models via Shuffle Inconsistency
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6b983d6-49ab-4cdd-967c-1d567a4bd97d · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message HTS-Attack: Heuristic Token Search for Jailbreaking Text-to-Image Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e668489-c586-4ce6-84b2-c3e279662485 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message SneakyPrompt: Jailbreaking Text-to-image Generative Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 634b7aab-52c0-4de4-91af-93bce1a07b10 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2067cff-6d94-47f2-8e03-bcdfb55f1ec1 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Gradient-based Jailbreak Images for Multimodal Fusion Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a6c5220-9606-4a96-be45-78998596fd50 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4e24e59-c187-4a25-b386-c1458f598470 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Red-Teaming LLM Multi-Agent Systems via Communication Attacks
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed12d4c6-a0e6-4378-b560-30a9911f5ae0 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message GPT-4 Technical Report
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e0cb174-ef09-44f7-a443-6fab71e75112 · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Consideration Set Sampling to Analyze Undecided Respondents
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0cf1581b-227e-41b4-9945-b6c7ee04bc4a · outbound
Trojan Horse Prompting: Jailbreaking Conversational Multimodal Models by Forging Assistant Message Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.