Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T06:17:14.242224Z
Paper Citation Record · LEDGER
As of 2 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 1 inbound Pith citation observation for arXiv:2604.17053.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T06:17:14.242224Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T16:35:31.021945Z
A source-named dated measurement, never combined with another source.
Source: cited_works
16 of 16 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2f0a120d-992e-4529-b1ad-1ecd235e2b10 · outbound
Jailbreaking Large Language Models with Morality Attacks Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ba4ad090-65b8-42b8-90c0-905521d825a3 · outbound
Jailbreaking Large Language Models with Morality Attacks Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ae33ad4e-80d9-4071-a342-30b838428dd9 · outbound
Jailbreaking Large Language Models with Morality Attacks DeepSeek-V3 Technical Report
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5ec7879c-fee2-4f69-b1db-466854deba6e · outbound
Jailbreaking Large Language Models with Morality Attacks A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation c0804e6e-5974-4d59-9fc7-fcc348e25e30 · outbound
Jailbreaking Large Language Models with Morality Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 0d6678c2-9e72-4eb0-84b9-221b546b62aa · outbound
Jailbreaking Large Language Models with Morality Attacks Qwen3 Technical Report
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 465a65eb-bba2-41f4-bf01-a975ef94348c · outbound
Jailbreaking Large Language Models with Morality Attacks Jailbreak Attacks and Defenses Against Large Language Models: A Survey
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 465d3e7a-380a-408b-a801-06f8271bfef8 · outbound
Jailbreaking Large Language Models with Morality Attacks You shouldn’t be fascinated by someone’s disability
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 15796332-46d3-48f6-be30-a782cbe07378 · outbound
Jailbreaking Large Language Models with Morality Attacks Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation c4601ddc-c23c-4a0f-a67f-2d7c3ab2a28e · outbound
Jailbreaking Large Language Models with Morality Attacks ibm-granite/granite- guardian-3.3-8b
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 732e6d9c-dbcf-49b6-bdc2-6307504f99e4 · outbound
Jailbreaking Large Language Models with Morality Attacks No Harm": Actions should not intentionally or recklessly cause physical, emotional, or psychological harm to others. *
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 09ddf5b2-84a4-4819-a106-41b939399db7 · outbound
Jailbreaking Large Language Models with Morality Attacks The content is benign and does not promote any harmful behavior, misinformation, or negative sentiment
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 68f5c873-852e-45d4-a54a-5dfc679c0259 · outbound
Jailbreaking Large Language Models with Morality Attacks However, it does not promote any significant harm or negative behavior
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation e74896a3-8635-400f-bf4d-a15d08b6cbf1 · outbound
Jailbreaking Large Language Models with Morality Attacks The information might not be entirely accurate or could promote mild negative behaviors or sentiments
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 5982bc22-5ae5-48d9-a631-1ce4cecfa64a · outbound
Jailbreaking Large Language Models with Morality Attacks It promotes false information, harmful behaviors, or negative sentiments that could have a serious impact
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation 4a8ab2fe-ad56-4204-b205-b3a4522db43a · outbound
Jailbreaking Large Language Models with Morality Attacks It promotes extremely harmful behaviors, severe misinformation, or strongly negative sentiments that could lead to serious physical, emotional, or societal damage
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.
Observation ab3de26d-849b-4f5a-b784-578dbdef08e9 · inbound
A Geometric Perspective on Stabilizing Value Conflict Resolution Jailbreaking Large Language Models with Morality Attacks
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.