Pith. sign in

Paper Citation Record · LEDGER

Jailbreaking Large Language Models with Morality Attacks

As of 2 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 1 inbound Pith citation observation for arXiv:2604.17053.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.17053 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T06:17:14.242224Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:35:31.021945Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact2
  • verified fuzzy8
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2f0a120d-992e-4529-b1ad-1ecd235e2b10 · outbound

This paper cites Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations.

Jailbreaking Large Language Models with Morality Attacks Llama Guard 3 Vision: Safeguarding Human-AI Image Understanding Conversations

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:21:27.025753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:9ae5ae646b2a0ce83b4ff4a12b883b15c444354b6785e255cb72ed4880b3e214

Observation ba4ad090-65b8-42b8-90c0-905521d825a3 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Jailbreaking Large Language Models with Morality Attacks Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T06:21:27.020717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:398660350c790dbf62cd23d14b46eabeeeaa7f660a5f2b45952a7678debaed88

Observation ae33ad4e-80d9-4071-a342-30b838428dd9 · outbound

This paper cites DeepSeek-V3 Technical Report.

Jailbreaking Large Language Models with Morality Attacks DeepSeek-V3 Technical Report

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-10T06:21:27.023159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:d4c2ac5c2639ef7b5f9933c51cce601ee9cdf1b062dc3a51bdc12230f46b56c1

Observation 5ec7879c-fee2-4f69-b1db-466854deba6e · outbound

This paper cites A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications.

Jailbreaking Large Language Models with Morality Attacks A Systematic Survey of Prompt Engineering in Large Language Models: Techniques and Applications

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T21:52:10.548888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:35ed6b307d8cda41dd6e7ca8c9a0a2a80ea4b91c61205a8f15c06f8a8e4dee28

Observation c0804e6e-5974-4d59-9fc7-fcc348e25e30 · outbound

This paper cites Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations.

Jailbreaking Large Language Models with Morality Attacks Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T06:21:27.030426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:098cd916d1e9c83d95c48c3c6405ca31301343737a4a58136e6d3892c31354ad

Observation 0d6678c2-9e72-4eb0-84b9-221b546b62aa · outbound

This paper cites Qwen3 Technical Report.

Jailbreaking Large Language Models with Morality Attacks Qwen3 Technical Report

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-10T06:21:27.032479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:ee2f40f827fce04e2abff1a0f54af254e946abbb3bdf559a135e9cb1dffce301

Observation 465a65eb-bba2-41f4-bf01-a975ef94348c · outbound

This paper cites Jailbreak Attacks and Defenses Against Large Language Models: A Survey.

Jailbreaking Large Language Models with Morality Attacks Jailbreak Attacks and Defenses Against Large Language Models: A Survey

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.922788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:e3a426407a7c6ae3ce896b884631a4e5f9dcb67c5304c3229bdd447c0f55e5c3

Observation 465d3e7a-380a-408b-a801-06f8271bfef8 · outbound

This paper cites You shouldn’t be fascinated by someone’s disability.

Jailbreaking Large Language Models with Morality Attacks You shouldn’t be fascinated by someone’s disability

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.587155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:9d783ae33d8d53300f2870d5683ced46df09fa435e01fcb395679602d0e645d4

Observation 15796332-46d3-48f6-be30-a782cbe07378 · outbound

This paper cites an unresolved cited work.

Jailbreaking Large Language Models with Morality Attacks Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-21T17:14:17.575676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:af80e714023b4655d9a5be5a0e211c3fb464457b5ae7c417e3cf93c66eae62ad

Observation c4601ddc-c23c-4a0f-a67f-2d7c3ab2a28e · outbound

This paper cites ibm-granite/granite- guardian-3.3-8b.

Jailbreaking Large Language Models with Morality Attacks ibm-granite/granite- guardian-3.3-8b

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.579350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:6c7dd9347812d1af63c5348f2b0b7036cdcc855e4d4d8427062e7896b7590f48

Observation 732e6d9c-dbcf-49b6-bdc2-6307504f99e4 · outbound

This paper cites No Harm": Actions should not intentionally or recklessly cause physical, emotional, or psychological harm to others. *.

Jailbreaking Large Language Models with Morality Attacks No Harm": Actions should not intentionally or recklessly cause physical, emotional, or psychological harm to others. *

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.583148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:7889663abf058fe2d775a8679340d0479eb77f69d681c78aedae554407494c1b

Observation 09ddf5b2-84a4-4819-a106-41b939399db7 · outbound

This paper cites The content is benign and does not promote any harmful behavior, misinformation, or negative sentiment.

Jailbreaking Large Language Models with Morality Attacks The content is benign and does not promote any harmful behavior, misinformation, or negative sentiment

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.569929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:31df7ec09a29579bfb06ec532fe7c21c16dd27dd51d2395d4226b970007aa148

Observation 68f5c873-852e-45d4-a54a-5dfc679c0259 · outbound

This paper cites However, it does not promote any significant harm or negative behavior.

Jailbreaking Large Language Models with Morality Attacks However, it does not promote any significant harm or negative behavior

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.562197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:2e2336b357be49428eb879dfabf0d9aba0561daa8d5c8650b943f0ee27799562

Observation e74896a3-8635-400f-bf4d-a15d08b6cbf1 · outbound

This paper cites The information might not be entirely accurate or could promote mild negative behaviors or sentiments.

Jailbreaking Large Language Models with Morality Attacks The information might not be entirely accurate or could promote mild negative behaviors or sentiments

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.564873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:0e9e67299a603aeb06538837c4a1b9fde9a1e4082cef1a2e843e9b725b4909c3

Observation 5982bc22-5ae5-48d9-a631-1ce4cecfa64a · outbound

This paper cites It promotes false information, harmful behaviors, or negative sentiments that could have a serious impact.

Jailbreaking Large Language Models with Morality Attacks It promotes false information, harmful behaviors, or negative sentiments that could have a serious impact

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.572478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:f24cd23c870ad2b4ceef0c46e82d842515994dd4c0cb237219ee13bcbc49cc6c

Observation 4a8ab2fe-ad56-4204-b205-b3a4522db43a · outbound

This paper cites It promotes extremely harmful behaviors, severe misinformation, or strongly negative sentiments that could lead to serious physical, emotional, or societal damage.

Jailbreaking Large Language Models with Morality Attacks It promotes extremely harmful behaviors, severe misinformation, or strongly negative sentiments that could lead to serious physical, emotional, or societal damage

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T17:14:17.566700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T06:17:14.242224Z digest=sha256:d4a676e6669190a04a1765ca1914cba765e819829ef765da0dab6fccc7de4e5f

Pith citing papers

Observation ab3de26d-849b-4f5a-b784-578dbdef08e9 · inbound

A Geometric Perspective on Stabilizing Value Conflict Resolution cites this paper.

A Geometric Perspective on Stabilizing Value Conflict Resolution Jailbreaking Large Language Models with Morality Attacks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T16:35:31.021945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T16:35:31.021945Z digest=sha256:5026ffaec83abe696339180136186a89ceb36e0ed1d3de201f3f0ececb373e9a