Pith. sign in

Paper Citation Record · LEDGER

Don't Say No: Jailbreaking LLM by Suppressing Refusal

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.16369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.16369 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:52.855847Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T00:38:39.819593Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 366c2e06-2992-430f-869e-3a8443a2753c · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.826939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:6ac89fd5a8170a9b9491bf74c7214255c437e533a2e6452c373b8237ba32aca2

Observation e224c9e7-a992-4dda-95ae-7a0912f68518 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.527408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:f37b3ad1fbb0fedbf4e8e19f28177ace80572ba65adaf45ed5d933fca7cb1d9a

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · inbound

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law cites this paper.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:2218fe749d088e4e7aa8dec9180ac74c58205b9ddcb5903ac66414c9284fd903

Observation 72ba783a-5b5a-4037-a109-1035b69cdfe8 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.621417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.621417Z digest=sha256:779ca27fd15ebb925f304d2ff0627cc3db0bf9fb1e4b75554dc8e55876b1a096

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:292038772e354f014b6fb5def2b9987c8440264ec79a7c14b85e49835681f94a

Observation 36a7ccd1-900a-40af-92c3-8adda79cb4ea · inbound

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages cites this paper.

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:41.418171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:41.418171Z digest=sha256:c0a0cab8245f50f0b48f54228f467bd29744e22c6f9240439e47c1d475f44c3c

Observation 17b10907-dc4b-451b-b535-e0841e28f91e · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.756873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.756873Z digest=sha256:f55d8dcc2b8da0b612f19fc788910df764aa15c7117a30883f156e746ff06604

Observation 0acbcdc5-64f0-4ef2-937b-b660adecc5c2 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.812481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.812481Z digest=sha256:cbf4f87b5f151f1ce9f8b5c49717a7efe17aedfaf8457fa87c6b6ec9b7c41c97

Observation 53dee416-6ad5-4d83-a501-a5a3c659e3de · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 246

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.741705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.741705Z digest=sha256:d9f1ec02d037821aa600b78b84de9be9303c5a36528ded8c808e3576d4bae342

Observation 85d5ee7b-4e79-40c2-a9b7-8a97940035a2 · inbound

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting cites this paper.

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T18:08:13.063281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T18:04:44.543311Z digest=sha256:a4ec22736e8e47f643d3a370c4472e9ef7c988e503d1c5888c26400509722a93

Observation a46b9f70-7727-424c-9626-25a7dfb622f3 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:03.956567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:00936cf7b91ce9cae31b6ad5d9b185b67e0a657a9d71c8daa4daf68582bb316f

Observation ed39c6ea-b358-4627-9245-5009824ddd61 · inbound

Adversarial Reframing: A Framework for Targeted Generation in Language Models cites this paper.

Adversarial Reframing: A Framework for Targeted Generation in Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:36:20.999209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:35:51.862736Z digest=sha256:2dd4d6b0e09b7e6c9ee8a0dd46727d3bde615aea6c91666fa16ece32d36253d8