Pith. sign in

Paper Citation Record · LEDGER

Don't Say No: Jailbreaking LLM by Suppressing Refusal

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2404.16369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.16369 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:28:52.855847Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T00:38:39.819593Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 366c2e06-2992-430f-869e-3a8443a2753c · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.826939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:8c893fe7f2551fcc77157bfa1e0c9dcd5c7349ab0e8defe59410d6cc84c88e07

Observation e224c9e7-a992-4dda-95ae-7a0912f68518 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.527408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:1292448200374c781d3394d08872ab219a0f43c44c796601563065f137ead5d6

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · inbound

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law cites this paper.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:90943e5f2b8e7c54a1664f207b9319cc31d2ec198e3d35f7f7a90daa2d474a56

Observation 72ba783a-5b5a-4037-a109-1035b69cdfe8 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.621417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.621417Z digest=sha256:7e82599c42c8579ae270cf8d5dba5d9ba4689a99cde0d70bafe80a25ede541fe

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:4ea71d89bf2e642918bc95907e35822354534ca54695c955256570a3aadb8ea4

Observation 36a7ccd1-900a-40af-92c3-8adda79cb4ea · inbound

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages cites this paper.

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:41.418171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:41.418171Z digest=sha256:f0aebacbe8b1617004f7d9a4dbb1cfcd6b975e6352475f409c487793db3d4c38

Observation 17b10907-dc4b-451b-b535-e0841e28f91e · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.756873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.756873Z digest=sha256:08929bd517a65b8ac1825d1bf54fc14e11f6a0bb43fc47f9f2b500c3f0a50b7e

Observation 0acbcdc5-64f0-4ef2-937b-b660adecc5c2 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.812481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.812481Z digest=sha256:ae9c3cb449b4a3ab2b752e04f8b2e3807f7285ec5ceab5d87bcf8ca79747b80a

Observation 53dee416-6ad5-4d83-a501-a5a3c659e3de · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 246

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.741705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.741705Z digest=sha256:2fabd3096df6ab99362b233c7125fc7933543524b5ffd7a174a25c14816b2856

Observation 85d5ee7b-4e79-40c2-a9b7-8a97940035a2 · inbound

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting cites this paper.

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T18:08:13.063281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T18:04:44.543311Z digest=sha256:3b9ceb37d81034d96683a1da1010d8294b2749d05765274dd19da54da22c1599

Observation a46b9f70-7727-424c-9626-25a7dfb622f3 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:03.956567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:40b1de54446a3b0642f878778d3bec2b2ee2c217de9eeab25cd9950e88874c09

Observation ed39c6ea-b358-4627-9245-5009824ddd61 · inbound

Adversarial Reframing: A Framework for Targeted Generation in Language Models cites this paper.

Adversarial Reframing: A Framework for Targeted Generation in Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:36:20.999209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:35:51.862736Z digest=sha256:fcc744d11147abfdf4a5e54147ba1045a9982dfa3d2456eac8207e280d1cc198