Pith. sign in

Paper Citation Record · LEDGER

Don't Say No: Jailbreaking LLM by Suppressing Refusal

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2404.16369.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.16369 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:10:31.116961Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T00:38:39.819593Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 366c2e06-2992-430f-869e-3a8443a2753c · inbound

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment cites this paper.

Uncovering Logit Suppression Vulnerabilities in LLM Safety Alignment Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:38:39.826939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-24T00:38:36.992597Z digest=sha256:45a05d6a05b08e2d956fa3d6b0ece7db15030fb99885f1b882b3ebb73ed75882

Observation e224c9e7-a992-4dda-95ae-7a0912f68518 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:20:44.527408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:6354a6bc970bc46f3d577aac96e8b1a956d718ce5a90bee2b9eec330fb36af78

Observation be30bdc0-80c3-4b36-8f45-0d7801a65a84 · inbound

Steering Language Model Refusal with Sparse Autoencoders cites this paper.

Steering Language Model Refusal with Sparse Autoencoders Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-12T18:45:48.618496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:45:48.618496Z digest=sha256:81be8a287d0143f5e580081b09fbb51aa2f682298ab4ae14db03a830757b41ce

Observation d3983639-afd5-4b66-b54b-f4fef6810f32 · inbound

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment cites this paper.

Immune: Improving Safety Against Jailbreaks in Multi-modal LLMs via Inference-Time Alignment Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-12T11:03:01.302592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T11:03:01.302592Z digest=sha256:998e08a316c3d36d0afd4393a660d66a5ffeb4cd7247634609938acfdd2c726f

Observation 06844bb0-0c0a-4e29-bc32-81682c8ded31 · inbound

LIAR: Leveraging Inference Time Alignment (Best-of-N) to Jailbreak LLMs in Seconds cites this paper.

LIAR: Leveraging Inference Time Alignment (Best-of-N) to Jailbreak LLMs in Seconds Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T20:53:38.180384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:53:38.180384Z digest=sha256:60db37e09fde78ce71c31b8e6f48c054cb552b18be2392430b926b5762887cc4

Observation b70321f9-361d-440e-bc8e-aa948d4f5c52 · inbound

SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector cites this paper.

SAIF: A Comprehensive Framework for Evaluating the Risks of Generative AI in the Public Sector Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T20:18:54.655981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:18:54.655981Z digest=sha256:ccd9076b6de2eaa4b3e55c09f3a933f36a58ce682f8a21f6d82344bd766c7fef

Observation 0a81a26c-05cf-4e47-ab71-dfbc0f7b0720 · inbound

DETAM: Defending LLMs Against Jailbreak Attacks via Targeted Attention Modification cites this paper.

DETAM: Defending LLMs Against Jailbreak Attacks via Targeted Attention Modification Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T12:10:31.116961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:10:31.116961Z digest=sha256:b07ddd814a35cbfd4adfb9becd5fe40907cc2a7c32e070163e99189598cbea6c

Observation 733aab27-5d09-4e3f-8a73-4d1720e9ffc3 · inbound

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models cites this paper.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.883975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.883975Z digest=sha256:9b8abe870dc45690d7d488b9b5625c6d6ff04a955142e3ebafb1681f9d7e7e8b

Observation 34896cab-3e45-46f0-b59a-f53c3fcb04d0 · inbound

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law cites this paper.

From Rogue to Safe AI: The Role of Explicit Refusals in Aligning LLMs with International Humanitarian Law Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:52.855847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:52.855847Z digest=sha256:ba637b5c500e75313e0365f154db5fcb641a8fc6024fbe96b1c2e883d0c42754

Observation 72ba783a-5b5a-4037-a109-1035b69cdfe8 · inbound

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures cites this paper.

Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:40:20.621417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:40:20.621417Z digest=sha256:da55ef1551cf312a95e049c7ea95ed5c96bd91184979a789bc637d2c4828f85a

Observation 42a48662-4628-4735-83ee-9602c01ffcd3 · inbound

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models cites this paper.

SoK: A Comprehensive Security Analysis of Jailbreak Resilience in GPT and DeepSeek Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T23:20:57.656839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:20:57.656839Z digest=sha256:b2bfcbadf8b8fc0e60656e11f22bbd63af487cab355529c1ca46e0ad43988f84

Observation 36a7ccd1-900a-40af-92c3-8adda79cb4ea · inbound

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages cites this paper.

Leveraging the Potential of Prompt Engineering for Hate Speech Detection in Low-Resource Languages Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T21:34:41.418171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:34:41.418171Z digest=sha256:4f440c1272d925c448710980a9b4988de8caeae2f7e07e03726894d371cf49d5

Observation 17b10907-dc4b-451b-b535-e0841e28f91e · inbound

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation cites this paper.

Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:33.756873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:33.756873Z digest=sha256:67ce628ddd5b2ab9bb9e6a4f4e9a24cd4ae6ece177612e8621df9d82d7ba9f8f

Observation 0acbcdc5-64f0-4ef2-937b-b660adecc5c2 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.812481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.812481Z digest=sha256:538fc40594dbdd0a8d2cfc4ff65731a863aa51ccb741dd9da84bf794b404e86b

Observation 53dee416-6ad5-4d83-a501-a5a3c659e3de · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 246

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:58.741705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:58.741705Z digest=sha256:2f7d5cbc33f2d28a7cac163da899858e89b45d3f415229f2146afad10075697e

Observation 85d5ee7b-4e79-40c2-a9b7-8a97940035a2 · inbound

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting cites this paper.

Break Me If You Can: Self-Jailbreaking of Aligned LLMs via Lexical Insertion Prompting Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T18:08:13.063281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T18:04:44.543311Z digest=sha256:9bf4af2ba24e2932779b299a8ff0a20a14886b1a5c8fdcc78b555f5a6cbfed9a

Observation a46b9f70-7727-424c-9626-25a7dfb622f3 · inbound

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems cites this paper.

The Salami Slicing Threat: Exploiting Cumulative Risks in LLM Systems Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:03.956567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T16:07:31.602378Z digest=sha256:d37c18e94caab5bc91c9e78c17d8b202cc53aafc982b9d240c583d83e38ee7ff

Observation ed39c6ea-b358-4627-9245-5009824ddd61 · inbound

Adversarial Reframing: A Framework for Targeted Generation in Language Models cites this paper.

Adversarial Reframing: A Framework for Targeted Generation in Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:36:20.999209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-22T09:35:51.862736Z digest=sha256:41b21c6767fba50595c0405bcb4e4bb59325632af672fe95cf8c508319a3b384