Pith. sign in

Paper Citation Record · LEDGER

SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2402.08983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.08983 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:18:49.555908Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:59:33.707048Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation cf1f22e1-580c-4e5a-ba21-6dc19222373c · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 102

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.438218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:85eb657b8b87b62b6e725325b8674946e20bed9b22d68a020915296b82926356

Observation fa55204e-3bbd-4601-95a4-cfe519237c62 · inbound

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments cites this paper.

ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-19T01:02:54.727920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-19T01:02:07.088724Z digest=sha256:9f644da17284dbb1a632a41c8cb2d867efe03a72aeef520bb4b7e7d3130bdaf2

Observation 0b0fb6d5-9e85-465a-88d3-191905b818b9 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:49.555908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:49.555908Z digest=sha256:a524cc907beb57caa345e76bb5e492af5540af33cfbc54ee43779e8cbc2fe248

Observation 7619697b-7203-48f8-8ed1-fb63a1dada03 · inbound

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models cites this paper.

A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 267

Resolution
unresolved
no resolver link, observed 2026-08-05T10:39:08.221533Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:39:08.221533Z digest=sha256:96da6f4f4babf20220e9bef9e56ee8382ab2df568b2085ec136f572f6b53ffb0

Observation 0986a88f-38f8-4ea1-9def-a33b45ba276e · inbound

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security cites this paper.

MoGU V2: Toward a Higher Pareto Frontier Between Model Usability and Security SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T23:09:41.478401Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:09:41.478401Z digest=sha256:7a473f0a5d45c30026f9458f4e289cd93b6cd1d584a7b5e75fb156858a38fc03

Observation a5d493be-74b1-4ce1-bc43-1ed95bf91700 · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:50.904105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:40955200682884379ca0f3a2a4bef56d21d3609676f894db2eb7e94e4d26488e

Observation c920d2a6-be03-4670-8fdc-06c68e8b3dfd · inbound

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion cites this paper.

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:46:04.241295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T15:19:46.920899Z digest=sha256:dee056323360b8268618a531c3da95cfef12e3e7f602eafc5825fbd80e6f9125

Observation 4a9abf7e-acab-4d20-9778-4bb2a44b8bd7 · inbound

Jailbreaking Frontier Foundation Models Through Intention Deception cites this paper.

Jailbreaking Frontier Foundation Models Through Intention Deception SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:14.065967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T03:17:51.039062Z digest=sha256:21cd80fdbf1c3f10fb0a038c47e381fd891532d516e703d80ced67cb8b67e4c6

Observation 50551200-c250-4ecb-ab89-7297a7cf4dcc · inbound

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models cites this paper.

Minimal, Local, Causal Explanations for Jailbreak Success in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:01:09.635553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-09T20:39:39.225898Z digest=sha256:9791e906a7144d23c9ee27f6bfe5e2706e60962e38f888eb69e154907aff37c5

Observation 3efa7aa7-9924-440f-9f25-3af5375d4347 · inbound

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models cites this paper.

SafeSteer: A Decoding-level Defense Mechanism for Multimodal Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T06:57:27.549492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T06:56:10.053418Z digest=sha256:e724c2a3caa9625b77918dfd4f47bfa7886fcd2104fd8f7849731c1e3963ea7f

Observation bd96da19-f6e6-4fbf-8aad-ec088501005c · inbound

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense cites this paper.

NeuroArmor: Safe-Variant-Guided Representation Consistency for Selective Re-Anchoring in Jailbreak Defense SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:46:32.739091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T09:42:58.615078Z digest=sha256:668bde2ca530e03b135bf28f34b5ab33393291fb160f38e2856d48d169cf01d7

Observation e7372632-2dd5-4540-8213-2e1e0ac76ba5 · inbound

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks cites this paper.

SlotGCG: Exploiting the Positional Vulnerability in LLMs for Jailbreak Attacks SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 62

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T13:26:59.257417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T01:16:07.252429Z digest=sha256:b4576d118f1399972acc3b8d2547572df0c309f726d23a52190b6f2e6c8010f2

Observation a385f6c9-16d1-429e-ae3e-d9e873aa672b · inbound

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling cites this paper.

SafeSpec: Fast and Safe LLM via Dynamic Reflective Sampling SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:59:33.709134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T17:23:06.382761Z digest=sha256:c11fd77cfb30afe5c9c95315c814f0b0adb671a4c1076f39648bbfa34d254377

Observation 1596d35c-477a-491b-9715-f47be97690bd · inbound

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models cites this paper.

Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T17:35:51.363638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-29T03:35:34.594617Z digest=sha256:f8bc3b2d7ebc1aa528937a18d3f7934af7050d20cbdc57049656365200acb77e