Pith. sign in

Paper Citation Record · LEDGER

Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2402.14968.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.14968 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:31:54.428284Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:58:26.356968Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bee83ff3-2209-43db-977e-908750aa1f22 · inbound

Jailbreak Attacks and Defenses Against Large Language Models: A Survey cites this paper.

Jailbreak Attacks and Defenses Against Large Language Models: A Survey Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 94

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:20:44.854675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T02:20:44.368219Z digest=sha256:bd29abb4267e18069ae83554c22f0b78b36864dff2b1572e4ba1e141bb4550cf

Observation 4a5adaac-d38b-41c5-8cec-76ee3c8bba3f · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 157

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.360026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:ad0145a515d7654e4fee6a083e93c615df706a40b5d6359b0b0d2cdd1361f2eb

Observation 9220a8f9-c9e5-46fc-8905-9519f83138e7 · inbound

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning cites this paper.

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:56.867587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:56.867587Z digest=sha256:d1f042d724119ef84a0b5a2f6d27df5d52a498e3b31c36c9edec5d715bbfaf2b

Observation 889c4057-f3ed-4ae3-9b97-1ffb190113d4 · inbound

Probe before You Talk: Towards Black-box Defense against Backdoor Unalignment for Large Language Models cites this paper.

Probe before You Talk: Towards Black-box Defense against Backdoor Unalignment for Large Language Models Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T19:31:54.428284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:31:54.428284Z digest=sha256:bc8159ac3800a5cedceffa4a423e4139e8e21d0e01633bb49036b3a495814573

Observation 37236ef3-5323-487e-b07e-d1321711ef35 · inbound

Safe Pruning LoRA: Robust Distance-Guided Pruning for Safety Alignment in Adaptation of LLMs cites this paper.

Safe Pruning LoRA: Robust Distance-Guided Pruning for Safety Alignment in Adaptation of LLMs Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T19:09:30.208410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:09:30.208410Z digest=sha256:66ddb010a33b7745ecf8f5d3caebc18904bb1594f7e7fcfa324b0a05f14e4247

Observation fcd0cd75-7cfe-43d2-a17c-c184cc4acc98 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment

Reference 184

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:56.274796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:56.274796Z digest=sha256:17295fa6bc113107b1621b7d9c28413d1d13b6bc9ada86f89fc6e7b1ce3785c5