Pith. sign in

Paper Citation Record · LEDGER

Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.09760.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.09760 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:55:13.865845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:58:25.971328Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a0e79c09-8bff-402d-acbd-ce8c0ffb11b1 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:25.975046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:8f2cd7f487924822a187fef25df6a81b7fce42746546303f41d705c1e8e27bad

Observation f0ca22bc-0259-4482-acad-375721e95898 · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 120

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:33.824855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:0e8828e6d1ecebccf05e376e1c336ed21a5354bd19ae04ba077a5928b42c1ded

Observation e769f352-d936-40d2-8c5e-4c6bcaea9fea · inbound

Secure LLM Fine-Tuning via Safety-Aware Probing cites this paper.

Secure LLM Fine-Tuning via Safety-Aware Probing Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:11:35.842510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-22T13:07:09.402763Z digest=sha256:263b99bd20a869d7e2ef9517752bf0b2051903fb056e991bef615a2fa4926f02

Observation d753062d-7496-4b61-aca7-e85fe104860f · inbound

SDD: Self-Degraded Defense against Malicious Fine-tuning cites this paper.

SDD: Self-Degraded Defense against Malicious Fine-tuning Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T17:55:13.865845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:55:13.865845Z digest=sha256:865a7182c16c987b8ae9aa5ed3b848d7fc10a6732dc80fc5b94bf5b0aed3fb70

Observation 45f09966-c506-4c95-a5c0-b781c742367d · inbound

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning cites this paper.

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:41:50.530339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-18T20:40:44.496392Z digest=sha256:ceadf86ed6d112f79433b6bad27a678a7b00260e9c37061d63cee4e526c7c47e

Observation 57bbbabd-bc5b-4b21-9893-180ba15357bf · inbound

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks cites this paper.

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:56:31.676636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-05-10T04:24:25.964495Z digest=sha256:22c23e4387eef972967059e39bce9b50cf70f9522b5908c3ec62dd767ca27b90