Pith. sign in

Paper Citation Record · LEDGER

Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2410.09760.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.09760 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:55:13.865845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T20:58:25.971328Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a0e79c09-8bff-402d-acbd-ce8c0ffb11b1 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 97

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:25.975046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:1a111375a7a4089ffcbfa1f7d8681751e8a9e0f8e6134a9706243853ac556fc9

Observation f0ca22bc-0259-4482-acad-375721e95898 · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 120

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:33.824855Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:67a4af8176b4e51c1e08ad54ee1d75e8a9a71b79e52131245644ca5945894596

Observation e769f352-d936-40d2-8c5e-4c6bcaea9fea · inbound

Secure LLM Fine-Tuning via Safety-Aware Probing cites this paper.

Secure LLM Fine-Tuning via Safety-Aware Probing Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:11:35.842510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-22T13:07:09.402763Z digest=sha256:92f9a7c0894659ad95bb60c7b6fe30cba34e8a5406a32b12b57abce5e9bbd211

Observation d753062d-7496-4b61-aca7-e85fe104860f · inbound

SDD: Self-Degraded Defense against Malicious Fine-tuning cites this paper.

SDD: Self-Degraded Defense against Malicious Fine-tuning Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T17:55:13.865845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T17:55:13.865845Z digest=sha256:d9ff0eab223c91753ec06597fe881a57164ec804501a514711fe7918dde48264

Observation 45f09966-c506-4c95-a5c0-b781c742367d · inbound

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning cites this paper.

Token Buncher: Shielding LLMs from Harmful Reinforcement Learning Fine-Tuning Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:41:50.530339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-18T20:40:44.496392Z digest=sha256:f0227e9f4851b03ca57483fba05dd8ce299b3471d50527f5db634770680caa4b

Observation 57bbbabd-bc5b-4b21-9893-180ba15357bf · inbound

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks cites this paper.

Different Paths to Harmful Compliance: Behavioral Side Effects and Mechanistic Divergence Across LLM Jailbreaks Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T11:56:31.676636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-10T04:24:25.964495Z digest=sha256:1387330c37f7c74eae5e1b9e07fd8d8b87b717ce96a6acd79cb3989584d75296