Pith. sign in

Paper Citation Record · LEDGER

Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2408.09600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.09600 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:15:18.856097Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:47:22.918490Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eeb9954b-7eba-4d31-904c-cca1ddc6ab21 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:25.832290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:2f6148fd0aaa2abc3445b7dbc921372291993a79eaf78c5ed8284db7e07cd9d2

Observation 36e46723-4c7c-4503-b2a8-e02bd740e5ec · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 123

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.050386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:5c18aedc5c02d7a45db0fef87e3f838bda8406aa7bf55f01f8c37bc4430c0952

Observation 5ed09cab-fdb5-4a0a-a16f-8cb48dd54f84 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:18.856097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:18.856097Z digest=sha256:bc4b2f32fdd58df6bcf88815f3134a9163ee4a916227f91e3a37e06ed754eda6

Observation 5ce7a588-a395-4725-802b-23fa97cb6001 · inbound

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning cites this paper.

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:58:29.556438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:58:29.556438Z digest=sha256:2c615c803738e002859c5a1959156ed03618bdd16f4300efbe126aa45c5ed872

Observation fa4eff3d-7ade-413d-a974-1b00d9d22bc5 · inbound

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning cites this paper.

LoX: Low-Rank Extrapolation Robustifies LLM Safety Against Fine-tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:59:56.106685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:59:56.106685Z digest=sha256:9d14f3acf8380c47cfc4c8a85bdba4d7a89376df318097e548e806af28e78534

Observation 497dbac8-a173-476f-b9d7-893127262d88 · inbound

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM cites this paper.

Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 195

Resolution
unresolved
no resolver link, observed 2026-08-05T23:13:05.149386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:13:05.149386Z digest=sha256:6f45cf65ba975f4e56a35be929239454abb1842261e1c312cb466aef8ba112b2

Observation 27f32768-a236-43d8-87c2-5e2116809be0 · inbound

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering cites this paper.

TamperBench: Systematically Stress-Testing LLM Safety Under Fine-Tuning and Tampering Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T03:46:14.818615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:46:14.818615Z digest=sha256:789d7dfc88bba59ef27cf006cc2ed03322964a605cead2860cef36ba5b5ceb03

Observation 0c3566bd-0122-4ae0-8728-8856981c14fa · inbound

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints cites this paper.

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:01.762843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T16:04:25.851592Z digest=sha256:ad5389dd9a4d83aeae5f2ff57d70ba879094f9e5f73b8a16121193012b2d7ef4

Observation c2d39d3c-2be1-4d64-98b2-42a633e1f123 · inbound

Generating Place-Based Compromises Between Two Points of View cites this paper.

Generating Place-Based Compromises Between Two Points of View Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:01:12.267420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T03:36:31.695964Z digest=sha256:1cc9270866f8e8a462759071907b38a414e6569ce085e5da85c0d0bded7ac060

Observation 1c0d34fb-72c6-4a6e-8509-cc129c2e3f80 · inbound

GradShield: Alignment Preserving Finetuning cites this paper.

GradShield: Alignment Preserving Finetuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T04:45:00.959333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T04:44:06.614390Z digest=sha256:bebdb0fa3721cc0148ab9e3bccbe59b08804c16c4b699fb9a0f3b66bb236329b

Observation 4eaea4bb-852b-4bed-a106-65420c0fad1f · inbound

SafeGene: Reusable Adapters for Transferable Safety Alignment cites this paper.

SafeGene: Reusable Adapters for Transferable Safety Alignment Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:36:29.093869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T09:58:16.956281Z digest=sha256:3764b7665f4da44f77f735cc0ae79de90b1d24fc9b830dc5810d5575af61bc32

Observation a746fa85-63be-4e64-b540-0ad1733a8c34 · inbound

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks cites this paper.

Defending Against Malicious Finetuning by Scaling Train-time Adversarial Attacks Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:22.920004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T20:10:25.375603Z digest=sha256:d132c54cfefcf34f78b3b4e29bdac1523145b33f931f488a81c836e2caf9ad36