Pith. sign in

Paper Citation Record · LEDGER

Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation

As of 13 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2501.17433.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.17433 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:26:13.652637Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:42:34.275997Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 40fb8e02-5ed2-48e4-9805-054843b93c9b · inbound

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety cites this paper.

Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:42:34.279284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-23T04:39:04.591722Z digest=sha256:0ab139386a83750ef52aae5fdd7ef8bb402761357b3275c3b24276892899c2aa

Observation f5ba8e02-9b42-4f11-a9f4-f1c4444f7204 · inbound

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation cites this paper.

Circumventing Safety Alignment in Large Language Models Through Embedding Space Toxicity Attenuation Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T19:26:13.652637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:26:13.652637Z digest=sha256:16fd9bd737201c47d9b87d7d4000cca4d9074070a860df6e5332682f8174d108

Observation 03448501-1d50-4b16-937a-8187e583662d · inbound

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems cites this paper.

LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation

Reference 131

Resolution
unresolved
no resolver link, observed 2026-08-04T17:46:19.330250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:46:19.330250Z digest=sha256:17b339bfed289b7e9ebb5c79540a0cc159409956a48fd56e1cb62fe1b174d4ff

Observation b0e32cd4-4516-4fef-a8a4-cb5c2fc4ca9c · inbound

SelfGrader: LLM Jailbreak Detection via Anchored Token-Level Logits cites this paper.

SelfGrader: LLM Jailbreak Detection via Anchored Token-Level Logits Virus: Harmful Fine-tuning Attack for Large Language Models Bypassing Guardrail Moderation

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:48:19.366394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-13T21:43:46.728512Z digest=sha256:5cf5779d9435f2127b7f44e62cffb1d82b297d889c9233a58e9315718e60ee6f