Pith. sign in

Paper Citation Record · LEDGER

Representation Noising: A Defence Mechanism Against Harmful Finetuning

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2405.14577.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.14577 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:34:31.243967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:59:47.035398Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f4b00d41-4523-4329-ba55-459af03b4950 · inbound

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey cites this paper.

Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 128

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:58:26.090038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-23T20:58:16.237327Z digest=sha256:b70a0c75d75611a2c407b79b4b56cad7938235c27240fd89508a15a65af6ebec

Observation 94794b8a-9fe5-48aa-937b-9906fb19ad23 · inbound

Obfuscated Activations Bypass LLM Latent-Space Defenses cites this paper.

Obfuscated Activations Bypass LLM Latent-Space Defenses Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T16:59:11.680413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T16:59:11.680413Z digest=sha256:a9abc6dc971778ea786f4e751bcf38bed73114229ba8f31326e32c456c08fea5

Observation d1c44e37-5f68-4494-8618-560ad26d944d · inbound

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities cites this paper.

Model Tampering Attacks Enable More Rigorous Evaluations of LLM Capabilities Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-09T14:47:15.278636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:47:15.278636Z digest=sha256:01310b2b986130cbd9b416e53a56450eb11627263dfd2acba4e76f16f22f8273

Observation 8c0bdbcb-74c2-45ee-80d8-7e72561f0d4f · inbound

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring cites this paper.

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-08T21:06:56.026593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:06:56.026593Z digest=sha256:e06d14a907f337d52d3ce92485c61399ab304fb899163ba11535aa6a69ebf705

Observation a6dcba4f-cdc1-476c-90c6-c5f3de018471 · inbound

Jailbreaking to Jailbreak cites this paper.

Jailbreaking to Jailbreak Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T17:02:47.412426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:02:47.412426Z digest=sha256:e5356919ffd7eaac540d917399520a0dc9ce71302fd0b142ab7dc0646f0cb685

Observation bd387456-43e6-40ce-8918-f4f43f6d5350 · inbound

Layered Unlearning for Adversarial Relearning cites this paper.

Layered Unlearning for Adversarial Relearning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:34:31.243967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:34:31.243967Z digest=sha256:9c688e4fb57dc033664bbcc916d3883fe6ac272ff478b85474d08a10376f4e56

Observation 6bfb566f-c754-4fbc-86e7-f52ceb0ed5e2 · inbound

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models cites this paper.

Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:15:20.532920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:15:20.532920Z digest=sha256:46fa1dd722bb91938e76151d71f11388939f27dd259b05ba4f3711e8e5fce065

Observation e024dd50-b677-4397-9582-d93faf791fef · inbound

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning cites this paper.

Vulnerability-Aware Alignment: Mitigating Uneven Forgetting in Harmful Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:58:31.656919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:58:31.656919Z digest=sha256:46ea8b80048a9d43362f9d07c34a925135235d2d8e7e81b329c662d42c43c7c5

Observation f4a944b9-7dd5-4bb7-9f03-2070e7a327f3 · inbound

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety cites this paper.

FORTRESS: Frontier Risk Evaluation for National Security and Public Safety Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T00:15:03.070633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:15:03.070633Z digest=sha256:b4ddcb82be66975e7cff88b75f852ed8e5fe96ae9d7b75dfc76f932586108ddc

Observation b8abaa4e-bfa9-4b77-ab58-0d51ba4db639 · inbound

Probe before You Talk: Towards Black-box Defense against Backdoor Unalignment for Large Language Models cites this paper.

Probe before You Talk: Towards Black-box Defense against Backdoor Unalignment for Large Language Models Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T19:31:54.376982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:31:54.376982Z digest=sha256:257e583c6cfe33b81621edcf4c8ddd7249a786d0ef3f8cfb5494dda0ccfe78ab

Observation a6b8f4de-7b43-4ce2-b7a4-f699d257bc52 · inbound

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning cites this paper.

Learning to Stay Safe: Adaptive Regularization Against Safety Degradation during Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:37.290144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T20:51:57.399260Z digest=sha256:16e3dd10da304b8b56ac9ce0b4ec513d863184567336e813cc1d07f71b805079

Observation ef32365d-7848-40d2-85c3-344861074321 · inbound

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps cites this paper.

Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:55:48.920053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T19:30:06.396482Z digest=sha256:474dba0139fe6ede0b4afafbcafdf2cbcde22baade2bcd580349b2409c1e1761

Observation d3769af6-9555-47b9-a9e3-0d98f9e839fd · inbound

Continual Safety Alignment via Gradient-Based Sample Selection cites this paper.

Continual Safety Alignment via Gradient-Based Sample Selection Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T07:16:54.733024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T07:16:53.472918Z digest=sha256:3a8c039f4b1259bc29f6cdf3ffb82cb8a8646e3adae705eb97549c1bb58d4cf3

Observation cff556fb-b075-47fc-a9bc-31fa87ab32f3 · inbound

Safety in Self-Evolving LLM Agent Systems: Threats, Amplification, and Case Studies cites this paper.

Safety in Self-Evolving LLM Agent Systems: Threats, Amplification, and Case Studies Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:59:47.037148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T08:11:30.091859Z digest=sha256:5426c145f3704779ab8fc8259293b2e7a21d06a6aac496d77f6d5f9831a66a51

Observation 54a12f0b-55c1-45d7-a659-426c261aec60 · inbound

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning cites this paper.

Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning Representation Noising: A Defence Mechanism Against Harmful Finetuning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.792708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:46:10.792708Z digest=sha256:d134c07d4fce3fd374c5ef3f8d2c7388c2f3178e60fff6617e2cb0ac9330f8bd