Pith. sign in

Paper Citation Record · LEDGER

Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2404.05880.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.05880 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:32:50.304538Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T22:26:18.073916Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2b14b808-5ddc-4f78-bcd0-a5c8461a1f6f · inbound

SEPS: A Separability Measure for Robust Unlearning in LLMs cites this paper.

SEPS: A Separability Measure for Robust Unlearning in LLMs Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:32:50.304538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:32:50.304538Z digest=sha256:3bd6284d463c2b2c5dd829063ef190840b60d3fa2e3103d935cdf2b12babb7a3

Observation 61797ac5-a8bb-4bb9-a17d-ea10b8b0ed05 · inbound

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge cites this paper.

Automating Evaluation of Diffusion Model Unlearning with (Vision-) Language Model World Knowledge Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:38.651564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:38.651564Z digest=sha256:fac787dcc0a20885c370acacf7545d656ba134e66c0385307e25e7caff05e071

Observation 452431f8-91b4-4951-8e0e-7add4f59f70e · inbound

A Survey on Generative Model Unlearning: Fundamentals, Taxonomy, Evaluation, and Future Direction cites this paper.

A Survey on Generative Model Unlearning: Fundamentals, Taxonomy, Evaluation, and Future Direction Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 153

Resolution
unresolved
no resolver link, observed 2026-08-06T13:54:40.043778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:54:40.043778Z digest=sha256:3d038ac270ae5b83446b76bca1eecda301f6153f102b41a5b026f4325b2fa061

Observation 21975885-5cd2-41c3-b0a7-6523cf50f573 · inbound

SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks cites this paper.

SafeLLM: Unlearning Harmful Outputs from Large Language Models against Jailbreak Attacks Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T18:05:32.075408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:05:32.075408Z digest=sha256:e3fe89916a99fd1701763948962c3c8818a181138e63a8c20e04ce4a34bdb542

Observation e75cb372-fbec-4995-bb46-54f0ab7f6f36 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:49.054581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:49.054581Z digest=sha256:8debe39637bbe742c1ed5f807d706c13de64f6fa102c2ba87dfe347e785693ed

Observation 8a472273-706c-43bb-a14f-55dae21abc06 · inbound

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting cites this paper.

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-18T06:51:00.698323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T06:48:49.492536Z digest=sha256:a8a16fc291ce1c088e83e77131671f05edbe118d3d98bb5f35ac99556b6e6efb

Observation 5460ab10-c69b-4656-8582-cead1e11ca9c · inbound

Exclusive Unlearning cites this paper.

Exclusive Unlearning Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:10:52.142037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:39:54.728106Z digest=sha256:56f27ff543e80797fd66ca5cfee0a72dd82968dabaaf090456360a9843793485

Observation c569a026-1139-43f3-8c96-e80baa2511a7 · inbound

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense cites this paper.

TrajGuard: Streaming Hidden-state Trajectory Detection for Decoding-time Jailbreak Defense Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:35:50.996117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T18:26:19.922383Z digest=sha256:8f217cf39d5604d830a8ca2963a507c36abcffc615e3ad6d90baa559b8d0f688

Observation 4ae492ea-516a-471d-b069-ef0fb166f5df · inbound

Jailbreaking Frontier Foundation Models Through Intention Deception cites this paper.

Jailbreaking Frontier Foundation Models Through Intention Deception Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:11:13.887943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T03:17:51.039062Z digest=sha256:e57646c2322c230d666ace350c7d7ad7f15e97441f1a089e2d9c29e2d5004c20

Observation ace9d281-c844-48ec-80de-49b4352ca8ff · inbound

Fast Unlearning at Scale via Margin Self-Correction cites this paper.

Fast Unlearning at Scale via Margin Self-Correction Eraser: Jailbreaking Defense in Large Language Models via Unlearning Harmful Knowledge

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:26:18.075413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T15:21:26.391287Z digest=sha256:c13bc21ffba8f14a9565e96d2b04cd997c841a881b36466802f8cb258aba12d5