Pith. sign in

Paper Citation Record · LEDGER

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models

As of 22 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 3 inbound Pith citation observations for arXiv:2504.21053.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.21053 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:33:11.895798Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T03:57:57.052252Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T22:31:19.414729Z

Reference resolution

30 of 30 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42c2dc98-afb5-43fe-9584-7d52fa3079b1 · outbound

This paper cites URL: " 'urlintro :=.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.716810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.716810Z digest=sha256:b3a189cc430e0ff84afb4c67612fe4eb04a22dbb7c680d8b9c7d424bb234fb0f

Observation 076c744c-46cf-46b0-9744-5fcfc66b1c5a · outbound

This paper cites write newline.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.724386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.724386Z digest=sha256:898bd0c3b92ccef59367fb9c1c13c094061c665f655dee39f42dbee71a83f5ae

Observation 347cc208-4f15-4768-8b24-90b690d60797 · outbound

This paper cites Jailbreaking Black Box Large Language Models in Twenty Queries.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Jailbreaking Black Box Large Language Models in Twenty Queries

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.730582Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.730582Z digest=sha256:75902203f71a5255f7a0b3b7748811881925ddeda3c8417af0f943e667384f80

Observation a2fe4983-8c5e-4a8c-a953-c5aadff40d70 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:33:12.476716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-16T05:33:11.737825Z digest=sha256:c7680750ef63f3bcd627405361a0c65e4afe2533642764848d30bb2b0567cf18

Observation e6f07c59-ba5c-4dc5-b3f2-57d72bf1a01c · outbound

This paper cites Chiang, Z.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Chiang, Z

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:33:12.459848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-16T05:33:11.743950Z digest=sha256:e766e8ae8199d0e5279c91506aad91e2f58e3cb6062d4ce75a32d18ea18d658c

Observation 48eb127e-111a-43bd-827c-0f1b9521e0b2 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.749823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.749823Z digest=sha256:d398bb4a1a14d42127fa83f8ec86d141de472b980927e585a44fa31c4d6b2bee

Observation bcb31f5c-0917-4cad-a08e-5fbb85964472 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.756846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.756846Z digest=sha256:e4a14b1e42a98b2c6e22f378d1e2752c08dfeed1dd512baa3699383015c3a343

Observation fa6cb716-eb98-4679-b867-6b9074b7c1fb · outbound

This paper cites COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models COLD-Attack: Jailbreaking LLMs with Stealthiness and Controllability

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.762587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.762587Z digest=sha256:056470f382c6286e81526c72c05a3aa0e0086a22154d3e04a5af9940082d421c

Observation 69031f02-918a-410f-85c8-01ce3f08d1d6 · outbound

This paper cites Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.776085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.776085Z digest=sha256:82f0d0f4ce91cd20bfd7e1a79e7b194b8ec81e121fb06958ebd803f784d709e8

Observation 79a6ef4f-f2dc-4c03-9743-c00caac43d70 · outbound

This paper cites Mistral 7B.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Mistral 7B

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.782192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.782192Z digest=sha256:095952e58d457d2b6951576984c4d62a95845c5782cd9a1f5511bf8f4639ed51

Observation e5ada9f8-ede8-478f-a75f-aeef5ccdc6a5 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.787854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.787854Z digest=sha256:f31fee0693937bf1c5d5537cf059188360a094cd6f568384f66be43b1014c5ba

Observation 7ac35d16-37d2-4ea5-8302-3b79771a2378 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.792894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.792894Z digest=sha256:6587f4efb7306d8ba1df07fb22ddb40ad67a944d04d155afccbe5ea7e14d6a84

Observation a64fafb6-1742-41cd-abc8-cd146e66c4dd · outbound

This paper cites LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.797685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.797685Z digest=sha256:639c8de3298ef1fe34b71890dee05aa1fb0daa0ce6abbca560a90efe80296111

Observation be1b1c36-5fe8-4511-bd4c-667aa5f72602 · outbound

This paper cites DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models DrAttack: Prompt Decomposition and Reconstruction Makes Powerful LLM Jailbreakers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.803668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.803668Z digest=sha256:109cdd85532e5ec5485d9d4484627dcb0dcefe4076e6406699a30a131fe8ad00

Observation bf2099da-6860-4bfc-8dc0-45fd9335f1ec · outbound

This paper cites AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.809302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.809302Z digest=sha256:4a9e3d13275123bac4faf0af83235db2ea132e42d92dcf02b0b597dbae0551ff

Observation 55208603-d1ca-4c78-b1f7-66c6ff447880 · outbound

This paper cites CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models CodeChameleon: Personalized Encryption Framework for Jailbreaking Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.814591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.814591Z digest=sha256:085900bff638348c3b4590d9c7baa690dc10a2a53fb2867abf791681af0a22f0

Observation d571d679-be0c-4c99-90df-625a98a6a874 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.819739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.819739Z digest=sha256:2bbf3535a2a0c9af062733e310ea4b4b95320b9123fdeef890c6875cb5a87ea8

Observation 82f49d09-3b5c-4066-875e-50d81a985c0d · outbound

This paper cites To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models To Forget or Not? Towards Practical Knowledge Unlearning for Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.825181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.825181Z digest=sha256:09cd81528accdb0503c5d78aff0818d66ca71212d203486cf0df2e1e3834c406

Observation 3fc416fb-f0c7-4594-adab-06930b48ebfd · outbound

This paper cites Evil Geniuses: Delving into the Safety of LLM-based Agents.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Evil Geniuses: Delving into the Safety of LLM-based Agents

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.831418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.831418Z digest=sha256:1ed7608e583400ca67089cd52c17547852f6329399135871a6bce68a1588db1c

Observation b531d147-0991-469f-91a4-ed6da32878cd · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.836648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.836648Z digest=sha256:115be076e5c0dbd4a9f41cf2f6bb66bbd27e37564a6663f5d2cd22a11f5255b2

Observation eec40907-6d21-4c45-946a-3b556c0aca3f · outbound

This paper cites Finetuned Language Models Are Zero-Shot Learners.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Finetuned Language Models Are Zero-Shot Learners

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.842583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.842583Z digest=sha256:4e2387e3817aae3d05a25ca65ef771100e362eb6d4ef0bea5ee6fb3871d16975

Observation 3b63665b-7c5a-40e5-81a5-57dea9952ae5 · outbound

This paper cites an unresolved cited work.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.849066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.849066Z digest=sha256:0b64a43767b1b4ba2067147c592f082024a3b82d203f71b5852283c3dc3a23cd

Observation 9842d1c3-4b93-464b-aded-7334c802a1be · outbound

This paper cites Low-Resource Languages Jailbreak GPT-4.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Low-Resource Languages Jailbreak GPT-4

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.854124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.854124Z digest=sha256:dc9330ffb286bf3a22b05742c851794100b67b60826285cab9e74f1a9a47eea6

Observation 57639e34-162d-43de-a771-01d083497ae9 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.859545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.859545Z digest=sha256:2a752250c5f1eb224476efb32504b73081bb40de4e8757ee58611019afb34173

Observation b45cb1cb-4d4b-4557-8dae-6afd2da702b0 · outbound

This paper cites GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models GPT-4 Is Too Smart To Be Safe: Stealthy Chat with LLMs via Cipher

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.865614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.865614Z digest=sha256:d88b650e360a9b2b0a79f85d456fc166d75820c46afc19af36950cdc3b5e25f6

Observation 19faeb0f-9330-440e-bb5c-f0295556c578 · outbound

This paper cites How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.871022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.871022Z digest=sha256:d520c47b6124db5f5ecabe86fcfe253d7aa05b3c6615f14d1118bf2bd230437b

Observation 1be42727-07c2-4019-9386-f7c37528ff82 · outbound

This paper cites Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Make Them Spill the Beans! Coercive Knowledge Extraction from (Production) LLMs

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.877687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.877687Z digest=sha256:6adb29fedfd61627ba7fddc6452da5dfc09727a1a29b86667fad2298c05b528d

Observation 733aab27-5d09-4e3f-8a73-4d1720e9ffc3 · outbound

This paper cites Don't Say No: Jailbreaking LLM by Suppressing Refusal.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Don't Say No: Jailbreaking LLM by Suppressing Refusal

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.883975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.883975Z digest=sha256:64c630c425180be6ff05f36bbdabddb258a830a304b30c286bd19fa73c66401d

Observation 6421b070-d893-495a-8809-42b896282eff · outbound

This paper cites Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Emulated Disalignment: Safety Alignment for Large Language Models May Backfire!

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.890422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.890422Z digest=sha256:e2c431ab67f86d32d86fa736c5b47223f4b090ebe46a4c33356b776a3161f24e

Observation 7ab447ec-7737-45e9-ba47-142a7673a71f · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T05:33:11.895798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T05:33:11.895798Z digest=sha256:7feb049767ebe486d18d4dee8b28880617501e3089a9ecde442d56843bb67e7a

Pith citing papers

Observation 58744180-127f-4436-9145-3444b7cf8dae · inbound

Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring cites this paper.

Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models

Reference 19

Resolution
malformed identifier
arxiv_id, observed 2026-05-16T22:31:19.417917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T22:28:41.134253Z digest=sha256:308aee67193423222a29f5963a7cd90c749fb3457d590ac8f3ec3e747da779e1

Observation 93ab2728-ee18-44e4-b6c9-379942709885 · inbound

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints cites this paper.

Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:01.722974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T16:04:25.851592Z digest=sha256:657c1dcf486342090322e7dc21a2dc712c571b0ddb137be4780ece76858aea4c

Observation c1e420db-5126-429d-bbdc-2a73b18457bd · inbound

Mask2Shield: Strengthening LLM Safety against Neuron-Pruning Attacks cites this paper.

Mask2Shield: Strengthening LLM Safety against Neuron-Pruning Attacks NeuRel-Attack: Neuron Relearning for Safety Disalignment in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T03:57:57.052252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:57:57.052252Z digest=sha256:40ac206d5c84757375ffc94801d69a94e6597318ebf2e195852564eb9eb9cfcb