Pith. sign in

Paper Citation Record · LEDGER

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning

As of 13 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2412.12497.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.12497 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:06:48.229957Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 40854e01-e74b-4965-b0d9-cbe40d469c66 · outbound

This paper cites , " * write output.state after.block = add.period write newline.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning , " * write output.state after.block = add.period write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.058083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.058083Z digest=sha256:2b5289614c2e4bd12f94545828c9fe900849345caa762d0ed4ba12ab6e8641ee

Observation f62d56ff-3c3c-462a-a48d-5c6e327e6c3f · outbound

This paper cites write newline.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.063975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.063975Z digest=sha256:f8fde050711e7438b0694d5933cb0adae60b0933d63c2451d3b68f823e861fb1

Observation 97e7f11b-9295-469f-b56e-e7c105f557f3 · outbound

This paper cites Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.069054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.069054Z digest=sha256:70a9567714e8fd1a6846cb9f003ce719743c3ad3a0580af901bddd41332d161c

Observation 8fe5e615-492f-4fec-9fef-51ca9752eaf5 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.075163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.075163Z digest=sha256:896642c8261fe2d98fc4ccb8e6d255747726608d46adc13918e048f1203f1937

Observation b08d789a-be3f-417c-8a36-623d179d9365 · outbound

This paper cites F.; Leike, J.; Brown, T.; Martic, M.; Legg, S.; and Amodei, D.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning F.; Leike, J.; Brown, T.; Martic, M.; Legg, S.; and Amodei, D

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:06:48.796303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.079786Z digest=sha256:737257b14ac40329bdc6b79e1774544dd78d0ebbb0afbe141cb12724b5bda1b8

Observation d9decf9f-77b8-401e-97d0-1fa8f5e22ec8 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.084434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.084434Z digest=sha256:8ccde6c733372b172649020a421a4092655a5916e8d095cf320e1de5cf66eec8

Observation 6bfafd4c-6f98-4d53-b514-05e918f83dad · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.089407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.089407Z digest=sha256:d34d4a52ae0ab5b0697265d0a599e5c1a6b7bfa0ec1aab70a0c078c8493d8283

Observation a64e96ea-2383-4421-981b-1f5701fa8642 · outbound

This paper cites DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.093168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.093168Z digest=sha256:d15c19215732ce44533b4cd1607e925daebeedcee1d969e14287e696a37b26d9

Observation dc1f3363-4a3d-41a8-bc03-4280d3fc4d82 · outbound

This paper cites KTO: Model Alignment as Prospect Theoretic Optimization.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning KTO: Model Alignment as Prospect Theoretic Optimization

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.097418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.097418Z digest=sha256:db954f0f5bbfe8debe7e7baf645be5c87af9c3c738679833d48d99a464c97ea4

Observation 0239bebb-fd78-40f0-bc2f-39878c464e92 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.776979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.101760Z digest=sha256:29ae15e92c4a39cafb4a8bf0d8980f4acdea939d70fbbd6f23336c00968fa0cb

Observation 03b1b44d-4afc-4701-97e0-e3ea5d04b89e · outbound

This paper cites ORPO: Monolithic Preference Optimization without Reference Model.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning ORPO: Monolithic Preference Optimization without Reference Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.105864Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.105864Z digest=sha256:8d90859067e046d6241bc6f16c32b4791edeeb4685fff3068ca3b8aa0ce8676c

Observation 60247d1a-0e15-4410-b08e-c23a53876c12 · outbound

This paper cites Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.111897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.111897Z digest=sha256:e20605b7aac530e2d39f92c8519e28b28cc8f93bbdbc4a80fcdb84df6cf2119b

Observation 3be4e412-dd8a-4cac-a863-aa41b3c84101 · outbound

This paper cites J.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; Chen, W.; et al.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning J.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; Chen, W.; et al

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:06:48.765676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.116771Z digest=sha256:7ca8b6a154beadaebbf9c5b1d08a971c88609606a19cee52ba8e220f7e116af3

Observation 1cf6fbb8-cf1c-47f2-a878-01a11a2c7bb9 · outbound

This paper cites Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.121324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.121324Z digest=sha256:070a3aa68bfa27da36905c834c3613b26dfa233122ce94805f750042bee5a0b9

Observation 690d51f3-c27d-461a-8c21-3a61329e228b · outbound

This paper cites F.; and Liu, L.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning F.; and Liu, L

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:06:48.752758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.125788Z digest=sha256:4be1fe4503af3aa079ae2a3ab8f469e950db11a2956b0070d8c009a62ade00af

Observation dbf4849b-520d-4b25-9590-fac183c3c6eb · outbound

This paper cites Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.130126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.130126Z digest=sha256:5cda7f3bf48eb1dae3aede985b20f4fc46beafd446c7da95e8e376b812d5ab9e

Observation 6365235b-d91b-493e-a612-5f4efede6104 · outbound

This paper cites Vaccine: Perturbation-aware Alignment for Large Language Models against Harmful Fine-tuning Attack.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Vaccine: Perturbation-aware Alignment for Large Language Models against Harmful Fine-tuning Attack

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.134674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.134674Z digest=sha256:296ceb6c9306a5c84d792ec945fcc6ace8e9f6573ee0fa70939b685d1112b85f

Observation 6bba3c2a-6f66-454b-a60c-6cccafaaa134 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.140584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.140584Z digest=sha256:72a26de67683594b70cabe859ffd4c1185ac60607954e0990631c64c21bd68f5

Observation 34ddb55e-4d52-4603-a3ab-5c1c2d3b468e · outbound

This paper cites Fine-Tuning, Quantization, and LLMs: Navigating Unintended Outcomes.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Fine-Tuning, Quantization, and LLMs: Navigating Unintended Outcomes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.144977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.144977Z digest=sha256:f94d2c5e11c326e6fc5ef75ecc6a0c68a44f4b5193a70370c6fb61ca31c2241a

Observation 37dba6a7-9873-44ef-8f81-ea53b7fdb5a3 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.733378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.150529Z digest=sha256:9932651f50aebc9ee3e2fa37d0153f357e968443bd6fcf5bf8aafba311454fe8

Observation 3ada8865-7075-4e2d-ac37-d98dea9d8925 · outbound

This paper cites SimPO: Simple Preference Optimization with a Reference-Free Reward.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning SimPO: Simple Preference Optimization with a Reference-Free Reward

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.154835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.154835Z digest=sha256:c8cf29375905fe95f06e7b830befab3717bf485b63646a60e3859d20d7e90105

Observation 11ee58a7-7eda-4938-b8ba-fe15bee946e9 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.720242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.159390Z digest=sha256:24b4543713a9998549e286133eee4936a2ba0938f44f23b3b76610c9a8fa361d

Observation 07953f59-80f5-44a2-ae5d-c74f2578d8fc · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.163494Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.163494Z digest=sha256:85add62752319cecbdee0ece0357266fd907f8e91c8718819e33e16554152a97

Observation d8a70c95-2273-4ece-b0bd-5ff11ef1c5f0 · outbound

This paper cites M.; Weber, L.; Choshen, L.; Sun, Y.; Xu, G.; and Yurochkin, M.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning M.; Weber, L.; Choshen, L.; Sun, Y.; Xu, G.; and Yurochkin, M

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:06:48.699002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.168044Z digest=sha256:5749602169c9a2db3ae2e866abc272137abbf8f189a021ecd842bbb05a02b061

Observation 23553a85-712b-41ce-81f0-b4d386c4eac7 · outbound

This paper cites Safety Alignment Should Be Made More Than Just a Few Tokens Deep.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Safety Alignment Should Be Made More Than Just a Few Tokens Deep

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.172421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.172421Z digest=sha256:52da860e6d08c7ee62ca3e7a92cd0cb0ee25e38a7c7ac6074a80237fc739e764

Observation bebc58f5-0fab-40a1-a6c5-74861db3b877 · outbound

This paper cites Learning to Poison Large Language Models for Downstream Manipulation.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Learning to Poison Large Language Models for Downstream Manipulation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.176729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.176729Z digest=sha256:ca1d3e5a4cb2272ad6b106487d3e8eb85a73559dfdb150a35861f5c9abb0d61c

Observation e48056dc-b2cb-4cd1-aff1-81eae5d6b4b0 · outbound

This paper cites D.; Ermon, S.; and Finn, C.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning D.; Ermon, S.; and Finn, C

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.180673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.180673Z digest=sha256:e144cbe4a699e537985db1bc3eeec883704a4b04b96d89fcd2de12ec14f85d6f

Observation 91be9f52-0582-4a44-8897-72b6035d46a0 · outbound

This paper cites Open Problems in Technical AI Governance.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Open Problems in Technical AI Governance

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.185516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.185516Z digest=sha256:f85743ada3dc075b1f33c9fb9d16995e185ccdc47b05a7f5bc436f091121836a

Observation 5884aaed-a0e4-4cae-a4f4-211157209f5e · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.665244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.190088Z digest=sha256:80ef794be1a01c950a81dc979d4901c7cef509814a6979bf9cc942fdba7932a4

Observation 155607c6-7134-4392-92ef-3e3f3826709b · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.648895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.193565Z digest=sha256:11982e354f638391a7d348d04ddc0f06573cb40ed48d6002ac1d39929674fd56

Observation 73d38436-1141-4c7c-9139-04a1373b2d4e · outbound

This paper cites D.; Ng, A.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning D.; Ng, A

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.197255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.197255Z digest=sha256:20a56ecbbf90cd572caff3506cd1f2189b63d1aba20674fa2f2a4216c593b2da

Observation 1ae85030-7255-4309-8c58-a4856bb3fc38 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.627338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.201967Z digest=sha256:db566340f4123ef08e37d447967ed3c815d1d72b18231b58ce76bd1d095f3552

Observation b0e13af2-798e-439e-bfad-f2e1609a6e33 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.614745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.205663Z digest=sha256:bc51f9ada97bc12a776839260832ac25841f2366ed87aa3194fe14c80453e56b

Observation cdbd60c0-94a6-4a42-a7d2-75353dabf277 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.601700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.209198Z digest=sha256:dcbfb3d5db5a4445c7ca37fb4e0c8ee2cc6f7eb3352f8cbafd3be097e87be96a

Observation 5c404844-43ab-4d8a-8466-32a1ffca7b90 · outbound

This paper cites Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.213034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.213034Z digest=sha256:90de5bfa2835e7a96b2867c67364b4decfee2f2718edc39bba1ead53eadf80d1

Observation 85bc9187-ca80-4bed-aeaa-c403d0630379 · outbound

This paper cites BEEAR: Embedding-based Adversarial Removal of Safety Backdoors in Instruction-tuned Language Models.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning BEEAR: Embedding-based Adversarial Removal of Safety Backdoors in Instruction-tuned Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.217261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.217261Z digest=sha256:dd0f4c6063d02af8994cc02deba47ab440103591ce25f7b0e3279d918731c5e9

Observation 8cc23ba2-e465-4ddf-a77b-9681f0fa6f87 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.221300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.221300Z digest=sha256:2b34dbf0c785f3ca12257c38da19b9b91f7af380aa36a02dff04d430d9594516

Observation cbd7fd4a-5d4f-467b-9214-c6580f013d38 · outbound

This paper cites Model Extrapolation Expedites Alignment.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Model Extrapolation Expedites Alignment

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T14:06:48.225243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:06:48.225243Z digest=sha256:d0f35431e880c3e0305fc9a8db3c97925740aa81dd8b78566b94bb5e74391cac

Observation ddb25a55-3154-46aa-a0c3-959550278c99 · outbound

This paper cites an unresolved cited work.

NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-11T14:06:48.577859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-11T14:06:48.229957Z digest=sha256:70ea64a3b634b59c40940b1a0c4f5dc16de0c6f88c413d5b7da528a6d056ef9b

Pith citing papers

No inbound Pith citation observations are available.