Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:06:48.229957Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2412.12497.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:06:48.229957Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 40854e01-e74b-4965-b0d9-cbe40d469c66 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f62d56ff-3c3c-462a-a48d-5c6e327e6c3f · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97e7f11b-9295-469f-b56e-e7c105f557f3 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fe5e615-492f-4fec-9fef-51ca9752eaf5 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b08d789a-be3f-417c-8a36-623d179d9365 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning F.; Leike, J.; Brown, T.; Martic, M.; Legg, S.; and Amodei, D
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation d9decf9f-77b8-401e-97d0-1fa8f5e22ec8 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bfafd4c-6f98-4d53-b514-05e918f83dad · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a64e96ea-2383-4421-981b-1f5701fa8642 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc1f3363-4a3d-41a8-bc03-4280d3fc4d82 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning KTO: Model Alignment as Prospect Theoretic Optimization
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0239bebb-fd78-40f0-bc2f-39878c464e92 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 03b1b44d-4afc-4701-97e0-e3ea5d04b89e · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning ORPO: Monolithic Preference Optimization without Reference Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 60247d1a-0e15-4410-b08e-c23a53876c12 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3be4e412-dd8a-4cac-a863-aa41b3c84101 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning J.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; Chen, W.; et al
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 1cf6fbb8-cf1c-47f2-a878-01a11a2c7bb9 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 690d51f3-c27d-461a-8c21-3a61329e228b · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning F.; and Liu, L
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation dbf4849b-520d-4b25-9590-fac183c3c6eb · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6365235b-d91b-493e-a612-5f4efede6104 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Vaccine: Perturbation-aware Alignment for Large Language Models against Harmful Fine-tuning Attack
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bba3c2a-6f66-454b-a60c-6cccafaaa134 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34ddb55e-4d52-4603-a3ab-5c1c2d3b468e · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Fine-Tuning, Quantization, and LLMs: Navigating Unintended Outcomes
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37dba6a7-9873-44ef-8f81-ea53b7fdb5a3 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 3ada8865-7075-4e2d-ac37-d98dea9d8925 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11ee58a7-7eda-4938-b8ba-fe15bee946e9 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 07953f59-80f5-44a2-ae5d-c74f2578d8fc · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8a70c95-2273-4ece-b0bd-5ff11ef1c5f0 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning M.; Weber, L.; Choshen, L.; Sun, Y.; Xu, G.; and Yurochkin, M
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 23553a85-712b-41ce-81f0-b4d386c4eac7 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bebc58f5-0fab-40a1-a6c5-74861db3b877 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Learning to Poison Large Language Models for Downstream Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e48056dc-b2cb-4cd1-aff1-81eae5d6b4b0 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning D.; Ermon, S.; and Finn, C
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91be9f52-0582-4a44-8897-72b6035d46a0 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Open Problems in Technical AI Governance
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5884aaed-a0e4-4cae-a4f4-211157209f5e · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 155607c6-7134-4392-92ef-3e3f3826709b · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 73d38436-1141-4c7c-9139-04a1373b2d4e · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning D.; Ng, A
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ae85030-7255-4309-8c58-a4856bb3fc38 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation b0e13af2-798e-439e-bfad-f2e1609a6e33 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation cdbd60c0-94a6-4a42-a7d2-75353dabf277 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 5c404844-43ab-4d8a-8466-32a1ffca7b90 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85bc9187-ca80-4bed-aeaa-c403d0630379 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning BEEAR: Embedding-based Adversarial Removal of Safety Backdoors in Instruction-tuned Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc23ba2-e465-4ddf-a77b-9681f0fa6f87 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbd7fd4a-5d4f-467b-9214-c6580f013d38 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Model Extrapolation Expedites Alignment
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddb25a55-3154-46aa-a0c3-959550278c99 · outbound
NLSR: Neuron-Level Safety Realignment of Large Language Models Against Harmful Fine-Tuning Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
No inbound Pith citation observations are available.