Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:26:52.695768Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 4 inbound Pith citation observations for arXiv:2412.11041.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T15:26:52.695768Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:24:12.897878Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-04T23:10:21.606208Z
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 20d9117e-33ae-44df-ad31-2e6bd620ebec · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 5780d874-7083-41a5-8e18-73b4ec64eb3e · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fef8731-9b0b-41a5-94a0-6c0780b08a18 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Language Model Unalignment: Parametric Red-Teaming to Expose Hidden Harms and Biases
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 436f0485-0fa4-4b07-adad-2a09eb473997 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 654620ec-49ff-41e0-a74a-981d3917502c · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models A Survey of Model Compression and Acceleration for Deep Neural Networks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3efd6f7c-b4a5-478f-9070-7d027b9122ad · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Training Verifiers to Solve Math Word Problems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f49de6f4-ab6e-4500-b55d-c41a1089914e · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models The Llama 3 Herd of Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cc6f91d-854b-4ab0-a6b9-6c922a17ac72 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ff76bc74-5778-4661-bc35-aaf94d76bde5 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 391e41f9-daed-4830-a854-9b8af45a00bf · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63112e83-81f7-4109-8e5f-7d8860b88571 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 80dcc775-929c-4de2-9da5-4ff0b50253e8 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fab30056-cded-43d6-9723-7be040661c84 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a760d5d0-e7b6-4a21-98fe-59cb815afeab · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c915d5a5-f44a-4c8d-a0be-236441c0e779 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 0f77d360-f070-4556-a17c-b8aa72b73249 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fad8359-bd68-4519-be87-115b513cca96 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 931c08c8-7044-4532-8f46-d42aae052d28 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfa19642-6d0a-4f8f-8501-83c2e24b0f64 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6c1b76a-ce66-49d3-ab6e-fa11365095f1 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a6e54d1-0d82-4003-9b6b-a209c0808896 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aacff39-88f3-4b8e-8c25-403d39615c4f · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d701fb8-7cf5-4415-aa4f-d1d4a996f0d1 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c6e39ee1-17d2-46d0-8fcc-3c8f291e07db · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18b0b759-7227-4c7b-97cc-2193bdc2ce81 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b11fa973-607a-4f66-ba87-ec89e27a2fbb · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 2072f9fe-7dc8-44c4-8eaf-652b005931a6 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71b2f8a6-e1a5-4c76-91fa-6ff7714bd496 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b96a578-b99d-4089-af2f-74fa2fb95082 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 8e94a77f-fed9-4900-aedc-b1add9661dbd · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1ef5781-803e-4503-8d5c-03e1987ac4b7 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Tree of Attacks: Jailbreaking Black-Box LLMs Automatically
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6ab3608-949e-4ddc-8b40-04df6f109706 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13a68553-b97c-4fa0-8eda-cdd37334623f · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc791333-22a8-4a41-b3b0-a80e663847be · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 919dc5fc-346b-4bf0-a327-d6d131a0d2ec · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffc2bcec-99fd-4bd8-b3dc-4ea21c9b08e1 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 779ba3a6-6c7b-420f-a757-b20b7808f7c4 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 42fb7418-b102-45af-a558-08a1e5addcbe · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation ef0157d1-605b-4477-b213-1200fe5daafa · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e38d6f70-3306-4e70-a03d-6c768dcb7548 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation cefd28b3-f2ed-4f9f-a5bd-c280c4de3775 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2646220-24a2-4db4-ac76-7331ee0c75ec · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0be89ffa-6e65-49eb-9746-298a163b9d48 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26b8fe06-b213-4a31-acf4-60fd49e7fbcf · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c61957f5-b7b2-4892-8fa7-5497a369b7eb · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models A safety realignment framework via subspace-oriented model fusion for large language models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbbac2c7-e774-45e1-a0df-c2981cd26f19 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d85f4b4-7b00-4176-96c9-c278d296ec5e · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7051e64-d50f-4ea8-bda3-95878fbca442 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14d701d1-5945-4b00-907e-6adc1869c781 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 17bd6627-4843-46e6-960b-24752bc1c028 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Learning and Forgetting Unsafe Examples in Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd9b0118-653e-4c34-9a46-6cd8665171e3 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Towards Comprehensive Post Safety Alignment of Large Language Models via Safety Patching
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0430171a-6480-41ba-9c37-005071b2f734 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Is ChatGPT Equipped with Emotional Dialogue Capabilities?
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05397acb-9944-42fd-b9fd-38a1a2238d2b · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56728b04-58a7-4bc0-ae78-6952d3a4eb67 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ec8dfd8-1bfa-42e7-bf65-c2ba6e371bbb · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64f15e3a-fa40-41f7-b966-d92ed54f55a7 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models To prune, or not to prune: exploring the efficacy of pruning for model compression
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 997d931e-c4e5-4aee-ad3a-b3cda03f1b24 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f700512-dfac-4be9-9541-26f38322202a · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models online" 'onlinestring :=
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc893d1-d193-4fb0-86e7-582ea9d73485 · outbound
Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models write newline
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3450871d-fe91-4623-85cd-6dcad79dd033 · inbound
On Almost Surely Safe Alignment of Large Language Models at Inference-Time Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30d6451f-7834-48bd-bf48-a3bd36654b79 · inbound
A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
Reference 294
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b9b2d63-4d52-4e54-86c1-3069ce4e4118 · inbound
Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 014e9948-822b-4be2-ab84-516643ec1fcd · inbound
Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.