Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:55:14.070225Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 3 inbound Pith citation observations for arXiv:2507.21182.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:55:14.070225Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T07:46:08.836451Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T04:45:00.946977Z
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19693b82-36a2-4966-b18f-52033639ce59 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a7d25de-d0b7-4356-9b7a-ba51c792bef9 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Towards Understanding Ensemble, Knowledge Distillation and Self-Distillation in Deep Learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ac7078a-6ac4-46c3-9844-d444b27a95fb · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7a6821dd-634d-4643-836b-c012c6f4e095 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Invariant Risk Minimization
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0928c9a-0e0b-4a46-8acb-16bd7f5c0ed8 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning A General Language Assistant as a Laboratory for Alignment
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e079766f-74e4-41ad-9939-9923773acebd · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9548ae66-4538-41e0-9c68-f900bab0d174 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Language Model Unalignment: Parametric Red-Teaming to Expose Hidden Harms and Biases
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3d19aae-fa94-479d-b7b8-225520d0a681 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0560f7f3-a974-4fb3-858d-fe4b17f882a0 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fea9462e-ecda-4490-b008-e3c95d2f2d9e · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning PaLM: Scaling Language Modeling with Pathways
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df6cb45f-5197-4121-9d2f-8cecba6eaa46 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 024a154c-afbc-4364-a7cd-85d5d7c6eb1b · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning BadLlama: cheaply removing safety fine-tuning from Llama 2-Chat 13B
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96a3381c-ad50-460b-9fc0-ac9a05f6c113 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f34a0f7d-d202-4194-ac34-981ac8a18d5e · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Will releasing the weights of future large language models grant widespread access to pandemic agents?
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cacd874-ab0d-458d-9ba2-2d0a5f4f251b · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Spear Phishing With Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f5cc971-ca58-4ecc-b5a0-41caa113420e · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Manning, Dan Jurafsky, and Chelsea Finn
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3cb7f07b-8f1b-4ce7-bc1d-aab912b3cfc7 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8bba5e72-74ee-4bf1-aed0-22d240228810 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Instruct2Act: Mapping Multi-modality Instructions to Robotic Actions with Large Language Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bd5f0f0-a413-460d-9fb1-a00c2f80df74 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c44ad53c-8969-49d4-b7f4-e87dee755ab3 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 49fba8c5-efad-467c-9814-a62acba02429 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Vaccine: Perturbation-aware Alignment for Large Language Models against Harmful Fine-tuning Attack
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a703f7f-b66b-4b07-9f4b-eb0a409e6e4e · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16691936-c84d-4cfd-b88f-ad1bdd547309 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2d34fdb0-117e-4693-9828-ae8a0c129d4b · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c433e800-24dd-42bd-97de-7f0a43cdf9db · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ef7c91a6-7714-45f7-84c4-ec3ad290b026 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Spurious Feature Diversification Improves Out-of-distribution Generalization
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d753062d-7496-4b61-aca7-e85fe104860f · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Targeted Vaccine: Safety Alignment for Large Language Models against Harmful Fine-Tuning via Layer-wise Perturbation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 218c90f4-0d96-4391-a0fd-c8f4b35fb6dd · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Robustifying Safety-Aligned Large Language Models through Clean Data Curation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 139c2e72-1d87-46fe-ba6a-a0cde68c243f · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning BioMedGPT: Open Multimodal Generative Pre-trained Transformer for BioMedicine
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0b3af2a-d5aa-4a63-8a1e-37716d4fdfa1 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning SimPO: Simple Preference Optimization with a Reference-Free Reward
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf9558ce-f833-4d55-9891-fdf2ad5b072f · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e04cf625-f0a3-44c6-9259-b55b0ac7672d · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6596feb-7fb1-477a-b1ef-337b08a99336 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation ac9dd152-c508-4102-a46c-86dd66dd27ba · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c54ba3fa-9e99-4458-8fb5-c8cae4f4c5c2 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation af6d1605-f5d1-47e7-898a-228304c67d56 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c28db790-7490-4122-ab14-0aea2495f2b1 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c4a1bf2-204e-4ff9-adb8-2e6c9304e885 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72a95aad-b02b-4994-98fa-7f5dfb343766 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Manning, Stefano Ermon, and Chelsea Finn
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation de66096f-486e-487d-aa45-cf8eb4c76568 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c19c6b36-8c91-4812-8500-d55804f2845e · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4fa50c34-ca7b-4f84-8ade-5d6d844bfa87 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning The Risks of Invariant Risk Minimization
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 759cadca-3f04-40f2-9ad3-3128302714fe · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a4dbacc2-2948-4795-a75c-1a05cf4278a3 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Ziegler, Ryan Lowe, Chelsea Voss, Alec Radford, Dario Amodei, and Paul F
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 40c438cb-afbd-44c6-8952-a8430d44453d · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 08fc65b0-8348-4838-b18b-88f08afcf5f1 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Hashimoto
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afc5bc9c-b228-49f0-8183-97d65de57d47 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning LLaMA: Open and Efficient Foundation Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76a7a51a-73fc-41d8-9c42-b5831e623813 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a36cc09-9f00-459e-8dc5-5a2558606ac1 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cd38159c-9305-4696-89fe-48a08a622172 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1f295d79-792d-48c6-b9de-c9f683ddced6 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Assessing the Brittleness of Safety Alignment via Pruning and Low-Rank Modifications
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5358107b-11fe-4fd8-9e1c-3ef9ada4c501 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c5aa7db9-83ae-4f37-8887-021ba9e8eade · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Some things are more CRINGE than others: Iterative Preference Optimization with the Pairwise Cringe Loss
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b9ab330-1cd2-4a80-b88a-dcb1d0c34424 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b7964131-1802-4170-9105-c6f4ee58b793 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d69086d-697c-4e8c-9585-f6b5bbf74fa8 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Self-Rewarding Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66e8f456-5755-460c-97bf-a83e487df1c3 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Removing RLHF Protections in GPT-4 via Fine-Tuning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b786a24d-dde6-4123-97a0-53bc3771afcf · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a0b91e01-9ded-4ea9-8f5f-030f3982a2c0 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning Making Harmful Behaviors Unlearnable for Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cea0e38b-c00d-4aa9-bf07-bfd4db651030 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning online" 'onlinestring :=
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91d9699d-f303-4c93-99f8-1c709b823b88 · outbound
SDD: Self-Degraded Defense against Malicious Fine-tuning write newline
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9621278c-7a10-4eb5-a00b-f6a485e84322 · inbound
Immunizing 3D Gaussian Generative Models Against Unauthorized Fine-Tuning via Attribute-Space Traps SDD: Self-Degraded Defense against Malicious Fine-tuning
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dcc4fd69-422b-40f8-8ead-9904507bdeae · inbound
GradShield: Alignment Preserving Finetuning SDD: Self-Degraded Defense against Malicious Fine-tuning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 988ed788-82f9-4433-bb8a-ba90cde03218 · inbound
Emergent Misalignment Recruits a Pre-existing Persona Subspace SDD: Self-Degraded Defense against Malicious Fine-tuning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.