Pith. sign in

Paper Citation Record · LEDGER

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models

As of 21 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 4 inbound Pith citation observations for arXiv:2412.11041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11041 v3

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:26:52.695768Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:24:12.897878Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-04T23:10:21.606208Z

Reference resolution

59 of 59 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 20d9117e-33ae-44df-ad31-2e6bd620ebec · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.387578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.442899Z digest=sha256:02224aaf26ebab48b9c423a12a3e41201cc52cb3130492083b3b891b8fdfe0c4

Observation 5780d874-7083-41a5-8e18-73b4ec64eb3e · outbound

This paper cites Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.448035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.448035Z digest=sha256:8b1cbbd9f726752f7513fe6f41ca61ac9bc22d01be52d2c343e626d4427e476c

Observation 0fef8731-9b0b-41a5-94a0-6c0780b08a18 · outbound

This paper cites Language Model Unalignment: Parametric Red-Teaming to Expose Hidden Harms and Biases.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Language Model Unalignment: Parametric Red-Teaming to Expose Hidden Harms and Biases

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.453386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.453386Z digest=sha256:acb54cff3d0f25160f1f670fc39ea0baf99136c723bd014df56fae489fe3c3c5

Observation 436f0485-0fa4-4b07-adad-2a09eb473997 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Evaluating Large Language Models Trained on Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.457711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.457711Z digest=sha256:6d9315a7a25c9268763b4387f11f1d19625c39302c9b9488dba086a2bfc330bb

Observation 654620ec-49ff-41e0-a74a-981d3917502c · outbound

This paper cites A Survey of Model Compression and Acceleration for Deep Neural Networks.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models A Survey of Model Compression and Acceleration for Deep Neural Networks

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.462508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.462508Z digest=sha256:ce802a3904e32527d5756bc272faa98717e590e33a3a29105cd9ab31bfe14c06

Observation 3efd6f7c-b4a5-478f-9070-7d027b9122ad · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Training Verifiers to Solve Math Word Problems

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.466747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.466747Z digest=sha256:e2bb4b78b5343aa7271e8f0c68f3226bfa166f0783625a072e2861840f1fd648

Observation f49de6f4-ab6e-4500-b55d-c41a1089914e · outbound

This paper cites The Llama 3 Herd of Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models The Llama 3 Herd of Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.472428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.472428Z digest=sha256:79b57d394c45dfde751da4f7ff1f9ba0857fc5aef760e1e0fbea98825b41d24f

Observation 3cc6f91d-854b-4ab0-a6b9-6c922a17ac72 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.373611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.477027Z digest=sha256:d64646c5c19a6dbc771bb47790e87e1369a995e43bcc0729276aa50303bf5264

Observation ff76bc74-5778-4661-bc35-aaf94d76bde5 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.481244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.481244Z digest=sha256:2637df78b140aae8b6a35afe017e86dd22c0213f996c837bb4da77587ad6857d

Observation 391e41f9-daed-4830-a854-9b8af45a00bf · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.485517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.485517Z digest=sha256:74ee9b1739ee705c1106332ac6855df5790542943f2826881f843eec105f727e

Observation 63112e83-81f7-4109-8e5f-7d8860b88571 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.490088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.490088Z digest=sha256:bd7692236a12842bdab8c499fe064da9dd77bd9572e8cead944dda5979b37112

Observation 80dcc775-929c-4de2-9da5-4ff0b50253e8 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.494407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.494407Z digest=sha256:ed98ab85dbf9542cd4f173e605c39a558012564002d574260c0cf9d81851efa6

Observation fab30056-cded-43d6-9723-7be040661c84 · outbound

This paper cites Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.498694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.498694Z digest=sha256:c203df29c9484974beb67c64e60944a6bd72d50a4266140b24b3a646329de5d4

Observation a760d5d0-e7b6-4a21-98fe-59cb815afeab · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.503117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.503117Z digest=sha256:a8baf650aba9e985de426d6eb4564f9f4bfe23ffef7fc682fbe2fdc049781199

Observation c915d5a5-f44a-4c8d-a0be-236441c0e779 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.323709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.507172Z digest=sha256:c099464bd4d7c19f0a34b39caae183e38f63048965b11dba22d3b021eaabb90d

Observation 0f77d360-f070-4556-a17c-b8aa72b73249 · outbound

This paper cites Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Safe LoRA: the Silver Lining of Reducing Safety Risks when Fine-tuning Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.511305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.511305Z digest=sha256:1640a17f140ecb6688d8733ea65d1ebd469c14bc2318734fd9f20a1a3cc7658c

Observation 5fad8359-bd68-4519-be87-115b513cca96 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.515631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.515631Z digest=sha256:12d122f2a0fce340f57626b2da073fe3bd506cfb61b70ccb6afc83a0e185434f

Observation 931c08c8-7044-4532-8f46-d42aae052d28 · outbound

This paper cites Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Antidote: Post-fine-tuning Safety Alignment for Large Language Models against Harmful Fine-tuning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.519325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.519325Z digest=sha256:1b723afef786640c47804361634b4122a98a3efccdbc913db00ce9f5b5caa0cc

Observation dfa19642-6d0a-4f8f-8501-83c2e24b0f64 · outbound

This paper cites Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.523412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.523412Z digest=sha256:5b00eb8c151bafe59a65dbdcafec480f7b97f0578336785c2d3c7c9e68641a8c

Observation b6c1b76a-ce66-49d3-ab6e-fa11365095f1 · outbound

This paper cites Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Lisa: Lazy Safety Alignment for Large Language Models against Harmful Fine-tuning Attack

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.527311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.527311Z digest=sha256:879662910d64ca2b91a78214ae4009ba46d1e45f811bc2144a5f76059f190620

Observation 5a6e54d1-0d82-4003-9b6b-a209c0808896 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.531191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.531191Z digest=sha256:fc9797d1e37418b3d03c53b7b84530d722a5ca81d5185fd2898d4d8a2de86d1c

Observation 1aacff39-88f3-4b8e-8c25-403d39615c4f · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.535003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.535003Z digest=sha256:376ad15745cc351ba09deaab50bc0aabf38951011bba27d9cbfc4c7bdb742f95

Observation 5d701fb8-7cf5-4415-aa4f-d1d4a996f0d1 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.538848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.538848Z digest=sha256:7bcc0d056231e1b313fd3d133ad9acc9cf1a13e885db7e98ce1b6339ce27ef22

Observation c6e39ee1-17d2-46d0-8fcc-3c8f291e07db · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.543060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.543060Z digest=sha256:f85286f2b1aede06c59cdc4f9868ac476c3d407fbda340edd17018d291538315

Observation 18b0b759-7227-4c7b-97cc-2193bdc2ce81 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.547425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.547425Z digest=sha256:222c53135bad5010f1724935bd9395543f6900f0577b63dc1518ae05fbab0134

Observation b11fa973-607a-4f66-ba87-ec89e27a2fbb · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.275194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.551762Z digest=sha256:4258270262d7b46da38c9e0b5eddf6391bbd72db1ba07def00316ae471d392a3

Observation 2072f9fe-7dc8-44c4-8eaf-652b005931a6 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.555942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.555942Z digest=sha256:1b2749dfb8370ea4047a66d85a170a3f7b6e046e92c43b9cafb44124048cae24

Observation 71b2f8a6-e1a5-4c76-91fa-6ff7714bd496 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.560449Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.560449Z digest=sha256:fbd248d2e7c201c3bef92c6325d1b8f07cfae34527653fe04d0b50abba090062

Observation 0b96a578-b99d-4089-af2f-74fa2fb95082 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.253717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.564431Z digest=sha256:554fc61c6da9f59df9134228f242fb2e5346559119ac35b5989d15d32b1c89fb

Observation 8e94a77f-fed9-4900-aedc-b1add9661dbd · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.568458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.568458Z digest=sha256:f6db16744c973e6da14f3b8b5084abdc1f45588a471ef6a49efb8cdba9bef5ab

Observation e1ef5781-803e-4503-8d5c-03e1987ac4b7 · outbound

This paper cites Tree of Attacks: Jailbreaking Black-Box LLMs Automatically.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Tree of Attacks: Jailbreaking Black-Box LLMs Automatically

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.572496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.572496Z digest=sha256:bf2c5c5d412c6aea2bea517b655430a3cdbf236a2ee656b75e14e6acb93c2ad1

Observation b6ab3608-949e-4ddc-8b40-04df6f109706 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.576698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.576698Z digest=sha256:ec90e94cbfca0747084952b1c27a92565a55d901c49536454e964e07abab7da7

Observation 13a68553-b97c-4fa0-8eda-cdd37334623f · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.580951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.580951Z digest=sha256:9827341ceb2e17ef134436f934395425396e74c7712f01782b27c87a15192872

Observation bc791333-22a8-4a41-b3b0-a80e663847be · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.585534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.585534Z digest=sha256:92cfa4651f5059f4fba49b91d7d907af39fe7dba1783ac52a4548f51181b55c1

Observation 919dc5fc-346b-4bf0-a327-d6d131a0d2ec · outbound

This paper cites Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.591441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.591441Z digest=sha256:bfc802f11a27750e51efaffa277e1610c6e7bcff02bd344d60621f40d3dc381b

Observation ffc2bcec-99fd-4bd8-b3dc-4ea21c9b08e1 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.595960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.595960Z digest=sha256:9cca380f769965f25c496553961e8d72ae5db7f30c408c9bb0c8df733a524689

Observation 779ba3a6-6c7b-420f-a757-b20b7808f7c4 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.207416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.600230Z digest=sha256:039ffcf4971f40dc6fff0259bad0e93bc1c5aca186e4bb2a058c9db666c182ce

Observation 42fb7418-b102-45af-a558-08a1e5addcbe · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.193886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.604498Z digest=sha256:8d291f451f224dc28bb9e6a8c55ad615e186df0cd1de6cdab858f670017178c6

Observation ef0157d1-605b-4477-b213-1200fe5daafa · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.609060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.609060Z digest=sha256:6b7d026325bdfe65ecdda395ea5c407b01259f06acf5fcb33b43c6f36c00f684

Observation e38d6f70-3306-4e70-a03d-6c768dcb7548 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.179469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.613531Z digest=sha256:2ac9ca7761906f9ac11b7187af96496c89f91cb2e3292a78e11a9eaa24fd599f

Observation cefd28b3-f2ed-4f9f-a5bd-c280c4de3775 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.617709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.617709Z digest=sha256:5d21bb0328849a4bd2dee4898e2db0ae6ea50b5fd90f4f8a2d0a5076a7f1c6ea

Observation a2646220-24a2-4db4-ac76-7331ee0c75ec · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.621427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.621427Z digest=sha256:4b8cccf77e132eaf8453baa58c0bd94b1264dae72dd9197d28c98a97650972fe

Observation 0be89ffa-6e65-49eb-9746-298a163b9d48 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.625679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.625679Z digest=sha256:69a32729966eb1aba216644b604d452aa5cc05ecb00ef91e8651be905311e488

Observation 26b8fe06-b213-4a31-acf4-60fd49e7fbcf · outbound

This paper cites Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Shadow Alignment: The Ease of Subverting Safely-Aligned Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.629530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.629530Z digest=sha256:9bf17201cdc4c06f696618b9f6f6c393de9e3758c08c1f002d66e404bf864379

Observation c61957f5-b7b2-4892-8fa7-5497a369b7eb · outbound

This paper cites A safety realignment framework via subspace-oriented model fusion for large language models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models A safety realignment framework via subspace-oriented model fusion for large language models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.633567Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.633567Z digest=sha256:5128b37b76332359185fbd35c9b459b4a9812733962d99ae6ba8904d2589e5a0

Observation cbbac2c7-e774-45e1-a0df-c2981cd26f19 · outbound

This paper cites GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models GPTFUZZER: Red Teaming Large Language Models with Auto-Generated Jailbreak Prompts

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.637604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.637604Z digest=sha256:29a73dca01a8be89de322cf38b83ea36640549b6140a6af1806b52131e4c22bf

Observation 8d85f4b4-7b00-4176-96c9-c278d296ec5e · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.642023Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.642023Z digest=sha256:6d2b1cd0bb912dc4f94594c7c86d0d26e925ecbe646f8c0461e061451046edee

Observation a7051e64-d50f-4ea8-bda3-95878fbca442 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.646182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.646182Z digest=sha256:2655b5d561ba2b5ca1c504ad09371e4bd7b188138d3809ded63a325839ac232d

Observation 14d701d1-5945-4b00-907e-6adc1869c781 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:26:53.131073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-08-11T15:26:52.650297Z digest=sha256:b782d186ab1ead1da92576aa25efa45a90d320ff70fdeb7dabcb2c69061d7975

Observation 17bd6627-4843-46e6-960b-24752bc1c028 · outbound

This paper cites Learning and Forgetting Unsafe Examples in Large Language Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Learning and Forgetting Unsafe Examples in Large Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.654464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.654464Z digest=sha256:407e3564d9064429dcd262fa287702d15e33ffe5529d5e4e7cae0993ffa2eeaf

Observation fd9b0118-653e-4c34-9a46-6cd8665171e3 · outbound

This paper cites Towards Comprehensive Post Safety Alignment of Large Language Models via Safety Patching.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Towards Comprehensive Post Safety Alignment of Large Language Models via Safety Patching

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.659260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.659260Z digest=sha256:e8237a57a83bbc162acb8fcd585ff0652974b946a9a52ed2d9fbaf9763d9bb35

Observation 0430171a-6480-41ba-9c37-005071b2f734 · outbound

This paper cites Is ChatGPT Equipped with Emotional Dialogue Capabilities?.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Is ChatGPT Equipped with Emotional Dialogue Capabilities?

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.663931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.663931Z digest=sha256:825841276371241edf5dbbb091da578a6ddc07e07228b1b5a47e662a36366d67

Observation 05397acb-9944-42fd-b9fd-38a1a2238d2b · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.668435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.668435Z digest=sha256:b2aa39cf6f4f2abdf5407cde3ac1b26a1c08689c63b4c6aa8388fb6938b5aa03

Observation 56728b04-58a7-4bc0-ae78-6952d3a4eb67 · outbound

This paper cites an unresolved cited work.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.672849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.672849Z digest=sha256:1e05beea3711343357e5c84082f9ce5e7df9b465d6959eea58505dac7dea4bb3

Observation 8ec8dfd8-1bfa-42e7-bf65-c2ba6e371bbb · outbound

This paper cites Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Model Tailor: Mitigating Catastrophic Forgetting in Multi-modal Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.677212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.677212Z digest=sha256:faf8ae41ca0a35508372a9ac7394538c55311e9fd0fe4f3732a3451739963ebd

Observation 64f15e3a-fa40-41f7-b966-d92ed54f55a7 · outbound

This paper cites To prune, or not to prune: exploring the efficacy of pruning for model compression.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models To prune, or not to prune: exploring the efficacy of pruning for model compression

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.681833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.681833Z digest=sha256:c1b2786772d5a023b3b38ebb273f787d9d8988bc6610a3f32ff055f325aa78c6

Observation 997d931e-c4e5-4aee-ad3a-b3cda03f1b24 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.686099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.686099Z digest=sha256:c083e3b660e73fbad9221fce459e8c7687cfb4302ccf6e676f695a0e87db8864

Observation 5f700512-dfac-4be9-9541-26f38322202a · outbound

This paper cites online" 'onlinestring :=.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models online" 'onlinestring :=

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.690576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.690576Z digest=sha256:3a3d28923a85bf9ce657c5b44de75ffee7307f689237f0a1f2fd3db31b3e3527

Observation afc893d1-d193-4fb0-86e7-582ea9d73485 · outbound

This paper cites write newline.

Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models write newline

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T15:26:52.695768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:26:52.695768Z digest=sha256:8a42cb2242f7a819b618c0651ac8d5e4c5d77954987566679df0c80dd08ed468

Pith citing papers

Observation 3450871d-fe91-4623-85cd-6dcad79dd033 · inbound

On Almost Surely Safe Alignment of Large Language Models at Inference-Time cites this paper.

On Almost Surely Safe Alignment of Large Language Models at Inference-Time Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-09T16:18:40.929011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:18:40.929011Z digest=sha256:19496750fc4e243dc1474e7a2e046a77835a223a9468204771b84303ad41e7dd

Observation 30d6451f-7834-48bd-bf48-a3bd36654b79 · inbound

A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment cites this paper.

A Comprehensive Survey in LLM(-Agent) Full Stack Safety: Data, Training and Deployment Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models

Reference 294

Resolution
unresolved
no resolver link, observed 2026-08-16T11:24:12.897878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:24:12.897878Z digest=sha256:745420e78ab78cf3d90022a8c9268b72cbac7e527a6feab9ce40bd49eea08046

Observation 6b9b2d63-4d52-4e54-86c1-3069ce4e4118 · inbound

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint cites this paper.

Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-08-04T23:10:21.613940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-04T23:10:21.317604Z digest=sha256:43d3a0a08a943398ceaa0226228884a4de1868b6535a913c6aa7edb6e5e44457

Observation 014e9948-822b-4be2-ab84-516643ec1fcd · inbound

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs cites this paper.

Deferred Exposure of Future Trajectories for Verifiable Reasoning in Autonomous Driving VLMs Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:55.402063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:55.402063Z digest=sha256:309ea760edc2b7962b132aaa12b41ae4d19d1a53215c49e5d0409469ece9e2ba