Pith. sign in

Paper Citation Record · LEDGER

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment

As of 16 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 0 inbound Pith citation observations for arXiv:2507.18631.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18631 v2

Coverage vector

measured 60 of 60 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:13:58.320855Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

60 of 60 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved59
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 853d5b88-5b37-4838-8b8d-3928eb6c2755 · outbound

This paper cites online" 'onlinestring :=.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.063350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.063350Z digest=sha256:ddc06d20840b135c1c3afd3322f012970be32d34408463953b60bd76d2db4949

Observation c4ea7ce4-9aba-42c9-b50b-17e999d18e22 · outbound

This paper cites write newline.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.068887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.068887Z digest=sha256:4bca881a8c1e7919432f304a95875b591bf2a1ed8102093226d63ba3c573763c

Observation f8963a37-f4d4-4fe0-a1fe-642bb77782a8 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.073741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.073741Z digest=sha256:58d165d37a700b987ec2fa8367801b922343d09bc51adee327d1fa89041ce17b

Observation 0feb4913-2f92-4f62-b2a4-f21cdc343e30 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.079059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.079059Z digest=sha256:4cfbad2e82513dca9f83fce12e3b9375b9625f7ad2339d0022a05fb1f93458aa

Observation 34d717fb-f455-4513-98da-761d3975516e · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Evaluating Large Language Models Trained on Code

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.083669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.083669Z digest=sha256:a352544b4c2c7ec5dd8306fc4a0222b63c882dd98ceb413ba555698de825c9be

Observation 95f71dd0-a009-47e1-9b09-0a6ea0771e2b · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.087943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.087943Z digest=sha256:247c6ef154dfb92520d1ea2dc35cb445e1d953a037945fdfe66d40d7bcbfd981

Observation eb39f237-eb76-4b3e-9b9e-f641630cc130 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.351858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.092309Z digest=sha256:a4c72631f4189f177d10feb3722f18a360ffc0f2c0e311b2a6e9b23e0ea4dc92

Observation d108daec-9c67-4b2a-95f9-5488ec0d0216 · outbound

This paper cites Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.096730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.096730Z digest=sha256:54acc6a89f2030edec777a9eae29cc2f911276b37e1a22ec54c3b2d838f5274c

Observation b0cb59f1-cc2b-48e8-9ece-a4a3dd1880a4 · outbound

This paper cites Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Safeguard Fine-Tuned LLMs Through Pre- and Post-Tuning Model Merging

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.102071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.102071Z digest=sha256:59f48f36d8ef537c74f0a95576719577061b551174f2c5ba2a60354526712cd3

Observation eab81e58-646d-4830-9ef8-64a26358ae24 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.107373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.107373Z digest=sha256:c9604746fbb9068b4746778477449eb2d2249e56bc8ae0dfff52a58563d64da2

Observation 3ec54ef0-dfa6-463a-a861-6c37894113b3 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.113830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.113830Z digest=sha256:be0b98ee1ea002dcc71366a51a219a0f48623dab2ea3acc748fe7655f75095ca

Observation 21ffb058-7435-4eee-b72a-03113c0aa761 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.330442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.119017Z digest=sha256:29f65760e39bee8abe318d9facd45c4b89fc5c4a5f64beec57caf6545d96e787

Observation 3e89adbe-4c88-49a1-a2a2-dbf903111736 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.317624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.122946Z digest=sha256:0506a1fa638fe5892cf4a42b32b2c1b9f3ae041331810e46edc45d47b391258e

Observation 6f82fae8-639f-4c48-80c1-85495d458c34 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.126797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.126797Z digest=sha256:d404c9860ac7b0112f87406d7db57e63c7c54039d030e41d93b3cfee5a3e9117

Observation 31ccae20-a47e-4daa-b67b-a85a597e9050 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.130907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.130907Z digest=sha256:392c9a16e806740b8d243e53edc62232d10b74e3c5be249f7b246bc72d9015ab

Observation 184995cf-ca86-4440-a72a-559a7f391ea4 · outbound

This paper cites VLSBench: Unveiling Visual Leakage in Multimodal Safety.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment VLSBench: Unveiling Visual Leakage in Multimodal Safety

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.134740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.134740Z digest=sha256:dd7ba6b1aa22402451d12fb85d994accd99b9fa9915773e57401763f71c405db

Observation bd18d47d-7abf-4508-898c-8623a8408cbe · outbound

This paper cites Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.138769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.138769Z digest=sha256:562b3f7db56cb27174dac0d76b38793ac34ceb15c624f82f0de0a4283ddc7f7d

Observation 73f76105-b10d-4c98-9208-e362b0c12fcd · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.287956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.143005Z digest=sha256:519091a10108bfa851917123941f8bb5caa952dd9b9b83a8b564a920dfd0a060

Observation 3de25048-435c-4f7c-883c-9b4a34220732 · outbound

This paper cites Mistral 7B.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Mistral 7B

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.147434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.147434Z digest=sha256:3cef1fb38ad117ea83c9ea0d180c91c88b54df1cca18a9ecce1d971c351c049b

Observation e0600772-ae63-4681-a634-df59de2a1b78 · outbound

This paper cites PubMedQA: A Dataset for Biomedical Research Question Answering.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment PubMedQA: A Dataset for Biomedical Research Question Answering

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.151662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.151662Z digest=sha256:ac4492892dacbb9edbe9fafdc0c31095b1aad0a01143137b9409a87d646e5d7f

Observation 0525ec50-605a-4d1a-a707-e5adc41d6dce · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.274830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.156078Z digest=sha256:fccfb2e423369c1c9888e18980cdf8794e9ce23237babc62a9062ff35cdb4220

Observation 1f89ed07-990b-4cfc-8f93-ba7549c2b795 · outbound

This paper cites SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.160012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.160012Z digest=sha256:cbe7c7cdf997fd953a459dbf94ad3bc039738a0d3f3f210dc5f20e137c2ffe58

Observation 155e2823-2219-4059-b2ed-1ce90deb75d9 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.261019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.164377Z digest=sha256:68d5e6213490aa35995e68c515152709afda0e019b11d9494f2e23980fee74a4

Observation 1827fb46-584a-4792-9e48-80aa1a195127 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.248134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.168244Z digest=sha256:d479bc7cd9a529b2d091b0a44dbfe12b06478276c8c1a83285e8c42d20db7e57

Observation 67c4c9a9-cb02-4d1c-ac31-595be8e96163 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.234462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.172127Z digest=sha256:354d94f761135708d66bf3f87fd6601b557ae1b710901d00cc083898b54eca10

Observation ef3e8b43-1d09-417c-92bd-386a0cc1f3b2 · outbound

This paper cites The Llama 3 Herd of Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment The Llama 3 Herd of Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.176161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.176161Z digest=sha256:748fc94e89e43218fd350cbe19267467c438c2cad339024e3725168536dec8fa

Observation 0ccec216-7161-42f3-8662-6d52a06b04c7 · outbound

This paper cites Le, Barret Zoph, Jason Wei, and Adam Roberts.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Le, Barret Zoph, Jason Wei, and Adam Roberts

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T18:13:59.221218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.181169Z digest=sha256:52696aca8df9549aaf8723312ffc81027dea35de73e49047774da771d35a7719

Observation 63b3335f-14ac-4fd6-b0a2-446ee4ea9f79 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.185761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.185761Z digest=sha256:a27ef9ac1b6fdcf001b9dde14e6030416a407dd79902acf29c1014fb2e4c09a3

Observation 92a0e61a-e8a1-4e22-a4f0-894b0f6f7b7c · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.207039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.189793Z digest=sha256:6532ff4ecbd7e55eeaa8211f4899a0d9db3bad3630483ec954ba93bc097c06b9

Observation f901c20c-96aa-45b4-b63b-0e2b2411d0a8 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.193740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.193740Z digest=sha256:7f742086397f288f90f73c75e8f28f5a7fc198aef77cfd3ae89cc2672f2b09d0

Observation 14f144ce-bfac-4c26-a57e-291a99e6f821 · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.197941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.197941Z digest=sha256:81edb3c5c8326df4d273ec537642f3bf2fb167b671139504f92f7acc434981b1

Observation 4a5a84a6-a340-46a4-8c79-755e0cf8f018 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.202174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.202174Z digest=sha256:af4d58e43f14e8970f0d22018010c951d5ce319512385b618bd64734e3946437

Observation bc8e3bf2-20bd-495b-b762-f8539d6ebb80 · outbound

This paper cites Orca: Progressive Learning from Complex Explanation Traces of GPT-4.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Orca: Progressive Learning from Complex Explanation Traces of GPT-4

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.206597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.206597Z digest=sha256:78d01403d82b5262dc645df317553bff57d095c1f52e71d1a39e11dc5e3ba09a

Observation 7ec92b30-77f1-4820-b4c2-2cdb605dc7f3 · outbound

This paper cites The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.211236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.211236Z digest=sha256:73d0931b3596ff138c9fad809470067e71975bfd7bcd2223b4f3f7fa5effe656

Observation 963cd09d-a9a1-409b-a56d-39695c76b58e · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.216544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.216544Z digest=sha256:37bd70d8af638fd8f917272e9f0ac6e02a25c5ac8fbda4cef172c249ea87bbf8

Observation 171f5cbf-5b79-4229-aa8b-5396ee8de8bf · outbound

This paper cites Instruction Tuning with GPT-4.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Instruction Tuning with GPT-4

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.220450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.220450Z digest=sha256:1f764d0be83ddd13de9fb7d323ed1cd1ebbc107ad56f7e713ba4792be5af36e4

Observation 96781a24-793d-4aca-b356-404b7ec5096d · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.224735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.224735Z digest=sha256:404a5644fa21c4545b59c8155f7749b90484f47caf437c47ed4f7da0f5865111

Observation ab883154-8edd-4b91-a4b1-73af5c88e90a · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.177364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.228636Z digest=sha256:e45c737dc09ac7c8d2b288359e918d45b483bbbe741dff328681168282ba6e70

Observation 451dedea-7b0b-4cc8-9c59-e863dfd9534a · outbound

This paper cites Qwen2.5 Technical Report.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Qwen2.5 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.232479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.232479Z digest=sha256:d0a81ad7f63469a9244bdf7e793a942b38d94ac75da2595b66560d4a2ba40192

Observation d49d37f8-03e6-4018-8feb-6f349ab08c85 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.236321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.236321Z digest=sha256:c12a2d7609842d0a4535c521eaf54b3c3932c712f3d099d8e4e40f33336a7feb

Observation 2888ba6a-04df-4405-b813-0250bd762f94 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.164425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.240260Z digest=sha256:784fdfb2b134990bd96ba323211ded29c0df1f67a788822f3c1fc338a357ac96

Observation 564797de-c99d-4454-b729-85aa6e7e69ba · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.150561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.244219Z digest=sha256:38091dffcba0a382cfb54a12e1349020c93a074c31351714f57b10d32c9ef640

Observation 5e4944a2-8146-4fa9-bc7e-422ed32d7d06 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.137665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.248560Z digest=sha256:792037a61bdc71ca2d3a56177b67e90ca25a3523296c0bc85ffc0a1f1cb643fe

Observation fb90d4b0-da16-44bc-bb0f-d2074f7446a6 · outbound

This paper cites Hashimoto.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Hashimoto

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.252634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.252634Z digest=sha256:48464fb4bcd69756bca1b84d6100f4f4bd2b7be746804490fe7be3ae8c5ad8a5

Observation 72a653c1-e72f-4547-8f1a-fbec0a5d1d4f · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.256725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.256725Z digest=sha256:3bf1396462532a678f1a360a97bf0ffa63dbdcb43704ff346813517d6a610729

Observation 53369d06-f6a7-46af-a47e-8343964a5297 · outbound

This paper cites DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.260498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.260498Z digest=sha256:ad055edf1304cdc8d68094e92600889dcd9afcacd73a071352c143d712927cea

Observation 59a330a3-7630-4926-a78e-eb2d59ded940 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.107927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.265181Z digest=sha256:6a32f807e668c2cfedf2745ee859a255dac7f544a6c32d71cfbff20bbecff32a

Observation f04d4964-7648-41d6-8fcc-161f0321a0db · outbound

This paper cites BloombergGPT: A Large Language Model for Finance.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment BloombergGPT: A Large Language Model for Finance

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.269384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.269384Z digest=sha256:a21c58baca95f9974432094eb71ab2e67eeee7c3d42f47d062d3a7edf17a0d5c

Observation f7aec56e-9a6d-4a00-a148-0ec82b4c2fd5 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.273453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.273453Z digest=sha256:ab8b896fdec73ea842b126f70e7ca2bae0d252e22ade9c68848e9fb6f8e60fdc

Observation 06b08e16-65e4-416d-9cca-039eeb1a402e · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.086526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.277662Z digest=sha256:d812f3b820c48c11b6908a8fdce5cd7fe8165ff1c295288606d98d4f1ec7e150

Observation 6ad2df5b-c7ae-4824-9c5c-3e0207162186 · outbound

This paper cites MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment MetaMath: Bootstrap Your Own Mathematical Questions for Large Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.282824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.282824Z digest=sha256:1ed6db28d07ff9bb1e8986db3be016ee8bb8725941b365c3db65de96bbb78cea

Observation c22a5ecc-d184-4848-8f80-e37a9ff148b5 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.287060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.287060Z digest=sha256:523f9b013d43d3cdd66e8678ca554d7f59f855a7c96ff27f09b2f0e76f45c33b

Observation 500277c0-7578-402f-91d5-72864343f2d5 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.073501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.291006Z digest=sha256:fe0a3b8996803f59493854f57f87bdf4e5b45bdc4819c96d223a0158e32febc2

Observation 448351e1-35eb-436c-8526-e03f6041d412 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.059765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.295289Z digest=sha256:cc4ea2134c6a744b971115f7051becc97338335a53387adaf7c0fa3b62b23393

Observation 9801f925-faa7-4bbf-8d7b-df8ab9294fd2 · outbound

This paper cites an unresolved cited work.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-15T18:13:59.045166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T18:13:58.299423Z digest=sha256:da4d8fb91f6c31f4d2f0191d951b880309b9f7857170432f5a10d6a3b76700d2

Observation 2f1efd34-6d5b-4aef-998c-1c6f65924b32 · outbound

This paper cites EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment EasyJailbreak: A Unified Framework for Jailbreaking Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.303164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.303164Z digest=sha256:b906588dac63b1717104d693d903e67265be3c44d54d6881d02a4f21b1f8efca

Observation 4c70829f-df76-4c64-aee7-e9700c6203d7 · outbound

This paper cites Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Safety Fine-Tuning at (Almost) No Cost: A Baseline for Vision Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.308125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.308125Z digest=sha256:7ae140bd171c7fced114269254c0ac115f600438757140c8b499ad87532edabb

Observation 172b5635-ec66-430d-811f-7942b5821db1 · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Representation Engineering: A Top-Down Approach to AI Transparency

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.312212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.312212Z digest=sha256:61433c1fca6d40559dd75b5fcb5d652fdaa01e99aff76ee3021dbc3a685daee1

Observation 123f864d-d679-420f-b0a0-ce10dd35ac5d · outbound

This paper cites Improving Alignment and Robustness with Circuit Breakers.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Improving Alignment and Robustness with Circuit Breakers

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.316618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.316618Z digest=sha256:52a47196925521caa8818b3bed1ae3cf4faf507ffabfc11e6a48fe2c6e44a34d

Observation 0fb29fb7-df3b-42eb-953d-146c7be5fd32 · outbound

This paper cites Universal and Transferable Adversarial Attacks on Aligned Language Models.

Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignment Universal and Transferable Adversarial Attacks on Aligned Language Models

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T18:13:58.320855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T18:13:58.320855Z digest=sha256:bcbab27ea482b0c96e62fc34ab5b41885426d8b8927976ee673af3e4cea09058

Pith citing papers

No inbound Pith citation observations are available.