Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2408.17003.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T23:08:02.302336Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation c9d6c6d8-a600-496a-9fc6-c334deb18f87 · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a175d9f-deb9-4e8c-8ceb-07f6bee309d0 · inbound
The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd8fe59-9419-43ff-9ff5-03d46611f82f · inbound
Reshaping Representation Space to Balance the Safety and Over-rejection in Large Audio Language Models Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc02672-ba13-4cb4-8174-ab22e8cf3187 · inbound
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24787934-f6cc-4630-b1bd-907232ad4328 · inbound
Depth Gives a False Sense of Privacy: LLM Internal States Inversion Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 584c895b-3c08-4de2-92b6-025324c04dff · inbound
AttenTrack: Mobile User Attention Awareness Based on Context and External Distractions Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69d67a9a-514b-438a-90db-0a0fd96c33e1 · inbound
ThumbnailTruth: A Multi-Modal LLM Approach for Detecting Misleading YouTube Thumbnails Across Diverse Cultural Settings Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e084c66c-68b0-4f02-874b-7e9a41c635bd · inbound
Anchoring Refusal Direction: Mitigating Safety Risks in Tuning via Projection Constraint Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9e438f-ad96-431b-a1f4-ed0d470b50f6 · inbound
You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c1476a3-099a-4a33-9b61-d5606d6db2b1 · inbound
You Had One Job: Per-Task Quantization Using LLMs' Hidden Representations Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3481c5ea-0f6b-4fd7-9030-6b18cc4775ad · inbound
A Lightweight Explainable Guardrail for Prompt Safety Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f77c2ff1-48ca-4d32-affb-ba58a833a3d3 · inbound
SALLIE: Safeguarding Against Latent Language & Image Exploits Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 139f1aa5-5823-4014-8e25-866da150b6a3 · inbound
Why Do Large Language Models Generate Harmful Content? Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1587fbe5-177a-4d03-94cf-e5ceb4ca1ca0 · inbound
Preventing Safety Drift in Large Language Models via Coupled Weight and Activation Constraints Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d09d643a-6e56-4864-a8ff-6dafdace5ff8 · inbound
LLM Safety From Within: Detecting Harmful Content with Internal Representations Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 754ebd1a-ce5c-4832-a16a-f49f72bf005a · inbound
Few-Shot Truly Benign DPO Attack for Jailbreaking LLMs Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9fb2d29e-5e4b-4ff3-944f-99559e102d70 · inbound
Defenses at Odds: Measuring and Explaining Defense Conflicts in Large Language Models Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21cc8787-5edb-4d31-ac65-ae4b255935fd · inbound
Defending Jailbreak Attacks on Large Language Models via Manifold Trajectory Kinetics Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b5a85a15-ae91-47a9-b963-c642c9be3b3e · inbound
Do Activation Monitors Survive Model Updates? Benchmarking, Predicting, and Repairing Activation-Monitor Staleness Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 986a73b8-f404-4c7f-a866-6e054d114540 · inbound
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9a408c1a-a20c-44a1-9484-7a1639be7c37 · inbound
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1bc1f87d-385a-4433-91f2-33b429b9c815 · inbound
A Dual-Hypothesis Reasoning Framework for LLM Guardrails Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f418fcab-282e-49ce-a78c-99e4c6b4dd68 · inbound
GhostPrompt: Cross-Image Adversarial Prompt for Vision-Language Models Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37fc86a9-f4b7-4be8-b4bf-417450e1c4e0 · inbound
Visual Token Compression Enhances Robustness of MLLMs Safety Layers in Aligned Large Language Models: The Key to LLM Security
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.