Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2405.18166.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T12:25:30.717668Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-19T11:53:03.533122Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 96d8eb54-9b08-4fa5-a3d1-5330150667ef · inbound
JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 968a4e6f-0bb3-495f-b052-a572df396cbc · inbound
Beyond Black-Box Obfuscation: Mechanistic Analysis and Defense of White-Box Monitors Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 011b081c-b633-4b13-b08b-47f08cc71277 · inbound
GloSS over Toxicity: Understanding and Mitigating Toxicity in LLMs via Global Toxic Subspace Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e9739da-bfe8-407d-8fbb-0b0a745ba207 · inbound
Disentangled Safety Adapters Enable Efficient Guardrails and Flexible Inference-Time Alignment Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3b6d11f3-dffd-46e9-91d9-fd8f4ce197b0 · inbound
PUZZLED: Jailbreaking LLMs through Word-Based Puzzles Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e9e3dab-f6bc-4b05-83ed-17188a1e8e3f · inbound
Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d3e94ee-1bdb-4344-8ae2-4535f55d15db · inbound
Silencing the Guardrails: Inference-Time Jailbreaking via Dynamic Contextual Representation Ablation Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 106f1c68-28ca-4dcc-9156-33c3e458f114 · inbound
Large Language Models Generate Harmful Responses Using a Distinct Mechanism, Shared Across Harm Types Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7645e6d1-bd18-4654-b0f6-c8e614faa764 · inbound
LLM Safety From Within: Detecting Harmful Content with Internal Representations Defending Large Language Models Against Jailbreak Attacks via Layer-specific Editing
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.