Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 25 inbound Pith citation observations for arXiv:2402.05668.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.788943Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-01T19:36:08.555578Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation a9b03fa3-e25e-4aab-91d9-3a40921e7934 · inbound
Refusal in Language Models Is Mediated by a Single Direction JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 73c472e5-a747-4521-b49c-e63f8ad69b10 · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4efbf67c-a5e7-4ffe-9e72-3869d0b34bee · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b2eaf344-2622-4f66-b217-4850c0eb5904 · inbound
Beyond Jailbreaks: Revealing Stealthier and Broader LLM Security Risks Stemming from Alignment Failures JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 266333ba-7c41-48f7-85d0-e27a04f77d93 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4547e9d-2bed-49f3-a872-041039dcc0f2 · inbound
Investigating Vulnerabilities and Defenses Against Audio-Visual Attacks: A Comprehensive Survey Emphasizing Multimodal Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92da5a23-91f1-46d2-80c5-233ed4293158 · inbound
InfoFlood: Jailbreaking Large Language Models with Information Overload JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53031d22-f6a6-49d8-938f-7cc3cf6f8c47 · inbound
Q-resafe: Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 2017
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 483123a3-0bfe-4cc5-9c74-2c4268e7b18d · inbound
Understanding How University Guidelines Address Privacy and Security Issues of Generative AI in Academic Settings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01705d48-dbcf-4a46-bc10-2a5b4228527b · inbound
Linearly Decoding Refused Knowledge in Aligned Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 155f482f-eb0c-4b09-930b-c5236f1f9e56 · inbound
CAVGAN: Unifying Jailbreak and Defense of LLMs via Generative Adversarial Attacks on their Internal Representations JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 833a0af3-8de8-4956-aa95-189fb01a3a44 · inbound
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1426feb-59eb-4b50-8b38-4e88560801cd · inbound
An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d1dba16-b206-4dc7-b45c-769d6d1a2b67 · inbound
Layer-Wise Perturbations via Sparse Autoencoders for Adversarial Text Generation JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c94b144-c84e-4a7f-92de-00f9a3c5adf4 · inbound
ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b71221d5-a147-49b2-8228-46fa8024f7d3 · inbound
SALMAN: Stability Analysis of Language Models Through the Maps Between Graph-based Manifolds JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cd21b06-c7ea-41e7-b016-ad9b2b1c5695 · inbound
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 914ffc20-a3ee-46f8-9239-f750fb86e931 · inbound
JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e95fd325-7c11-47a5-ae6b-5c7da99a67f0 · inbound
LLM in the Middle: A Systematic Review of Threats and Mitigations to Real-World LLM-based Systems JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4efc1b01-efce-4b62-be4b-0c1e0775fafb · inbound
How Well Do AI Systems Solve AP Physics? A Comparative Evaluation of Large Language Models on Algebra-Based Free Response Questions JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83b253df-03b6-4e6f-9d82-15dc863af4a2 · inbound
The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb8ba95f-ef39-43f5-b0d1-985a43a95377 · inbound
SoK: Robustness in Large Language Models against Jailbreak Attacks JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 356aab86-984e-425c-80a4-1165029df8ba · inbound
Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6453eee7-1ef5-458a-9d47-4ea67e510ce6 · inbound
SCARCE: Scalable Cascade Analysis for Rare-event Characterisation via Embeddings JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 30413506-c871-4d98-ae6d-1776f7b318c1 · inbound
AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models JailbreakRadar: Comprehensive Assessment of Jailbreak Attacks Against LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.