Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2401.18018.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:35:39.020367Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
5
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 0de766c3-c60f-451e-94c9-6ae8a1726782 · inbound
Refusal in Language Models Is Mediated by a Single Direction On Prompt-Driven Safeguarding for Large Language Models
Reference 206
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 377af2ab-81c0-42a1-a840-06d03b8365ca · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey On Prompt-Driven Safeguarding for Large Language Models
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77e6bcea-72a4-4f40-aedc-3c1047fb9f50 · inbound
COSMIC: Generalized Refusal Direction Identification in LLM Activations On Prompt-Driven Safeguarding for Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da8b92c4-2591-4935-b4ff-28cb1244cdb6 · inbound
ReGA: Model-Based Safeguard for LLMs via Representation-Guided Abstraction On Prompt-Driven Safeguarding for Large Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36cde0a9-52cb-47df-8885-9a0c480d83cd · inbound
Just Enough Shifts: Mitigating Over-Refusal in Aligned Language Models with Targeted Representation Fine-Tuning On Prompt-Driven Safeguarding for Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c46fd5ac-48fa-416e-9127-e34bc783e919 · inbound
Defending Against Prompt Injection With a Few DefensiveTokens On Prompt-Driven Safeguarding for Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 408b7a6e-c13e-4e9d-85a8-fd40974b14f4 · inbound
Beyond Surface-Level Detection: Towards Cognitive-Driven Defense Against Jailbreak Attacks via Meta-Operations Reasoning On Prompt-Driven Safeguarding for Large Language Models
Reference 197
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c96e3b12-54fb-4287-967a-e210244f832d · inbound
Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection On Prompt-Driven Safeguarding for Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af93b3cf-a963-466e-813d-761b4e2907a2 · inbound
Probing the Difficulty Perception Mechanism of Large Language Models On Prompt-Driven Safeguarding for Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9629203d-6a04-42d9-ba97-2cbeaf89f271 · inbound
Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks On Prompt-Driven Safeguarding for Large Language Models
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4569ed4-f15f-4ffc-8149-6e402c9d083f · inbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses On Prompt-Driven Safeguarding for Large Language Models
Reference 245
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afa072e4-9ff1-4b24-850d-2432afd9bbf9 · inbound
RACC: Representation-Aware Coverage Criteria for LLM Safety Testing On Prompt-Driven Safeguarding for Large Language Models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d5160d1-3eed-4b90-a28b-f1a98939264c · inbound
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment On Prompt-Driven Safeguarding for Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a8bc067e-a858-481a-a400-4b0f18743e42 · inbound
Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs On Prompt-Driven Safeguarding for Large Language Models
Reference 131
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c8a09e59-74ee-4937-83ad-b7ada4e165b4 · inbound
A Single Neuron Is Sufficient to Bypass Safety Alignment in Large Language Models On Prompt-Driven Safeguarding for Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 155d4921-eef9-4a92-9c50-46cb5fab9cf2 · inbound
CLAP: Contrastive Latent-space Prompt Optimization for End-to-end Autonomous Driving On Prompt-Driven Safeguarding for Large Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a30150e8-b607-4b84-bdb7-75acaad5a53b · inbound
Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction On Prompt-Driven Safeguarding for Large Language Models
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16e62ae5-3a5f-4a29-94b2-79b8b61791a7 · inbound
Do Activation Monitors Survive Model Updates? Benchmarking, Predicting, and Repairing Activation-Monitor Staleness On Prompt-Driven Safeguarding for Large Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2fbc2da-a71e-4a66-933c-d61fc06f346f · inbound
Visual Token Compression Enhances Robustness of MLLMs On Prompt-Driven Safeguarding for Large Language Models
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b1d80d9-b084-414a-9a8f-51ecc7ee23d1 · inbound
Efficient LLM Adversarial Training via Low-Rank Defense and Circuit-Guided Surrogates On Prompt-Driven Safeguarding for Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.