Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2406.05946.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T09:47:24.461017Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 18e07020-4633-4cca-ac77-ddc13a509810 · inbound
Harmful Fine-tuning Attacks and Defenses for Large Language Models: A Survey Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 122
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0a3315e6-5034-416a-a3a2-574c9ae1d9e1 · inbound
LLM-Safety Evaluations Lack Robustness Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 69257a21-2ff3-458d-9a67-ac2adfc991f8 · inbound
Towards provable probabilistic safety for scalable embodied AI systems Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 138
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 31a3c0fc-9898-45f0-8b3b-81da89cdccf5 · inbound
ReasoningGuard: Safeguarding Large Reasoning Models with Inference-time Safety Aha Moments Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b18694bb-f4ed-470d-95df-23f3ff646588 · inbound
Attention Misses Visual Risk: Risk-Adaptive Steering for Multimodal Safety Alignment Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a85dacc-f99a-4773-9cb3-851478a24a55 · inbound
Reasoning Up the Instruction Ladder for Controllable Language Models Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfbb6043-d7a0-4edb-9bf6-cf0e2bbfa64c · inbound
SCOUT: A Defense Against Data Poisoning Attacks in Fine-Tuned Language Models Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 796e4efe-4123-4436-9ed9-85fff8302d09 · inbound
Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c534c6af-9e58-40e0-bbba-33d68b94a85d · inbound
Robust Policy Optimization to Prevent Catastrophic Forgetting Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation aefd7fba-2534-46b6-8ec3-068284866237 · inbound
IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3731c79d-7bc6-4d28-9d1f-c74def9e77ad · inbound
IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2dfbc05-f563-47c7-aa30-380075ba452c · inbound
Are GUI Agents Focused Enough? Automated Distraction via Semantic-level UI Element Injection Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c548783a-0369-45e7-9738-d533958b6ae5 · inbound
Large Language Models Generate Harmful Responses Using a Distinct Mechanism, Shared Across Harm Types Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5f398ca0-4b21-4259-9724-775fd701064c · inbound
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ccc0fa3b-dfe9-4e8b-b4b9-e7df34205fd6 · inbound
MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a2869571-ad0f-4ee1-8546-bbc2319adc3a · inbound
RouteHijack: Routing-Aware Attack on Mixture-of-Experts LLMs Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation af70cf46-cef4-42b9-b766-97bcd16061eb · inbound
Internalizing Safety Understanding in Large Reasoning Models via Verification Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f8c0cbb3-a741-458e-9e23-509d52f38403 · inbound
REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 191
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b8910e1c-87c9-4d1a-ab7a-1ba54351b815 · inbound
Reducing the Safety Tax in LLM Safety Alignment with On-Policy Self-Distillation Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 50336ac5-2974-406f-ac2d-16be30884918 · inbound
Ablating Safety: Mechanisms for Removing Alignment in Language Models for Security Applications Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 32b70579-c676-4e14-9cb8-9813176209ca · inbound
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0c86dbcc-5cd4-4e1c-85ed-99cd3007d537 · inbound
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 45b8de59-d327-4b5c-a66e-34520b9d6e45 · inbound
REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e74004bf-4bd9-4474-990e-04ed45a96364 · inbound
Detecting Unfaithful Chain-of-Thought via Circuit-Guided Internal-External Discrepancy Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eeb1c9dd-8132-4ef8-9bf8-8e403c29f35d · inbound
MESA: Improving MoE Safety Alignment via Decentralized Expertise Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 54d52469-bcff-45c6-be6b-ec2b969b7c84 · inbound
MENTIS: What Belief Changes Under Alignment? Measuring Multi-Scale Latent Torsion in Language Models Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 40e2d7f2-51b7-4498-9ae6-418d195d0f72 · inbound
MaskForge: Structure-Aware Adaptive Attacks for Jailbreaking Diffusion Large Language Models Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d974c386-bb2b-494d-bd0a-1181b7910d2b · inbound
When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation db63a811-404a-41be-84c2-f94e3316018c · inbound
Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 105
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3509ee17-abb0-4704-ad67-20241ed6968e · inbound
Understanding and Mitigating Prompt Leaking Attacks in Real-World LLM-Based Applications Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d4d13048-04c3-4378-bfaa-27f6348b12c1 · inbound
Forget, Anticipate and Adapt: Test Time Training for Long Videos Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 822981c2-9a40-4540-a0b8-1fb41754b31a · inbound
Forget, Anticipate and Adapt: Test Time Training for Long Videos Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 72b76032-4b80-44a6-9452-18ff921c8d81 · inbound
Forget, Anticipate and Adapt: Test Time Training for Long Videos Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08176c57-f6af-417e-bb42-155045e9ae68 · inbound
Defending Against Harmful Supervision Hidden in Benign Samples Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 52b6147b-d107-47f2-99c0-493177db0293 · inbound
Breaking Safety at the Token Boundary: How BPE Tokenization Creates Exploitable Gaps in LLM Alignment Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13b80043-0285-4561-8a97-cda8b8ec98cb · inbound
Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5 Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b0e5569-2443-405d-8e81-3555d29cf70a · inbound
Pretraining Curricula Enable Selective Fine-tuning Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3df96879-e39b-4505-8a87-b1cbe822d585 · inbound
An Emergent Mirage: Is Emergent Misalignment and Realignment Indeed a Robust Phenomenon? Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54e89284-a98d-4c11-8add-40007f5d5d99 · inbound
Breaking Refusal in the First Half: A Mechanistic Study of the Prefill Jailbreak Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c5b5df1-bb7b-4950-8fc5-89572e115e3b · inbound
Normalized Rewards for Preference Optimization Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e94f6a3c-d03e-4e58-8f17-d7a4b741ab86 · inbound
How Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak Contributions Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18bb393f-b765-4321-a81a-61f84477da3b · inbound
Reason Before You Retrieve: Agentic Planning for Multi-modal RAG Safety Alignment Should Be Made More Than Just a Few Tokens Deep
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.