Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 54 inbound Pith citation observations for arXiv:2404.09932.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T22:53:20.428520Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
14
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation e47d2fa6-25b8-4557-8bcc-89e4f3513b11 · inbound
Scaling and renormalization in high-dimensional regression Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 07a1da5b-b4f0-44e8-b987-1a99bf9035de · inbound
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9c9485fd-3d9b-48a2-a535-f7d52dd85d64 · inbound
A Method for Enhancing the Safety of Large Model Generation Based on Multi-dimensional Attack and Defense Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b11812df-a75e-499f-a9e1-35055655b0f2 · inbound
Safeguarding Large Language Models in Real-time with Tunable Safety-Performance Trade-offs Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edaf3004-9b57-4970-a514-98992a9076e3 · inbound
Mechanistic understanding and validation of large AI models with SemanticLens Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb033e1-ce8e-48c9-a83d-809d6234bd9a · inbound
Self-Instruct Few-Shot Jailbreaking: Decompose the Attack into Pattern and Behavior Learning Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a0ec520-0d0d-45c1-97ad-35ebf5e33ec5 · inbound
Clone-Robust AI Alignment Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e88b23e9-baaa-4345-a6a3-3101215e6698 · inbound
Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a148bf4-fa45-45ed-9888-668ed4c51d37 · inbound
Episodic memory in AI agents poses risks that should be studied and mitigated Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 926cd96f-cc6f-4f4b-ae9a-489852782b6f · inbound
CASE-Bench: Context-Aware SafEty Benchmark for Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e708e35-c5da-4e6d-831e-ba710e2645fa · inbound
The AI Agent Index Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d821892-22fe-43fa-9cf9-a296911e69e9 · inbound
MEETING DELEGATE: Benchmarking LLMs on Attending Meetings on Our Behalf Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67090237-d39b-4ec6-85bb-953388fac2cc · inbound
JBShield: Defending Large Language Models from Jailbreak Attacks through Activated Concept Analysis and Manipulation Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 276b0fa3-77a0-4277-a525-fee2aa77b9ef · inbound
A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 167
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2bc9d3-9b3a-4bd8-a2df-7f37c0a94918 · inbound
Mitigating Deceptive Alignment via Self-Monitoring Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eaaae2ea-aa11-4887-9fc5-b7df25e0b9b0 · inbound
ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6604e4c7-4876-4fbf-a6c5-4753b52ba655 · inbound
Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d1cbc67-e62a-4ab9-8b73-d7ef4d75dfb6 · inbound
The State of Multilingual LLM Safety Research: From Measuring the Language Gap to Mitigating It Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75c56f71-475a-42c5-8fca-4ab6831a08d1 · inbound
Risks of AI-driven product development and strategies for their mitigation Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ffa341c-45b3-4e5d-9adc-c2fa387f1b01 · inbound
Linear Spatial World Models Emerge in Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14fbc27-e4c3-4582-9af6-2e467cb1714f · inbound
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e62633f9-e83d-4a72-b74f-41265a3c4fd3 · inbound
We Should Identify and Mitigate Third-Party Safety Risks in MCP-Powered Agent Systems Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28c35c6b-b441-4b2e-a413-138b5172816b · inbound
Probing the Robustness of Large Language Models Safety to Latent Perturbations Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 555ad9b2-e220-459a-b447-9abe1b5f31e0 · inbound
Agent-to-Agent Theory of Mind: Testing Interlocutor Awareness among Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d718aa0a-df81-4905-89ce-89b731533a5d · inbound
Deprecating Benchmarks: Criteria and Framework Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e01fb38-ff2f-4f15-9450-db40c886c0bf · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 292
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2e208163-7343-40e9-b35a-595ddf22ab0f · inbound
Against racing to AGI: Cooperation, deterrence, and catastrophic risks Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 372bda71-ed0d-4395-82e4-c3299076289b · inbound
Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0f6b72d-cb2b-44a1-a86b-39d5e21a3c38 · inbound
Beyond Solving Math Quiz: Evaluating the Ability of Large Reasoning Models to Ask for Information Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecc3fec7-e818-46c6-b563-51e46ebf87d9 · inbound
Large Language Models for Next-Generation Wireless Network Management: A Survey and Tutorial Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8c68cc1-77b3-4e49-af8b-3d246cf85831 · inbound
Scheming Ability in LLM-to-LLM Strategic Interactions Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c691439b-6c02-4607-97ce-bce0e5225742 · inbound
Value Drifts: Tracing Value Alignment During LLM Post-Training Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6777e281-4e8f-406e-9a6a-71c8f5b6d78c · inbound
Phantom Transfer: Data Poisoning can Survive Data-Level Defences Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41cc4a35-8e10-49c1-a665-dadfb8cdaed7 · inbound
Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4bc4ceda-8734-4625-addb-79961c6c573b · inbound
NEST: Nascent Encoded Steganographic Thoughts Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21fcd595-f29d-4fd0-8931-13f04dc59457 · inbound
BarrierSteer: LLM Safety via Learning Barrier Steering Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f1f45249-5e71-4b83-be97-9d66264b82b8 · inbound
Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 228
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ffd801da-ca43-4ed6-b6f6-99f976febded · inbound
Belief or Circuitry? Causal Evidence for In-Context Graph Learning Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2bcdda77-c8c6-4b91-8ba4-782e7a398082 · inbound
Interpretability Can Be Actionable Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cf4c3b87-5ed5-45d5-b964-48f9624777db · inbound
Stories in Space: In-Context Learning Trajectories in Conceptual Belief Space Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 144
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e4284353-edd1-4007-8b40-4521a20c3b61 · inbound
Prediction-Powered Inference Across Many Tasks for AI Evaluation & Social Science Research Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b7dfc7f5-81bb-491b-ae4d-bf9f74b8ffa8 · inbound
Dissociative Identity: Language Model Agents Lack Grounding for Reputation Mechanisms Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ca044699-010f-4031-944a-97e2aeaea091 · inbound
The Surface You Test Is Not the Surface That Breaks Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ee43106b-6bdd-435c-b545-49b57857f71f · inbound
VET: A Framework for Analyzing AI Discourse Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 166
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation fe12e385-1113-407c-b812-b8418590317b · inbound
Consistency Training while Mitigating Obfuscation via Rate Matching Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2e78d5e7-5ce1-4849-b17c-6a011716ecbf · inbound
Black-box, Adaptive, Efficient, Transferable, Harmful, Applicable... Attacks Are All You Need to Break LLMs Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 46a39709-f2b5-41aa-892b-53ffb6583773 · inbound
A Geometric View for Understanding Concept Learning and Neuron Interpretation in Sparse Autoencoders Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dcbdf9e3-9228-4c45-bb78-1a1947a62269 · inbound
Auditing Proprietary Alignment in Large Language Models: A Comparative Framework Without a Ground-Truth Standard Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5eb8546a-7b7b-4288-9804-0fd91738197e · inbound
Distilling Safe LLM Systems via Soft Prompts for On Device Settings Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e2cc9341-e13b-4344-aebc-f97c301d251d · inbound
Investigating The Security of Modern AI and Cloud Infrastructure Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2811f24f-910a-4c3e-b60d-9ea0d54caf92 · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 580bb6d7-6b73-421a-b14f-02db6794d56d · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 078c6946-2dbb-4e66-bc77-84bc0ee45a14 · inbound
(Towards) Scalable Reliable Automated Evaluation with Large Language Models Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b30a7b2d-a07f-4801-bbb5-93ca6fd4e08b · inbound
Taxonomy-Driven Analysis of Open-Source AI Risk Mitigation Tools Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.