Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2311.08370.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T23:33:40.319869Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:59:32.652255Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 4f5f243a-d99e-431f-9ae2-b485bdeb8495 · inbound
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f83ef09f-72de-4e00-92e2-0405adcdb8d1 · inbound
Fine-Tuning Lowers Safety and Disrupts Evaluation Consistency SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09a1df70-ccc8-4ba9-9cf1-87fe13daffc1 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 290
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd6466c5-a3a4-419c-afb5-52a2f4965a5b · inbound
Scaling behavior of large language models in emotional safety classification across sizes and tasks SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7923d2dd-4db1-40a6-ae09-5a6959bf521e · inbound
YuFeng-XGuard: A Reasoning-Centric, Interpretable, and Flexible Guardrail Model for Large Language Models SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a269938-0b2e-41e3-89cb-813df70be459 · inbound
Predict, Don't React: Value-Based Safety Forecasting for LLM Streaming SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 541b6e26-2efd-412e-afee-8da6d76b1fcd · inbound
Blind Refusal: Language Models Refuse to Help Users Evade Unjust, Absurd, and Illegitimate Rules SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cc973d17-47c9-489b-82d9-1ab742d3aec2 · inbound
Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3e4d273c-0bd7-403f-bafe-ec9e1246d42a · inbound
LLM Safety From Within: Detecting Harmful Content with Internal Representations SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9ea79794-54cc-4934-974d-0f9707be9d06 · inbound
Chain of Risk: Safety Failures in Large Reasoning Models and Mitigation via Adaptive Multi-Principle Steering SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bb71f8f2-96a5-477f-90a6-42c06ca75125 · inbound
How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b388b35e-96a5-4604-a763-b991bc3f650c · inbound
How Many Iterations to Jailbreak? Dynamic Budget Allocation for Multi-Turn LLM Evaluation SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93da4356-bb7f-465f-8d4c-42dab9e7f2f1 · inbound
GLiGuard: Schema-Conditioned Classification for LLM Safeguard SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 33dfac57-ee03-403b-82b1-6f1f32950cde · inbound
Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d56d2ff7-7585-4262-bde0-e4d5dfb61862 · inbound
Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 127
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 98611d6b-ce7f-4d3c-a713-58155635b318 · inbound
Opir: Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 038c79bb-5e1e-48f9-99d3-ef3b851d43f2 · inbound
Beyond Safe Data: Pretraining-Stage Alignment with Regular Safety Reflection SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3136f6e0-4c54-4f1d-9371-5d9978f37405 · inbound
CATCH-ME if you RAG: a dataset of Contextually Annotated multi-Turn Counterspeech against Hate and Misinformation Exchanges SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 138
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7cfaa08-b924-4810-9d1c-e6f39525d57e · inbound
Efficient Safety Benchmarking via Item Response Theory SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b6c5d441-8210-43b1-b07c-0172b1668320 · inbound
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 88e4652b-bc78-4315-926b-f45ec0be14ba · inbound
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 431a847b-be93-46c0-b3b2-f91be2408787 · inbound
When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard: A Fast Vision-Language Guard for Real-Time Moderation SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f208a3a-3e98-4988-bd5a-a9a9a43da872 · inbound
Semalith v1.4: A Calibrated 184M Safety Classifier Achieving State-of-the-Art Prompt-Injection Detection at 44x Fewer Parameters than Llama-Guard-3-8B SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.