Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2308.13387.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-09T13:14:34.116590Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
4
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 49da06db-1474-4a09-96ad-1dec12bda0b5 · inbound
Low-Resource Languages Jailbreak GPT-4 Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d214990f-e87f-4a7c-bd7c-551b0ac587d2 · inbound
DeepSeek LLM: Scaling Open-Source Language Models with Longtermism Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 151
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f80178cc-aec5-402b-b6c5-a92bbc41ba05 · inbound
TrustLLM: Trustworthiness in Large Language Models Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dbae861b-3cd9-4b8d-ad84-67e64f05caba · inbound
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4b290f1a-438c-44dc-a7ec-95464f90b851 · inbound
Jailbreak Attacks and Defenses Against Large Language Models: A Survey Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aebf8f1a-3c1c-4fe8-a5c7-c90c725ec6d2 · inbound
Vulnerability Mitigation for Safety-Aligned Language Models via Debiasing Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 025fa25f-3b8a-4744-8796-914ab383d9b6 · inbound
Safety Reasoning with Guidelines Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8dea832e-a657-4a75-9287-1323bcb16dd5 · inbound
SweEval: Do LLMs Really Swear? A Safety Benchmark for Testing Limits for Enterprise Use Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4e6b0ea-e62c-4c59-a12d-910c41a6b0cb · inbound
Xinyu AI Search: Enhanced Relevance and Comprehensive Results with Rich Answer Presentations Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d3a899b-48da-4516-b36b-31962dfbab41 · inbound
Beyond the Surface: Measuring Self-Preference in LLM Judgments Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a77cdee8-0522-41fa-84d8-875e81486a24 · inbound
Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 234
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0c3a4973-30b9-47b5-9a48-97af1325e9d3 · inbound
SEALGuard: Safeguarding the Multilingual Conversations in Southeast Asian Languages for LLM Software Systems Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e3d5fd9-d925-426f-aeed-56809efd5295 · inbound
Guardians and Offenders: A Survey on Harmful Content Generation and Safety Mitigation of LLM Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 124
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dde2c32f-5be4-4115-a64b-ff810cd02d84 · inbound
Revisiting Robustness for LLM Safety Alignment via Selective Geometry Control Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aff94550-37b0-4bdf-8fcb-7c74bfd3e047 · inbound
Cooking Up Risks: Benchmarking and Reducing Food Safety Risks in Large Language Models Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation c4abdfe4-cb30-48df-a7cd-305efd06f8d7 · inbound
Guardian-as-an-Advisor: Advancing Next-Generation Guardian Models for Trustworthy LLMs Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 931bb8ab-2a49-4933-bdbc-cea2a466d8e6 · inbound
The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 329b1e73-7a7a-4401-ac4b-213d93b31f5d · inbound
Dialect vs Demographics: Quantifying LLM Bias from Implicit Linguistic Signals vs. Explicit User Profiles Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 67893cbf-c80c-41c7-b2d2-f3776cc73794 · inbound
MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6bacf8e7-54c3-4907-94b3-e089133574ce · inbound
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation b4835099-28c6-415c-8219-aaf1e81b83c2 · inbound
Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8189446b-33c3-4263-aab8-515445b7585c · inbound
Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1123835b-527d-428a-a568-12d987057cf5 · inbound
REFLECTOR: Internalizing Step-wise Reflection against Indirect Jailbreak Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 41351988-a409-43be-b92d-e6ba900d5688 · inbound
JT-SAFE-V2: Safety-by-Design Foundation Model with World-Context Data Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation db495455-5492-4c3a-8908-647e36b023b8 · inbound
RouteScan: A Non-Intrusive Approach to Auditing MoE LLMs Safety via Expert Routing Telemetry Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7bafba10-78c1-43df-945c-79fc7fd30192 · inbound
IDEAFix: Evaluation Framework for Creative Defixation Prompting in LLMs Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8a5fb43b-54f1-4e5c-a1d3-f03a2f75e43e · inbound
When Behavioral Safety Evaluation Fails: A Representation-Level Perspective Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d8086c03-e6eb-4e6f-8c34-302b925b7b0b · inbound
Unsupervised Causal Abstractions Discovery Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 96db16a7-6e3b-479b-81d1-551b69abd473 · inbound
FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation aaa08779-a722-40ec-ae7e-ba503da2235b · inbound
Efficient Safety Benchmarking via Item Response Theory Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e7a258d6-01cd-4f4c-b84e-22a31e3655f6 · inbound
Discriminatory Compliance: How LLMs Answer Queries from Protected Groups Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ea5bf22f-5b06-4ed8-9db0-30d635e15b83 · inbound
Do Encoders Suffice? A Systematic Comparison of Encoder and Decoder Safety Judges for LLM Adversarial Evaluation Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f8ddaa2-f460-4ba9-9473-dbf00df5c8e5 · inbound
Symbolic Mechanistic Data Attribution: Tracing Training Influence to Learned Behavioral Policies Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 166868af-f1aa-4a2d-bc06-e82283a56383 · inbound
BioTIER: A Refusal Benchmark for Targeted Biological Risk Mitigation Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc180841-96ec-4227-8e0e-e3a769d177a1 · inbound
Reason Before You Retrieve: Agentic Planning for Multi-modal RAG Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 106
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f60688c9-3ec5-4555-84be-6cf231031183 · inbound
The Mirage of LLM Guardrails: A Case Study in AI-Assisted Medical Note Manipulation Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ebea8a40-8b99-4396-8cd8-b56c0929cae9 · inbound
Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e98565-810a-4f74-8e17-bd34845b0412 · inbound
Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52acaad3-523c-437f-b54c-ec44f120bc2a · inbound
Item Response Theory for AI Safety Do-Not-Answer: A Dataset for Evaluating Safeguards in LLMs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.