Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2407.17436.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T10:39:08.125748Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 32bb5059-d424-47b3-a16e-80c5ac345f4a · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 242
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25b02937-0332-4627-a663-341e2264d9ac · inbound
SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 230
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd6579a4-a8d8-4ddf-9357-46a11e763ff8 · inbound
OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 93a34e3b-59e6-4fa6-92b9-1d52fe91019f · inbound
Seed1.8 Model Card: Towards Generalized Real-World Agency AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f3818a60-a2e4-466c-80fa-cf4cc7520b4a · inbound
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ad4d6a5d-b8e0-490c-9650-c701483c320c · inbound
Beyond Fixed Benchmarks and Worst-Case Attacks: Dynamic Boundary Evaluation for Language Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7ebd75bf-b23f-4b84-97de-3bd1b6993d3f · inbound
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a768578b-b939-42c5-a201-531f761575b0 · inbound
ROK-FORTRESS: Measuring the Effect of Geopolitical Transcreation for National Security and Public Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 361f769d-7269-4c3e-8ecf-29bfd8626531 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5671be64-514b-49dc-a6b5-e5b8aac7edb1 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4b7a78d5-fc87-4552-9dfc-399aa26317b1 · inbound
SafeGen-Bench: Benchmarking Safety in Image-Conditioned Text-to-Video Generation AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d8ecb8ee-7117-4d70-ba5a-7ac8bfe58c4c · inbound
Safety Measurements for Fine-tuned LLMs Should be Grounded in Capability AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a49d69b0-35de-48a1-93c7-5d5d442eaa76 · inbound
Beyond Single-Policy: Evaluating Composed Organization-Specific Policy Alignment in LLM Chatbots AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 821c38c6-b39a-4790-b855-e6e2e115e6e1 · inbound
RiskNet: A large-scale dataset of AI risk incidents from news with alignment and multi-dimensional annotations AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cd4ad914-2d01-4bd9-8315-73e47c8324af · inbound
Culturally-Adapted Red-Teaming Across East and Southeast Asian Contexts: A Methodological and Comparative Analysis AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b466eda9-ecf2-4ae4-9d68-1757f587d565 · inbound
FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3eccc86e-7846-4ba5-a478-6d7a2ec44d1e · inbound
Efficient Safety Benchmarking via Item Response Theory AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7c3bdb28-b686-421f-ac4d-5bb66a412136 · inbound
How Jailbreak Attacks Inform Safety Alignment: A Defender-Centric, Shapley-Based Evaluation of Jailbreak Contributions AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76cd7933-fe44-421e-b840-c1cef0f98e17 · inbound
AIR-BENCH Live: An Evolving Safety Benchmark for Foundation Models AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a2bfb26-0fbe-4c53-955a-2c7da7224d28 · inbound
Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.