Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 45 inbound Pith citation observations for arXiv:2312.04724.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:33:32.868576Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
18
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation d71ed716-aa49-42ae-a6c4-491f44c60ad0 · inbound
StarCoder 2 and The Stack v2: The Next Generation Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 164
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 4299a779-96e4-4419-9105-def9b99382a0 · inbound
Precision or Peril: A PoC of Python Code Quality from Quantized Large Language Models Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8ff9165e-9e70-4832-be1e-67814bb5b3d5 · inbound
CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f303afb9-da19-4460-9313-d870f3e455bd · inbound
Humanity's Last Exam Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 301c7b7c-519f-472b-9145-8a07f717f1db · inbound
LLMSecConfig: An LLM-Based Approach for Fixing Software Container Misconfigurations Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e89f275-13dc-49a2-8358-52d8b34ab118 · inbound
Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 259fba11-9ee5-4183-a195-facc5c07f03b · inbound
Training Language Models to Generate Quality Code with Program Analysis Feedback Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f606523-bef9-46e5-b411-10bf0bb1bb23 · inbound
MCP Safety Training: Learning to Refuse Falsely Benign MCP Exploits using Improved Preference Alignment Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a740e82-3223-4dfb-9470-ccff15569fbc · inbound
Developing a Risk Identification Framework for Foundation Model Uses Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e5f17a6-a27d-47b9-b7eb-835e73519e26 · inbound
SCGAgent: Recreating the Benefits of Reasoning Models for Secure Code Generation with Agentic Workflows Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78188daa-cfd8-411e-b556-0bf09ef0c4f5 · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 381c30dd-e45d-4868-af24-85cbb8ecd9f9 · inbound
Guiding AI to Fix Its Own Flaws: An Empirical Study on LLM-Driven Secure Code Generation Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf6837e0-a1cf-456c-b41d-df3c068a6907 · inbound
MGC: A Compiler Framework Exploiting Compositional Blindness in Aligned LLMs for Malware Generation Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97794215-ada6-49f4-9f3e-40a02bb0b665 · inbound
Understanding the Supply Chain and Risks of Large Language Model Applications Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dfe7ef9-3eb8-4fa5-8515-d47d7e71faec · inbound
Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24a861b7-a4e2-4ef7-8c4e-c1211673c999 · inbound
Towards terahertz nanomechanics Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75bcb548-44e5-45d4-8578-254e6d951780 · inbound
ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f80b99b-ddea-4efa-bee6-62a473fe1d3f · inbound
Secure Code Generation at Scale with Reflexion Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 375b5ec1-aada-443f-8522-c7944673692a · inbound
BEAVER: An Efficient Deterministic LLM Verifier Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 12076545-9314-47ab-8e36-96d721f8ed1e · inbound
Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c6d0de2-b199-4885-869e-e2348c008998 · inbound
"Tab, Tab, Bug": Security Pitfalls of Next Edit Suggestions in AI-Integrated IDEs Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 94d15de5-8ac2-4d03-9fd8-6d3157135cbb · inbound
Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation bebf90b3-ee10-42c9-9deb-daf54a51f640 · inbound
Large Language Models Generate Harmful Responses Using a Distinct Mechanism, Shared Across Harm Types Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 7fbbb1e4-aa8e-49f4-b162-b551afa5aa3e · inbound
Adversarial Arena: Crowdsourcing Data Generation through Interactive Competition Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 2c5d9f52-8c89-4d6e-9660-35e553710ce1 · inbound
Towards Optimal Agentic Architectures for Offensive Security Tasks Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3eda2f9c-237c-4b36-803e-1bdc8981c0d5 · inbound
A Validated Prompt Bank for Malicious Code Generation: Separating Executable Weapons from Security Knowledge in 1,554 Consensus-Labeled Prompts Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 89af28ed-bcf5-40d3-8611-5630b43ad2ef · inbound
Redefining AI Red Teaming in the Agentic Era: From Weeks to Hours Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 15bad8b2-2d34-44a3-928e-b4a531ea6e94 · inbound
Bridging Generation and Training: A Systematic Review of Quality Issues in LLMs for Code Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 593fda4e-e88b-4287-a655-07e4b936771f · inbound
SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation d2708aee-a774-4683-95be-2aeb7caa45d7 · inbound
LLM-Agnostic Semantic Representation Attack Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation de0b055b-e787-4943-bf8c-d0ea299bb904 · inbound
Ablating Safety: Mechanisms for Removing Alignment in Language Models for Security Applications Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e9453341-ea23-434c-9bd0-857ce0aca0b2 · inbound
Refusal Evaluation in Coding LLMs and Code Agents: A Systematic Review of Thirteen Malicious-Code Prompt Corpora (2023-2025) Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 014b4ae9-6471-42c6-9e79-be13ffc29e50 · inbound
HIDBench: Benchmarking Large Language Models for Host-Based Intrusion Detection Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 07a96601-9bbd-4e36-bdfd-14c9a7bae55b · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 50142792-7fa7-41bc-bfa7-d3d3e189d660 · inbound
Boiling the Frog: A Multi-Turn Benchmark for Agentic Safety Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 8497cf12-ab6e-4c82-9ac5-566da15afb6a · inbound
Security of LLM-generated Code: A Comparative Analysis Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0e47e741-ab83-4a5a-aa3c-4c83b18a6e4d · inbound
SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e7a4f50f-7229-44a5-a12a-25664d9e796a · inbound
Helpful or Harmful? Evaluating LLM-Assisted Vulnerability Patching via a Human Study Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 9d048307-d19b-422f-a380-eb09b9fd25ee · inbound
Direct Causation in International Humanitarian Law and the Challenge of AI-Mediated Civilian Cyber Operations Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f2caab62-2233-4b1d-a14f-600eca1074a7 · inbound
Not All Refusals Are Equal: How Safety Alignment Fails Cybersecurity at Scale Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d6179b5-8f42-4710-af49-7b22b97a1710 · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4b6b2a4-983a-46fb-a945-a3a224793c59 · inbound
Refused in Chat, Written in Code: Workflow-Level Jailbreak Construction in IDE Coding Agents Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 237e36ec-cebd-45cf-9c43-8a1d2b584881 · inbound
Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2eb78e9-4e54-4e10-a6a0-07da0442523b · inbound
Cyber-Capable AI Agents: Vulnerabilities, Evaluation Containment, and Defensive Response Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cdc2bd7-d8d6-4bb1-8996-da6ea68649a7 · inbound
Trident : How to Break Deep Reinforcement Learning Cyber Defenses (Agentic) Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.