Pith. sign in

Paper Citation Record · LEDGER

SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2506.05692.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05692 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:29:57.375647Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T04:09:34.777136Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bd32eeba-1950-480b-aa1a-c8119ef59ad0 · inbound

Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software cites this paper.

Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T05:18:43.734775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:18:43.734775Z digest=sha256:474963922754c36e9ba438f6856ca401960f557b3fcbff4f5b1be63694e7d1f7

Observation 80fe48cd-4c88-4d9f-b7c7-86f8fcf6fa8e · inbound

False Security Confidence in Benign LLM Code Generation cites this paper.

False Security Confidence in Benign LLM Code Generation SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:21:26.746518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T06:20:52.774912Z digest=sha256:080fb12882148f694455e0b2684b9b42675412c2841b53cdf8197b9d2464b1ad

Observation 90c05cd9-b579-4d74-b9d7-94e4ee5fbf69 · inbound

Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning cites this paper.

Taint-Style Vulnerability Detection and Confirmation for Node.js Packages Using LLM Agent Reasoning SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-10T00:59:49.633068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T00:56:42.767929Z digest=sha256:341846bfa3c9e7e68c4228e232cac00cda05d85102aa9b9c83e592b892244fe1

Observation 547f6267-2fff-48c6-a572-25ea7a88cb65 · inbound

On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies cites this paper.

On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:26:09.981257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-08T09:12:15.486780Z digest=sha256:30038be3b496afad1af720b2b30e64ccbb15a148bac7d3e259ae42be49ac8afb

Observation a3915f51-e380-45ed-95b0-7fae6f255723 · inbound

CoT-Guard: Small Models for Strong Monitoring cites this paper.

CoT-Guard: Small Models for Strong Monitoring SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T19:39:23.772865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T19:39:16.264272Z digest=sha256:bc93ff396c153fa97316b2add184164dbb96e799a93e55f06ad120197fa16048

Observation 4fd05067-2171-419b-aed8-656c01488385 · inbound

ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage cites this paper.

ASSEMBLAGE-DEEPHISTORY: A Cross-Build Binary Dataset with Temporal Coverage SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-22T09:41:21.765966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T09:40:21.746161Z digest=sha256:89830f1e5e452df7bb8747a1d5bfbd536ac23e2578a5fed6a524842c73a28e93

Observation 4da95db5-56eb-40e7-b1fa-eb2d852ed576 · inbound

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming cites this paper.

FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T04:09:34.779477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T17:11:40.088809Z digest=sha256:2ae524ee92084fed2e819f200b46338d90f584a2d5eb74456571c3a9ee1e9193

Observation 9391a7b8-e6cf-48a0-8c26-f446b61fa838 · inbound

An Empirical Study of Security Calibration in Large Language Models for Code cites this paper.

An Empirical Study of Security Calibration in Large Language Models for Code SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:45:43.045640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T05:04:41.223743Z digest=sha256:5a3da6089f1b3cbbb6438252477cb54c89a19a714354e8b9a30cfdf500eb58c4

Observation 75240d8a-1d5d-40ba-b438-c96b1600794c · inbound

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code cites this paper.

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T09:44:42.646014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T09:44:42.646014Z digest=sha256:5a99d86064cdeb85d09f21a4dc77c9633547faf76e1463bca210a157137b4d53

Observation 91d0a92e-fce5-4d73-85f0-61222953f4a8 · inbound

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code cites this paper.

The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T04:38:06.679657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:38:06.679657Z digest=sha256:8bc79a7372df9ea066b25db6f9db003bacbddc01339704e1807a4305fc27bccc

Observation c8762db4-355e-454e-90e2-635ce74067b8 · inbound

The Patchwork Problem in LLM-Generated Code cites this paper.

The Patchwork Problem in LLM-Generated Code SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T01:16:30.101870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T01:16:30.101870Z digest=sha256:458df49463365d91b1194bf796ec5df51ee1c815708c32c2570cb9ef787e25c3

Observation b212ebb8-7ab2-4f5b-a7ae-95296fd38590 · inbound

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios cites this paper.

Poster: Rethinking Security in LLM Code Generation through Real-World Risk Scenarios SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-01T03:40:38.189512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T03:40:38.189512Z digest=sha256:60d5b46e477502f5f7ed35adafc94fbcdfa65e37ef2b65a6dc9e4e826805f75b

Observation c6ca3504-9d98-4ad8-8114-0bcf1183b914 · inbound

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation cites this paper.

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T00:29:57.375647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:29:57.375647Z digest=sha256:d74ae94eb5274305b0617ebeeaca1993862dbb367efe133792112d865bdc34ca