Pith. sign in

Paper Citation Record · LEDGER

CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2501.08200.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.08200 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:10:30.472485Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:40:02.845349Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04dc944f-0a6a-4956-b526-84e6a7eb3d8a · inbound

XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants cites this paper.

XOXO: Stealthy Cross-Origin Context Poisoning Attacks against AI Coding Assistants CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-22T23:57:17.057288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T23:55:14.887237Z digest=sha256:50d35db90d9b7ab6c143d913698abbecde5720c5fe67e3cb6b6e4e6a8ec39194

Observation 56edde5f-2352-4f73-bd48-d5a899c4f206 · inbound

Teaching an Old LLM Secure Coding: Localized Preference Optimization on Distilled Preferences cites this paper.

Teaching an Old LLM Secure Coding: Localized Preference Optimization on Distilled Preferences CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 768

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:30.472485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:10:30.472485Z digest=sha256:0faee70df1bb0504d4373faecece1e9fcb1564e3d35dfd2de6862384cf6c293e

Observation 28da7e5f-0abb-41d8-9b7b-0d3af409a614 · inbound

SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code cites this paper.

SafeGenBench: A Benchmark Framework for Security Vulnerability Detection in LLM-Generated Code CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T10:22:12.808455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:22:12.808455Z digest=sha256:84c3ed9ed2a0d1c93eccfa17719ed9ffb4940c328c19b7acc0f055b0c5372a68

Observation 497e7f3b-bebb-42e6-9b67-9d7e4a6855ad · inbound

SCGAgent: Recreating the Benefits of Reasoning Models for Secure Code Generation with Agentic Workflows cites this paper.

SCGAgent: Recreating the Benefits of Reasoning Models for Secure Code Generation with Agentic Workflows CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:54.920034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:42:54.920034Z digest=sha256:9f5cc4556d9d7789c5a88a1a02415c0d054aba29470ec4b020b2336be91f123c

Observation 17f77bbc-21f9-4670-ab67-ab1744fc08c6 · inbound

Adversarial Attack Classification and Robustness Testing for Large Language Models for Code cites this paper.

Adversarial Attack Classification and Robustness Testing for Large Language Models for Code CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T05:28:48.272888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:28:48.272888Z digest=sha256:18540d83d1fb95cd7a1425ce5fedbfae10e8db02bd4a3cf3c5491f676a92aa27

Observation 970c6897-f15b-47b5-b8f9-42b23192b9c6 · inbound

AutoBaxBuilder: Bootstrapping Code Security Benchmarking cites this paper.

AutoBaxBuilder: Bootstrapping Code Security Benchmarking CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 3

Resolution
malformed identifier
arxiv_id, observed 2026-05-22T12:16:31.806244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T12:15:45.967861Z digest=sha256:f3081a9a05e9628157f9da7ea19e8870a4e0c13f05d6a74c8a4fa46684250d89

Observation b499969b-e70b-4065-90a3-1ad874e09844 · inbound

Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software cites this paper.

Extracting Recurring Vulnerabilities from Black-Box LLM-Generated Software CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T05:18:44.414943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:18:44.414943Z digest=sha256:8eb4a45d803c861bee3815a1a0355f030e282580d05108fbc525fc5c0e1ea8db

Observation ea40145c-09e4-4958-8e26-55e546e045ee · inbound

False Security Confidence in Benign LLM Code Generation cites this paper.

False Security Confidence in Benign LLM Code Generation CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:21:26.741830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T06:20:52.774912Z digest=sha256:120ad11fdb47c53ffb1ebe5266917e6a81dbe55e11385a9060839fa450142f83

Observation 68f89573-743c-4267-b3b1-75c45a31535c · inbound

Constrained Code Generation with Discrete Diffusion cites this paper.

Constrained Code Generation with Discrete Diffusion CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:27:47.909009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T21:24:10.211816Z digest=sha256:b068ef56de5dc237b616818c514f8e63040744a57f29cfbb9dde7e9a49831b10

Observation 64c52ece-aa3a-4ded-8255-cb99dcc484ec · inbound

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities cites this paper.

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:06:48.119951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T06:22:44.483045Z digest=sha256:15bb1f289fa0ed57a833073b2a3e4a3e6657806cedf79ba6799bf35cdbbd6fe2

Observation 55a94dff-5902-4d0f-a14c-e69dd910b58a · inbound

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities cites this paper.

CyberGym-E2E: Scalable Real-World Benchmark for AI Agents' End-to-End Cybersecurity Capabilities CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T12:28:25.132320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T12:28:25.132320Z digest=sha256:b47ac4e7c7f01a219f838cfc832c5774071216db5cb39f438eb4654e66ca8411

Observation c09d2929-ae56-4c12-b317-8a3c9ba71977 · inbound

SoK: AI Secure Code Generation: Progress, Pitfalls, and Paths Forward cites this paper.

SoK: AI Secure Code Generation: Progress, Pitfalls, and Paths Forward CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:40:02.846847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T22:41:15.200590Z digest=sha256:84202632e01fd0da3b8a66303f98e3dfd13b81e4c91b50dcf5d36d22235e0543