Pith. sign in

Paper Citation Record · LEDGER

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases

As of 14 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2608.07776.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07776 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:17:02.693029Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact2
  • verified fuzzy17
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation de1814ca-d3a9-452e-8bde-aa0b9b0da5d3 · outbound

This paper cites Asleep at the Keyboard? Assessing the Secu- rity of GitHub Copilot’s Code Contributions,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Asleep at the Keyboard? Assessing the Secu- rity of GitHub Copilot’s Code Contributions,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.127782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.591829Z digest=sha256:79e3e23e38d760d4fb838de3497f2e595c2d9f9d3afdb911f905c31f173dfce4

Observation 8a3f4160-bcc3-418a-84ea-0e0553b79503 · outbound

This paper cites Do Users Write More Insecure Code with AI Assistants?.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Do Users Write More Insecure Code with AI Assistants?

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.114919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.596517Z digest=sha256:0629507c2aaedfd692b2095b2f3e21595e843da14cb011d1644ba84cf3de5043

Observation 576b1c4f-1e25-49c1-821b-214628929bff · outbound

This paper cites Lost at C: A User Study on the Secu- rity Implications of Large Language Model Code Assis- tants,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Lost at C: A User Study on the Secu- rity Implications of Large Language Model Code Assis- tants,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.097596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.600703Z digest=sha256:b82d477ad6488ef47a5cc5af5470ceb165fe17aa31fe1b65496fef6e8a1e67a2

Observation 08e42133-959e-4cc6-8186-4a664a0bed7c · outbound

This paper cites LLMSecEval: A Dataset of Natural Language Prompts for Security Evaluations,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases LLMSecEval: A Dataset of Natural Language Prompts for Security Evaluations,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.086302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.604734Z digest=sha256:5c9607197fb92e490f65ea788883ca27a6655e36b4714250184e028d32b5d700

Observation cb069cce-fc0e-43ed-826d-3cd11c23ced6 · outbound

This paper cites SecurityEval Dataset: Mining Vulnerability Examples to Evaluate Machine Learning-Based Code Generation Techniques,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases SecurityEval Dataset: Mining Vulnerability Examples to Evaluate Machine Learning-Based Code Generation Techniques,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.075181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.608726Z digest=sha256:97631d7f98c17b6d9d983559ee20f8a44ce731fe101b4397dfdd773877411a4c

Observation cb5446c9-1651-4e8b-85c4-84b83507b890 · outbound

This paper cites SALLM: Security Assessment of Generated Code.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases SALLM: Security Assessment of Generated Code

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.612913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.612913Z digest=sha256:1e6c2ddc334be86313e766fa3c2447e37023406dbf2fb011861cc028f392967f

Observation a874bdae-d925-45c1-9d81-71beefb0d23a · outbound

This paper cites CodeLMSec Benchmark: Systematically Eval- uating and Finding Security Vulnerabilities in Black-Box Code Language Models,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases CodeLMSec Benchmark: Systematically Eval- uating and Finding Security Vulnerabilities in Black-Box Code Language Models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.064829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.617635Z digest=sha256:e055ae85c6fd4ad09198c4ee3e6a60d64a038a7f6396ea958a85f6327bdecb41

Observation 158d1fab-94aa-4335-86c7-be2f50db70f3 · outbound

This paper cites Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Purple Llama CyberSecEval: A Secure Coding Benchmark for Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.621342Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.621342Z digest=sha256:87c3907e045f3bada7373d01fb9d0d0ea2bd50141f4bc36f0dc5f4e0ce3202eb

Observation 8e65020e-9877-44ec-9f3c-fd501a302fee · outbound

This paper cites CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases CWEval: Outcome-driven Evaluation on Functionality and Security of LLM Code Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.625355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.625355Z digest=sha256:ea0039b3d8fed17c6050185d52e2b2ab3b9a50b7f9089970564ad1b00e873564

Observation 33176de4-a6f0-4c6f-b7e6-27424973ee9f · outbound

This paper cites Prompting Techniques for Secure Code Generation: A Systematic Investigation.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Prompting Techniques for Secure Code Generation: A Systematic Investigation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.629367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.629367Z digest=sha256:19163211d4a0a35ac7c521cae340bbeff6727fa695bf73b60bf6951f706a004d

Observation 4a141209-5c76-4eed-9920-2784f3840aa6 · outbound

This paper cites Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Benchmarking Prompt Engineering Techniques for Secure Code Generation with GPT Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-11T04:17:02.855913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.633989Z digest=sha256:e3e23ba6adde067650ac971b9745d7174a4f2df38c2b996030434962b9ccf346

Observation f31bfd67-f09e-42ff-86c0-bdadcda6ce27 · outbound

This paper cites How Secure is Code Generated by ChatGPT?.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases How Secure is Code Generated by ChatGPT?

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.053971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.638113Z digest=sha256:180ef9d77c18fc13f9926083078ac94d571cb1300070e0ad3b91a4c2e021ce95

Observation e02bb8c1-d777-4d80-bbf6-aa1c3a5f54a2 · outbound

This paper cites Judging LLM-as-a-Judge with MT- Bench and Chatbot Arena,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Judging LLM-as-a-Judge with MT- Bench and Chatbot Arena,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.041879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.641707Z digest=sha256:8357f7978b7fbdd546f696fc7144fe511a796e402cd6045527fa8ab56e40d1c1

Observation f9eaad64-d1d5-43a7-a31d-c1d65a944fcd · outbound

This paper cites G-Eval: NLG Evaluation using GPT-4 with Better Hu- man Alignment,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases G-Eval: NLG Evaluation using GPT-4 with Better Hu- man Alignment,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.029917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.645527Z digest=sha256:ee8dfab6699628713d96f0bf7ea8539947dc22014bb9843dd20eb12e060f3a20

Observation 8adef996-bee4-46f3-bc8a-4a884abc126c · outbound

This paper cites Large Language Models are not Fair Evaluators.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Large Language Models are not Fair Evaluators

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.649319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.649319Z digest=sha256:a3d0e8a333a9a2e4f080d16b93664b7a75f9146c93b8579121110f4a64c08bdc

Observation d5bdf449-c31f-465a-9d47-52d7df05c8b8 · outbound

This paper cites How Secure is Secure Code Generation? Ad- versarial Prompts Put LLM Defenses to the Test,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases How Secure is Secure Code Generation? Ad- versarial Prompts Put LLM Defenses to the Test,

Reference 16

Resolution
verified exact
raw_fallback, observed 2026-08-11T04:17:02.826409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.653497Z digest=sha256:c8f1e226ad4cb0b178118641c7ab7cf94bfc171b33b58be5e9f2e9dc7cbcb7fb

Observation 7c355cd8-7397-412f-b000-01362f457e3f · outbound

This paper cites Why Don’t Software Developers Use Static Analysis Tools to Find Bugs?.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Why Don’t Software Developers Use Static Analysis Tools to Find Bugs?

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.015745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.657190Z digest=sha256:1e834ea52abd3cfd1ac3263a02fba47482a8a820f75d46bf2328ad4d5e36bf2d

Observation 23720a4c-e661-42bf-ac00-4b7c64fda719 · outbound

This paper cites A Few Billion Lines of Code Later: Using Static Analysis to Find Bugs in the Real World,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases A Few Billion Lines of Code Later: Using Static Analysis to Find Bugs in the Real World,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:03.002174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.660773Z digest=sha256:208c1110bc4627b23416b1db895a93ce2e6bbf4b326a064ce84291c60c4bca2b

Observation 3f6868b5-5210-48a9-8a0e-985a8acffc9b · outbound

This paper cites A Coefficient of Agreement for Nominal Scales,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases A Coefficient of Agreement for Nominal Scales,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.988312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.664726Z digest=sha256:27b179feef2a88f1213d832e61c67172420fc8fb0c4e54d0dd644cf2af914d91

Observation 16352237-93f9-4479-996a-e48b9408e220 · outbound

This paper cites Checkov: policy-as- code static analysis for infrastructure as code,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Checkov: policy-as- code static analysis for infrastructure as code,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.974374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.668388Z digest=sha256:6447a49cc63e63dcc1d75607bc83843e9d62388b542511f38bb90b6671a39458

Observation 8c1e08f3-530d-45ef-bec3-66118f495ec2 · outbound

This paper cites tfsec: security scanner for Terraform code,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases tfsec: security scanner for Terraform code,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.963251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.672203Z digest=sha256:f5402f8c6b25f6f9bf3f40c3c1b3ff4b593ac2c8a0da9a570407f44ca84b5ded

Observation 062ffe4b-d578-4a21-b1cc-e5479ac34457 · outbound

This paper cites OPA / Rego policy en- gine,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases OPA / Rego policy en- gine,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.950493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.676062Z digest=sha256:8e368e41b3967572a121f487f25cf419a95b44ee253ca3f8a9da42e996eb4f93

Observation ce510f33-eb7a-4238-85e1-9dac1e920a49 · outbound

This paper cites Large Language Models for Code: Security Hardening and Adversarial Testing,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Large Language Models for Code: Security Hardening and Adversarial Testing,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.934682Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.681251Z digest=sha256:42ce4cfbf6662262f207a4de0087a37b2b28b6f647bbb9cf7977f1425b1e6678

Observation ceee5fb5-2ea0-4a91-a177-9d70def2ff01 · outbound

This paper cites Instruction Tuning for Secure Code Generation.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Instruction Tuning for Secure Code Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.685067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.685067Z digest=sha256:41d87dd23f71aec13ce56fcb5209ea9faf688849624ee1723fc863f01a1adaf1

Observation 3b25a36a-1c4a-4f9f-b882-3d478bf15978 · outbound

This paper cites Security Weaknesses of Copilot-Generated Code in GitHub Projects: An Empirical Study.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Security Weaknesses of Copilot-Generated Code in GitHub Projects: An Empirical Study

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T04:17:02.689121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:17:02.689121Z digest=sha256:35793fa384df60cfe396dbfc01a0e991132e9a5bbecfcf3beda1de297e420b35

Observation cfc2d66e-1ef4-4517-ab80-e5f84fbab76e · outbound

This paper cites Lessons from Building Static Analysis Tools at Google,.

Can AI Write Compliant Code, and to What Extent? Evaluating SOC 2 Compliance of Claude Fable 5, Claude Opus 4.8, and Claude Opus 5 Across Four Use Cases Lessons from Building Static Analysis Tools at Google,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T04:17:02.921515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-11T04:17:02.693029Z digest=sha256:524681e5ff69f34aacf4dc3af565d2d984e92ef247f8b5ab50a27c1490dd19da

Pith citing papers

No inbound Pith citation observations are available.