Pith. sign in

Paper Citation Record · LEDGER

VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2406.07595.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.07595 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T05:55:13.013887Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 95362f83-f121-4289-9324-e241f0d50311 · inbound

Shaping the Safety Boundaries: Understanding and Defending Against Jailbreaks in Large Language Models cites this paper.

Shaping the Safety Boundaries: Understanding and Defending Against Jailbreaks in Large Language Models VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T05:55:13.013887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T05:55:13.013887Z digest=sha256:4cff34b495890f6a66213bd0a07e78434f99b6024fcc4b655f17eb1e6966201e

Observation c566f049-a1ea-498e-9dcf-c46f97da6194 · inbound

Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" cites this paper.

Large Language Models for In-File Vulnerability Localization Can Be "Lost in the End" VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T17:26:43.470202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T17:26:43.470202Z digest=sha256:2c957996868c4449893f6243b8e6a3a8d1a5c4a586379957a1419bccb69d4425

Observation 1d5ac9fb-acaf-424c-a717-11fb4dd31a52 · inbound

LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights cites this paper.

LLMs in Software Security: A Survey of Vulnerability Detection Techniques and Insights VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-08T13:58:13.954423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:58:13.954423Z digest=sha256:5323f95a1db2c79e87fa3b762c1245ab4b5bb68f6c4cf11cbbefef283258ab2a

Observation b9fba813-ad88-44ac-b018-df23aa2bd02b · inbound

HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding cites this paper.

HiLDe: Intentional Code Generation via Human-in-the-Loop Decoding VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:03:33.764455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:03:33.764455Z digest=sha256:61577c5c04d5785681b01b1a671ce25b6eff3dad654481b003fd26a22ba3fc2f

Observation 19069f42-e98b-4e82-9d9b-095ac0e0c833 · inbound

Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond cites this paper.

Mono: Is Your "Clean" Vulnerability Dataset Really Solvable? Exposing and Trapping Undecidable Patches and Beyond VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:01:16.118464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:01:16.118464Z digest=sha256:13c9ca389393a6a144c2d128627432bbebc3fa95a96faa79ec6d4084323faf28

Observation 225aad1e-5c95-464b-9018-e735d5372cc9 · inbound

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild cites this paper.

QuiLL: An LLM-Based Vulnerability Assessment Framework for the Wild VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T11:01:17.182065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T10:56:28.973065Z digest=sha256:5cfa6d708b77e56d4b3642f7e22f46d73c45846f82a75d6fdefb700f2ccbd7ce

Observation 3bfc7b60-35f8-4e8e-8cbb-633859e1ff5a · inbound

How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection cites this paper.

How Code Representation Shapes False-Positive Dynamics in Cross-Language LLM Vulnerability Detection VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:00:12.651412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-07T05:58:06.190344Z digest=sha256:ed4b070f270f937caaa681430f917d81086be77e809755a38ab5033f1801ac76

Observation 02dd1c1b-5378-46f7-930b-191bc8e04dcd · inbound

Calibration Without Comprehension: Diagnosing the Limits of Fine-Tuning LLMs for Vulnerability Detection in Systems Software cites this paper.

Calibration Without Comprehension: Diagnosing the Limits of Fine-Tuning LLMs for Vulnerability Detection in Systems Software VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:29:35.045754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T16:59:07.712294Z digest=sha256:664745de782f4f0d70f431a090c1497f3e8cec96ace1a24ba8b14fa7380babdf

Observation 1e111069-6050-4813-adcb-4620dcaf2905 · inbound

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection cites this paper.

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-06-30T04:54:16.785146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T04:52:35.384665Z digest=sha256:a09ead4047d3253afc5b80deac0a7cc216c8c94f51b13dabe596fc0007f28249

Observation 058933ad-70f3-41cf-a329-554b9b14ad00 · inbound

DREA: Decoupled Reasoning and Exploration Agents for Repository-Level Vulnerability Detection cites this paper.

DREA: Decoupled Reasoning and Exploration Agents for Repository-Level Vulnerability Detection VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T05:13:52.726660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:13:52.726660Z digest=sha256:86e78c18870e6b082d63dff9c12e93752806c0579fe50ad4b8f04b5805652a8c

Observation 0049addd-b049-47cf-9033-f47aa6ea3e62 · inbound

VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection cites this paper.

VulnGym: Benchmarking Coding Agents for Repository-Level Vulnerability Detection VulDetectBench: Evaluating the Deep Capability of Vulnerability Detection with Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T16:59:23.765035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:59:23.765035Z digest=sha256:3979dbd72f66bed10dc95c65b9ce700c4deeb52846f42708e7c99a7dea55a73b