Pith. sign in

Paper Citation Record · LEDGER

CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2406.07599.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.07599 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:55:16.115892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:24.102448Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6d295178-0456-4dd9-83e5-f2fb80e74d11 · inbound

SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis cites this paper.

SV-TrustEval-C: Evaluating Structure and Semantic Reasoning in Large Language Models for Source Code Vulnerability Analysis CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T13:55:16.115892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:55:16.115892Z digest=sha256:03942a736bc7ec65c4bf10d4644486649eaa510c829e5e39c6d3656336a59da8

Observation 9160eeab-babb-4fc3-99d2-a6de7021e998 · inbound

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure cites this paper.

Military AI Cyber Agents (MAICAs) Constitute a Global Threat to Critical Infrastructure CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:40.069631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:27:40.069631Z digest=sha256:e9f89bf16d614ad9a4d8558621118e1138344082781c44e2300891af4323b076

Observation f6ef7597-a1e8-4265-aae1-a122bea82018 · inbound

A Practical Guide for Evaluating LLMs and LLM-Reliant Systems cites this paper.

A Practical Guide for Evaluating LLMs and LLM-Reliant Systems CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:55.150430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:41:55.150430Z digest=sha256:2ff4bb012087c8c2bfb304b8a213517cc787f92d437db2695a14deb1e7a485aa

Observation f012fec6-dbe4-4299-8217-d0d6fef50081 · inbound

False Alarms, Real Damage: Adversarial Attacks Using LLM-based Models on Text-based Cyber Threat Intelligence Systems cites this paper.

False Alarms, Real Damage: Adversarial Attacks Using LLM-based Models on Text-based Cyber Threat Intelligence Systems CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:55:32.994446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T07:54:51.925530Z digest=sha256:1d2f2f5b6498f58a457e0ee584bab0a5fdb98f1aacbafbad3e903c5592d39b75

Observation 3ca6a682-1b67-478e-a1cf-776abf18eac4 · inbound

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation cites this paper.

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.763400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T04:37:33.942379Z digest=sha256:7ca6cf3c11c5a27dff560f7d22fc77d42168dd561cf1492de0199c48b980cbe8

Observation 084bd0b0-62d6-400c-bd10-d6b3261f90af · inbound

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting cites this paper.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.064442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.064442Z digest=sha256:3f05e31f82662dab46aab6c765f7db55d7bffe6d84c45cf95ecab0d242392295

Observation f36749e2-0a77-47fa-b9e9-3261b9bb0c94 · inbound

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence cites this paper.

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:47.986398Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:47.986398Z digest=sha256:edb816d41d7a025cfb954af0aa950a6ee1706f308fedc19045f8be463ca2a5f8

Observation a701c1a1-be47-4628-a3aa-b149cbabf5f4 · inbound

Dynamic Cyber Ranges cites this paper.

Dynamic Cyber Ranges CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:16:52.816852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T03:04:03.611481Z digest=sha256:27470acfc81cbc500d27577d113766894952c643086d6c6786fd318b5ffd0152

Observation 0028edf9-0d23-4139-b133-f812aff1a3ec · inbound

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents cites this paper.

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:05:02.502191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T05:00:26.307473Z digest=sha256:ded5cbfc7f2ab662e4850f1f5517a5c9a24f6673c1960cdd810e22b967c4a585

Observation 6a284f94-2916-4a7f-9a6f-b2d187e1342e · inbound

Cybersecurity AI (CAI) Dataset cites this paper.

Cybersecurity AI (CAI) Dataset CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:13:26.887469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T12:07:49.656453Z digest=sha256:7c768b2432d8e8fcc7d59fc08979804b7b88e384dc07dab8e31698414d564e92

Observation b523346f-d1aa-492a-99e5-00b393894364 · inbound

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning cites this paper.

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.105488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T14:03:41.903170Z digest=sha256:f15c576b3c3ae14f810c5ad9f54a5ac9cf0af623fbdab320f20595141c1c6008

Observation 3073133f-b7f5-4a69-a307-ee116e01a7b7 · inbound

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense cites this paper.

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T10:23:20.740447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:23:20.740447Z digest=sha256:35f0a1c46bc2415671a2d8e1be2da30ab02ade9edba823cc0f020f8f0c3fa468

Observation 8848ad00-0978-448f-8d2f-90257fdb5331 · inbound

Antares: Foundation Models for Agentic Vulnerability Localization cites this paper.

Antares: Foundation Models for Agentic Vulnerability Localization CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T07:50:50.560574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:50:50.560574Z digest=sha256:07176abd164f201ad2e685769a7f6892aeeb96faaa774707bbc11c0a0dc3a060