Pith. sign in

Paper Citation Record · LEDGER

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2509.23571.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.23571 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T14:43:00.540074Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-12T04:42:49.379223Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T06:01:22.935133Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed287650-8aae-4599-8832-05804fc4637c · outbound

This paper cites Exploring LLMs for Malware Detection: Review, Framework Design, and Countermeasure Approaches.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Exploring LLMs for Malware Detection: Review, Framework Design, and Countermeasure Approaches

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:55.997378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:55.997378Z digest=sha256:58039df255591100b02064510d14c20e15e2c3c05db8f5f5da551d7fe559399a

Observation 879a5909-2399-4b79-aa26-cb68a5cbd7b9 · outbound

This paper cites Recent activity by APT29 involving phishing attacks.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Recent activity by APT29 involving phishing attacks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:00.300356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:43:00.300356Z digest=sha256:5752e9ed144984956bbf42f3338680fb3039c8f654efb810acfcc8f6c3ef60a8

Observation 76eeb581-2705-45ad-9a38-86094e7e1a23 · outbound

This paper cites PentestGPT: An LLM-empowered Automatic Penetration Testing Tool.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting PentestGPT: An LLM-empowered Automatic Penetration Testing Tool

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.442650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.442650Z digest=sha256:8ed9cc8991b4df69f81b4d8f18b3a047bd9bcd551fb870808234e62b32e1988f

Observation c843c349-e698-40e1-83e5-a2563f7cc3cf · outbound

This paper cites Given a vulnerability description and metric values (Confidentiality, Integrity, Availability, Scope, Attack Vector, etc.), compute the CVSS v3.1 Base Score.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Given a vulnerability description and metric values (Confidentiality, Integrity, Availability, Scope, Attack Vector, etc.), compute the CVSS v3.1 Base Score

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:00.368756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:43:00.368756Z digest=sha256:705e846dd40bc79d12977d2fc80e04eddf818ecfe71985ef13cd9cf1e448deb3

Observation 73607cfe-56a0-4dad-bcf7-da49325fa311 · outbound

This paper cites LawBench: Benchmarking Legal Knowledge of Large Language Models.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting LawBench: Benchmarking Legal Knowledge of Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.728972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.728972Z digest=sha256:3cedb85a150d2ba16050f0c79d615ff7b469601e0986b4ff8989bcffd620ed5a

Observation ff3934ee-a634-4147-8c59-bf63a2da6dad · outbound

This paper cites Getting pwn’d by ai: Penetration testing with large language models.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Getting pwn’d by ai: Penetration testing with large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.871031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.871031Z digest=sha256:d8328076b3eda109ee9f2ff2d600a67a8accf90d7ca3555ae84ad8ee2de97f16

Observation 3479835f-b7dc-4ee1-915f-dd801fba241d · outbound

This paper cites Linking Threat Tactics, Techniques, and Patterns with Defensive Weaknesses, Vulnerabilities and Affected Platform Configurations for Cyber Hunting.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Linking Threat Tactics, Techniques, and Patterns with Defensive Weaknesses, Vulnerabilities and Affected Platform Configurations for Cyber Hunting

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.074742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.074742Z digest=sha256:a66f68bc34ef12927c3e90a4f50dc48ba6790821f5f372cdc0b2312802fea5af

Observation 86b985b7-1372-49a6-9559-b40879ee6b34 · outbound

This paper cites Double Backdoored: Converting Code Large Language Model Backdoors to Traditional Malware via Adversarial Instruction Tuning Attacks.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Double Backdoored: Converting Code Large Language Model Backdoors to Traditional Malware via Adversarial Instruction Tuning Attacks

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.277985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.277985Z digest=sha256:05ec967683927e4fddbc497f24e88c54d32aa261ab7b317901ad8f9130c34062

Observation a74e8a31-41a2-46ee-b059-42dd4d490b5e · outbound

This paper cites AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.447644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.447644Z digest=sha256:a789a867d842d6dddad5a1d570d2e9588a39f37e84d248ec9ba2e20fde7f1979

Observation f0a9743d-324b-46dd-9305-bad0f4975483 · outbound

This paper cites SEvenLLM: Benchmarking, Eliciting, and Enhancing Abilities of Large Language Models in Cyber Threat Intelligence.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting SEvenLLM: Benchmarking, Eliciting, and Enhancing Abilities of Large Language Models in Cyber Threat Intelligence

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.638474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.638474Z digest=sha256:f74468c038ed5f82c7a8af64de94a18291a94ada400d19232fa9f2f2c18c2fe7

Observation f0696039-ceff-457a-b349-8a9394cff6ca · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.860062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.860062Z digest=sha256:b3b486d80c305639dc7e818acb28eeb46c0a7448fe2370c61cafa8216e78e219

Observation 74321f04-6cec-4cf2-bdbb-52392101377d · outbound

This paper cites From ML to LLM: Evaluating the Robustness of Phishing Webpage Detection Models against Adversarial Attacks.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting From ML to LLM: Evaluating the Robustness of Phishing Webpage Detection Models against Adversarial Attacks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.029258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.029258Z digest=sha256:6802fb142f9e68a73b35cb297d3644285aa7fb4fe2d07b9c7b74eddcfcdef513

Observation afcbea36-528e-46c1-a069-673915217a58 · outbound

This paper cites Mary M Lucas, Justin Yang, Jon K Pomeroy, and Christopher C Yang.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Mary M Lucas, Justin Yang, Jon K Pomeroy, and Christopher C Yang

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.251543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.251543Z digest=sha256:dc52b2286fd38e896c0a84c2775b1736ea3df09fd9035bcc5ae58188cfc5dc89

Observation 1deb6f73-04d7-4ccd-8be9-d519bfa4a402 · outbound

This paper cites John Morris, Eli Lifland, Jin Yong Yoo, Jake Grigsby, Di Jin, and Yanjun Qi.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting John Morris, Eli Lifland, Jin Yong Yoo, Jake Grigsby, Di Jin, and Yanjun Qi

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.347252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.347252Z digest=sha256:92ebe8d0eef4ce4646bd3ae126a087234bc845a18bc224fdcbd08282c91c5ac8

Observation 84e9fd00-88f7-4d68-9944-69a8ba80a76d · outbound

This paper cites Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.510053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.510053Z digest=sha256:f40ce2085c95a2e1208d692960cd179f3fdf8823e13fef9087d3fd0401082530

Observation 56861ea8-a6ca-447d-b970-86941d6a8b75 · outbound

This paper cites HackSynth: LLM Agent and Evaluation Framework for Autonomous Penetration Testing.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting HackSynth: LLM Agent and Evaluation Framework for Autonomous Penetration Testing

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.660386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.660386Z digest=sha256:24c6fc4e30e3260e312591cab3eebacba95766a537abc8d634644d316496988e

Observation 3540c3ce-97a8-4aa0-b976-37c5f23aec20 · outbound

This paper cites LAMD: Context-driven Android Malware Detection and Classification with LLMs.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting LAMD: Context-driven Android Malware Detection and Classification with LLMs

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.798445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.798445Z digest=sha256:ae42ae2ca793c342d66109978506391f217b42ec1687d1d05de8c3785a9657d2

Observation 822c3d1c-f698-4902-be20-5fe59bf02018 · outbound

This paper cites PentestAgent: Incorporating LLM Agents to Automated Penetration Testing.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting PentestAgent: Incorporating LLM Agents to Automated Penetration Testing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.126030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.126030Z digest=sha256:ac8b9b641c55cd2b317749652a5def4b685030833cb829b56b5eb11a8cbcb166

Observation 181ddbef-3acf-404c-a71e-400fc87569da · outbound

This paper cites LProtector: An LLM-driven Vulnerability Detection System.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting LProtector: An LLM-driven Vulnerability Detection System

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.281707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.281707Z digest=sha256:77ebeeaca7b0f17e77ef9b42a56f047416a670a6c194519cb3f516bcab576e66

Observation fc089923-f062-4713-9b86-731138440fd7 · outbound

This paper cites Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.437743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.437743Z digest=sha256:64ece5bea2509ce9bd1c3150b97cd69be05cc8b2c25eb415c6c4898d089565e9

Observation 24a463b9-bcb9-479c-9c74-10d3354b88aa · outbound

This paper cites Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Towards Understanding Chain-of-Thought Prompting: An Empirical Study of What Matters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.598970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.598970Z digest=sha256:237aa7e38c4f04363cf4a1a3cc037749c8b7b8bb917075b9b42f471fa0208a8e

Observation 97b00db1-7dc6-49d9-8677-e69b0a1ddb9f · outbound

This paper cites Towards Better Chain-of-Thought Prompting Strategies: A Survey.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Towards Better Chain-of-Thought Prompting Strategies: A Survey

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.770384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.770384Z digest=sha256:2a14e6c751d212b47bcd8f81f02f07864555e0cc435481f5b974dade59ae94e6

Observation de27c9e2-631d-4ce6-92fc-0ad621c483f7 · outbound

This paper cites BERTScore: Evaluating Text Generation with BERT.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting BERTScore: Evaluating Text Generation with BERT

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:59.893756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:59.893756Z digest=sha256:03dae0bd23c5c151fd2ed4597ff593755aba0b1a993b89563af5a06caba26a8b

Observation cf5585ba-a656-413d-a23f-c5741f87fc15 · outbound

This paper cites Each CVE entry is enriched with exploitability status, malware connections, and actor attribution.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Each CVE entry is enriched with exploitability status, malware connections, and actor attribution

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:00.068603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:43:00.068603Z digest=sha256:4736d1da604d0527634544dbcb982952f911e1f9200ae49296ba122afff122dd

Observation d1cfb6ad-fc9c-41ea-93df-b837e26610b4 · outbound

This paper cites Entries often include exploitability scores, attack vectors, exploitation status, and tags related to malware or campaigns.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Entries often include exploitability scores, attack vectors, exploitation status, and tags related to malware or campaigns

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:00.190538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:43:00.190538Z digest=sha256:f38d1e759cee1d65773fe903623be480ec860763d17de6947029f7eff356aec4

Observation 713946d3-1953-4df6-bfb2-9bbf9ac0e35d · outbound

This paper cites \n", "Q:.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting \n", "Q:

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T14:43:00.540074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:43:00.540074Z digest=sha256:55bb652e003ccda3321fc9f036a4e0fb7656060c93595ba650a8f4d260bf76d1

Observation 8b7d6512-f21c-434b-a8e0-1e0f7890470d · outbound

This paper cites EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents

Reference 2013

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.148795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.148795Z digest=sha256:aad816ca614b56e89173203a7f8b66591842d1863c4c88971d5eda8e7154f8c9

Observation 25fa5b4f-dfe7-49d1-8644-beaec5f741f1 · outbound

This paper cites A Survey on In-context Learning.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting A Survey on In-context Learning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.514657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.514657Z digest=sha256:a9fc24b2468ead3cd5b2c3b56f91443929fa533e33d20754d974e3d51ea8c415

Observation 57876458-b78d-4858-a0d9-49f8537d52c8 · outbound

This paper cites Improving Discovery of Known Software Vulnerability For Enhanced Cybersecurity.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Improving Discovery of Known Software Vulnerability For Enhanced Cybersecurity

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:58.909486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:58.909486Z digest=sha256:31f810080b800d01ebd41443340fe449f128db48ab11019cad868d16e289ab21

Observation 2f3222c0-e045-4213-a748-ff069c3b7153 · outbound

This paper cites Turning the hunted into the hunter via threat hunting: Life cycle, ecosystem, challenges and the great promise of ai.arXiv preprint arXiv:2204.11076,.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Turning the hunted into the hunter via threat hunting: Life cycle, ecosystem, challenges and the great promise of ai.arXiv preprint arXiv:2204.11076,

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:57.173909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:57.173909Z digest=sha256:cb8b89e97eeae0f73762c9509427630f4fa6559d4ce5ae36096bef5b5975311c

Observation 9ccc2245-e5a5-4540-bb65-45706cfcee4b · outbound

This paper cites ThreatZoom: CVE2CWE using Hierarchical Neural Network.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting ThreatZoom: CVE2CWE using Hierarchical Neural Network

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:55.916346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:55.916346Z digest=sha256:524d1a8bf63cae4f09974ad96db0e0cde5f48e59a1709edbb56b8190cec91aa7

Observation c3d143e5-2835-4012-a3a8-43833b7994f0 · outbound

This paper cites ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting ReSpAct: Harmonizing Reasoning, Speaking, and Acting Towards Building Large Language Model-Based Conversational AI Agents

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.627111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.627111Z digest=sha256:3abee11555ae5020f001c23745dc5290d092808fbaca2924719a22a183f07971

Observation b3979bdf-9f6e-4b69-b65c-3e4e97cb0781 · outbound

This paper cites Frameworks for Querying Databases Using Natural Language: A Literature Review.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Frameworks for Querying Databases Using Natural Language: A Literature Review

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.347322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.347322Z digest=sha256:a74e94f0e4b5ba44bb9bfbeddff815c6aa8797d8b40f4e8a481e9b853785ff3b

Observation 084bd0b0-62d6-400c-bd10-d6b3261f90af · outbound

This paper cites CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting CTIBench: A Benchmark for Evaluating LLMs in Cyber Threat Intelligence

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.064442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.064442Z digest=sha256:3f05e31f82662dab46aab6c765f7db55d7bffe6d84c45cf95ecab0d242392295

Observation 0150a05d-0931-4a38-aa7d-b870dc710369 · outbound

This paper cites Leshem Choshen, Ariel Gera, Yotam Perlitz, Michal Shmueli-Scheuer, and Gabriel Stanovsky.

Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting Leshem Choshen, Ariel Gera, Yotam Perlitz, Michal Shmueli-Scheuer, and Gabriel Stanovsky

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T14:42:56.259736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T14:42:56.259736Z digest=sha256:af1e46b4df50360d3a362cc481f511a9a09b946124e7d96b391be40eeeb38108

Pith citing papers

Observation fadfd38b-f307-4367-a530-2dee41952e18 · inbound

The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory cites this paper.

The Trap of Trajectory: Towards Understanding and Mitigating Spurious Correlations in Agentic Memory Benchmarking LLM-Assisted Blue Teaming via Standardized Threat Hunting

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-29T02:04:56.816440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:42:49.379223Z digest=sha256:92a898c59de49e0bb1abc404d74c53962e484bc7808349dc2bf0bf4aa5589505