Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T13:10:41.870343Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 0 inbound Pith citation observations for arXiv:2605.26548.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T13:10:41.870343Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
24 of 24 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6cd2488a-f694-4ced-898c-f6f4d4c8cbf8 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57c6980f-9f27-46d7-b3e5-6fb1eb5cb49b · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? not on target page
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bedb53be-0590-4fac-ab8d-c6e1fd5599ad · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b050198-709c-47a4-aef4-f42089960311 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Haoyu Li, Xijia Che, Yanhao Wang, Xiaojing Liao, and Luyi Xing
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ecd2d0d-a224-486d-a18a-db12c78eb25f · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? A Dual-Loop Agent Framework for Automated Vulnerability Reproduction.arXiv preprint arXiv:2602.05721,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5277e74-807e-4536-a468-587eb7e192b3 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Automated Vulnerability Validation and Verification: A Large Language Model Approach.arXiv preprint arXiv:2509.24037,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a148fe4-ec80-4b6a-9cba-682de91936f7 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? ARVO: Atlas of Reproducible Vulnerabilities for Open-Source Software
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b97b4a4f-fc81-439c-b441-46569dc84bd0 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Codex Security.https://help.openai.com/articles/20001107, 2026a
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8314cb41-31ea-4907-946f-eb1ec8b816b6 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Patch-to-PoC: A Systematic Study of Agentic LLM Systems for Linux Kernel N-Day Reproduction.arXiv preprint arXiv:2602.07287,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca79777c-087b-4274-9059-94f8acdc096b · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? To Err is Machine: Vulnerability Detection Challenges LLM Reasoning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dbca145-14d2-462f-976e-139eadd99315 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? From Naptime to Big Sleep: Using Large Language Models To Catch Vulnerabilities In Real-World Code.https://googleprojectzero.blogspot.com/2024/ 10/from-naptime-to-big-sleep.html,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d49d6713-073a-4bee-8f1a-bf369c216c04 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Coskun, and Gianluca Stringhini
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11ee100e-dcd4-4f1c-8d52-2f76fa6c5f7f · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? From CVE Entries to Verifiable Exploits: An Automated Multi-Agent Framework for Reproducing CVEs.arXiv preprint arXiv:2509.01835,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6040990-466e-48c2-9882-f4293c255258 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Zichao Wei, Jun Zeng, Ming Wen, Zeliang Yu, Kai Cheng, Yiding Zhu, Jingyi Guo, Shiqi Zhou, Le Yin, Xiaodong Su, and Zhechao Ma
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4101701b-1883-46c1-ab19-33749dc1f3b6 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Why Is CSP Failing? Trends and Challenges in CSP Adoption
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8877253a-ba49-450b-be39-ee7688061346 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? ProgramBench: Can Language Models Rebuild Programs From Scratch?
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5df73039-f2a9-4449-9229-9451da11a600 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Bhatia, Vikram Sivashankar, Yuxuan Bao, Dawn Song, Dan Boneh, Daniel Ho, and Percy Liang
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11cfe1db-ee4d-4bb6-ac90-edee42727ef3 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? A Systematic Study on Generating Web Vulnerability Proof-of-Concepts Using Large Language Models.arXiv preprint arXiv:2510.10148,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a058625f-3062-4926-af18-8be7bda83573 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? AnyPoC: Universal Proof-of-Concept Test Generation for Scalable LLM-Based Bug Detection
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f986cd1-df99-4d0d-b9d3-14d0fac539bd · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? FaultLine: Automated Proof-of-Vulnerability Generation Using LLM Agents
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94eda813-e819-4078-9c09-76ab404abcdb · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? ZeroDayBench: Evalu- ating LLM Agents on Unseen Zero-Day Vulnerabilities for Cyberdefense.arXiv preprint arXiv:2603.02297,
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dfba7bd-f1fb-4ba8-ba56-651694406fa0 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? PoCGen: Generating Proof-of-Concept Exploits for Vulnerabilities in Npm Packages
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41e5fc11-617f-4f8b-b02b-316fb5fafc8a · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Context Length Alone Hurts LLM Performance Despite Perfect Retrieval
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7d3e999-f2db-44ae-b16f-49e938e61a34 · outbound
SEC-bench Pro: Can Language Models Solve Long-Horizon Software Security Tasks? Introducing Claude Opus 4.6
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.