Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:01.393974Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 1 inbound Pith citation observation for arXiv:2505.19973.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:01.393974Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-10T15:21:11.287455Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-11T10:41:07.580051Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 8e7400f4-1739-431f-b46a-09b65903ab69 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Advances in Neural Information Processing Systems 37 (NeurIPS 2024), Datasets and Benchmarks Track (2024)
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7c401da4-505c-4185-88f3-09e4dca02f06 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response The DeepSpeak Dataset
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce3740db-48bf-4642-80df-11a6ad02864e · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Proceedings of the 8th International Con- ference on Information Systems Security and Privacy
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 42bfa150-170b-405d-be2d-e18e8abd04f9 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 56870997-740c-4c89-8fb4-37d16de7c9d1 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c8fbc282-17a1-4067-81e6-1c9bb0a5c761 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Al-Onaizan, Y., Bansal, M., Chen, Y.N
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 686507d1-4c77-4c22-a672-4981dd42c0e7 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Internet of Things and Cyber-Physical Systems5, 1–46 (2025).https://doi.org/10.1016/j.iotcps.2025.01.001
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 522d94ff-2cf8-4a0a-b126-814867508d6f · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response IEEE Access12, 23733–23750 (2024).https://doi.org/10.1109/ACCESS
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4857d67b-0746-4bc9-ad44-c737b17a1ac4 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response FrontierMath: A Benchmark for Evaluating Advanced Mathematical Reasoning in AI
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5079905-c200-4c49-bdcf-a3a00b93d006 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Forensic Science International: Digital Investigation38, 301264 (Sep 2021).https: //doi.org/10.1016/j.fsidi.2021.301264
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ac2167b3-5f91-457d-99b2-919a03ed2e68 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Packt Publishing, Birm- ingham, England, 2 edn
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e35c8751-ab69-4505-8fde-fc46e66d96df · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c4b10b2-35ba-456a-a8fa-bc92a2cab5d7 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9e7c063d-4d7d-4325-bfa6-57f05b040eb7 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response IEEE Networking Letters4(3), 162–166 (Sep 2022).https://doi.org/10.1109/ LNET.2022.3185553 14 B
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 72a33af0-4852-4e67-99d9-f8b2a6b444e6 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: GLOBECOM 2022 - 2022 IEEE Global Communications Conference
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a0a9d54-458c-4caf-b6ba-91cb97dc44b2 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Computers14(2), 67(Feb2025).https://doi.org/10.3390/computers14020067, number: 2 Publisher: Multidisciplinary Digital Publishing Institute
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 22996b99-e82f-405c-810c-79745a1871aa · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Forensic Science International: Digital Investigation48, 301683 (Mar 2024)
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e6d8b77f-7fc2-4c84-bca1-87ff85087c27 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Ad Hoc Networks174, 103840 (Jul 2025)
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 044247eb-f8c9-4e0d-8efa-6cdf17a71cdd · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Computer Networks227, 109688 (May 2023).https://doi.org/10.1016/j.comnet.2023
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e2e559d2-a13e-442e-b6e0-aaa4a46e923e · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: 2024 5th International Conference in Electronic Engineering, Information Technology & Education (EEITE)
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ac27fb3-1426-47d9-9540-d256df0ae217 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response IEEE Software40(3), 4–8 (2023).https: //doi.org/10.1109/MS.2023.3248401
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3630059-5892-48b5-a009-7b699e10ae81 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Proceedings of the 2016 Conference on Em- pirical Methods in Natural Language Processing
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a1959e6e-bbdc-4ade-a614-fe6958d5d6b9 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Forensic Science International: Digital Investigation46, 301609 (Oct 2023).https://doi
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c74d775-9016-4dce-a428-bfd8ed546358 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Forensic Science International: DFIR-Metric: A Benchmark Dataset for Evaluating LLMs in DFIR 15 Digital Investigation52, 301872 (Mar 2025).https://doi.org/10.1016/j.fsidi
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85a067d4-6cd1-4b97-82c1-89b6252237ea · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response EURASIP Journal on Information Security 2017(1), 15 (Oct 2017).https://doi.org/10.1186/s13635-017-0067-2
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation f040064d-e91f-473a-a599-44e0a74bf403 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Computers and Electrical Engineering124, 110307 (2025).https://doi
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0247944f-51a0-41c1-bd3e-770c7270a65d · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response https://doi.org/10.48550/arXiv.2505.03100
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7d854e0d-28d3-41c4-81fc-89bda38cd9f1 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: 2024 IEEE Interna- tional Conference on Big Data (BigData)
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ec31bb7-ac50-4844-8102-82982d5a4821 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: 2024 IEEE International Conference on Cyber Security and Resilience (CSR)
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b4fafc2-d381-4dc3-a0d5-33a84d82a272 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Ideas That Cre- ated the Future, pp
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 35d555da-003d-4796-af14-32b1c69a8536 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Proceedings of the 31st International Conference on Neural Information Processing Systems
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d499b04-88be-4839-bfbb-c7e9997de976 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Linzen, T., Chrupała, G., Alishahi, A
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19180516-dd0c-4787-8f96-833fe3e02d5a · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Proceedings of the 31st Inter- national Conference on Computational Linguistics
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4c631ead-22b0-454f-bfe9-5ca53e402b26 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: Proceedings of the Digital Forensics Doctoral Sym- posium
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1497f6d1-788d-4d2c-90cb-1fd6a276aad2 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response In: 2024 12th International Sympo- sium on Digital Forensics and Security (ISDFS)
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eaaa9b9-46bd-4c3d-98e5-f2b8a6a11824 · outbound
DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response Digital Forensics in the Age of Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ee7520f-eeab-46dd-a785-7edb08ccdea1 · inbound
SIR-Bench: Evaluating Investigation Depth in Security Incident Response Agents DFIR-Metric: A Benchmark Dataset for Evaluating Large Language Models in Digital Forensics and Incident Response
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.