Pith. sign in

Paper Citation Record · LEDGER

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning

As of 7 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2607.09328.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.09328 v2

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T07:40:58.826285Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4165941-492b-4e1c-b4a2-6a4d33f0e1ac · outbound

This paper cites Raw-source evaluation.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Raw-source evaluation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.826285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.826285Z digest=sha256:2cef0898da57a77eaf4831c55b94c0fa4635f3828e10640736366ea7ab98c842

Observation 3a4b5c20-6e72-4afc-8839-fd879b986223 · outbound

This paper cites URL https://arxiv.org/ abs/2606.19348.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning URL https://arxiv.org/ abs/2606.19348

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.860969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.860969Z digest=sha256:b60f2808d04050fc7a2688e5d7d5e392dd68570ef1381595abbfa101dfda560c

Observation 4622702c-c9ff-45af-95ab-d7f71a5845ed · outbound

This paper cites DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning DeepResearch Bench: A Comprehensive Benchmark for Deep Research Agents

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.937791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.937791Z digest=sha256:adeddf841ebff51ba937b3c5fce7e97a40dd39dd94a0175706e95955c6c3daf8

Observation 197b182d-0c74-49dd-adfd-2f2f0e5128d3 · outbound

This paper cites DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning DocScope: Benchmarking Verifiable Reasoning for Trustworthy Long-Document Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.029050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.029050Z digest=sha256:52b9c9da9cd174fa5e1b2294de570b23c2dcb8ab7438017b96efb7bc00995c8f

Observation 7ca26883-e3ce-42d8-9d41-76d8e40b4db9 · outbound

This paper cites Gemma 3 Technical Report.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Gemma 3 Technical Report

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.168353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.168353Z digest=sha256:11c9d5ab4f6474618135ed92e8eb5b97d3c6a8376819d80998f2dd17d62dca78

Observation 8af18281-6573-4540-b62e-5ed686f74739 · outbound

This paper cites Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.445058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.445058Z digest=sha256:7203a89e2d203e2d8f011e395750bc89c8f5f4763ad7a8e98883714dffb51f91

Observation 4fa34d82-c498-4dd2-b4a5-5e6b5208bdea · outbound

This paper cites an unresolved cited work.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.611667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.611667Z digest=sha256:279c5bde4f5ec8c9a96a12c7de94c3a28dcd4afa6c4c36cbf78293d968adb449

Observation 2f2da1a5-1f07-41fa-8233-ec0c4fa9635c · outbound

This paper cites Nelson F.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Nelson F

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.670841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.670841Z digest=sha256:33bd7dac314a6d5a70973862b1a896957a366a153a7d954b9039d9e094fa6f53

Observation 7260b785-016b-44bb-a2ff-a3db3dd3e8e7 · outbound

This paper cites Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.900586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.900586Z digest=sha256:a7290fb39c13085dad1bc55e5228d58dd9993689e8afbedff83732f3dd70f6b9

Observation 0d42286d-b6b7-4820-98a6-6a6c136b73a7 · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.024753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.024753Z digest=sha256:e1ec0e073b9e95eeca18d5f7aef52697a458c92d38f650059278947df20b4649

Observation d22003ee-8a96-45b7-926b-793b11998dfc · outbound

This paper cites Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.166701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.166701Z digest=sha256:62d6025a8d021304084d1f1b68f2b294fcc92fff1a93a260a2982019952684cc

Observation 3296dad3-f356-467d-b8b3-19ec7a5bf16f · outbound

This paper cites Humanity's Last Exam.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Humanity's Last Exam

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.554756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.554756Z digest=sha256:90339b5db4d2c0dad053993985da5ab4a352678ad4ef8093322406d04f0ced58

Observation 120ccf54-e084-4efd-8057-16fdb228a777 · outbound

This paper cites NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning NovelQA: Benchmarking Question Answering on Documents Exceeding 200K Tokens

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.650142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.650142Z digest=sha256:b9121ff0777bf6ec6a1c378003a4af3e386aebd33835ccdcc8e9d763739b3e30

Observation 9a7f0693-f639-49ef-8d61-87bfdc6691d9 · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:57.984751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:57.984751Z digest=sha256:053a06c93e9b3156fc1a2e9a967986bd78bee4e0f835ca50f5b1b451796855f5

Observation fe763d1b-0d53-4554-bd91-4ce47b9a6495 · outbound

This paper cites Kai Yan, Zhan Ling, Kang Liu, Yifan Yang, Ting-Han Fan, Lingfeng Shen, Zhengyin Du, and Jiecao Chen.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Kai Yan, Zhan Ling, Kang Liu, Yifan Yang, Ting-Han Fan, Lingfeng Shen, Zhengyin Du, and Jiecao Chen

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.114779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.114779Z digest=sha256:c24c05eb390b95656b43e2747944c2ebf5437de336cc05f09fb2dc4e2dd3b3bf

Observation b118c4f5-32ae-4d1a-aa9d-fbb97ea9b043 · outbound

This paper cites 100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning 100-LongBench: Are de facto Long-Context Benchmarks Literally Evaluating Long-Context Ability?

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.250915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.250915Z digest=sha256:a6debd2b9a5ef78cbc94d0760282ba7fa046a13b5544c4976daffb8f2dc129dd

Observation 7b9d11d5-bfea-4a30-a41a-fdc19305862f · outbound

This paper cites HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning HELMET: How to Evaluate Long-Context Language Models Effectively and Thoroughly

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.349130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.349130Z digest=sha256:521f172f3ba89d7e5693d8591589c08d8f3c53786726709a277ca4e45335d917

Observation ef65d078-8b46-4f19-b04d-d11b6c463472 · outbound

This paper cites Academiceval: Live long-context llm benchmark.arXiv preprint arXiv:2510.17725,.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Academiceval: Live long-context llm benchmark.arXiv preprint arXiv:2510.17725,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.484748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.484748Z digest=sha256:0c4641879ea6380e26ee0dd82d294deefe42fc93bd131237fae36941639dc0f8

Observation b8133e90-1526-4ae2-bf84-6b0bf42b83dd · outbound

This paper cites $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning $\infty$Bench: Extending Long Context Evaluation Beyond 100K Tokens

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:58.784842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:58.784842Z digest=sha256:4bb61d631246217c960dd084cab197181a334ea6cbfa3e7c1edf0a56d7d48a4b

Observation 5b0fcf00-ff01-4511-b418-633c4506c10f · outbound

This paper cites BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning BABILong: Testing the Limits of LLMs with Long Context Reasoning-in-a-Haystack

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.540931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.540931Z digest=sha256:b8e7e09c940b3b8b0f28f03416684cc2d50bd4e00c406c91bb375febfb77b90f

Observation f9655137-add8-40a1-8f72-c849a75f5c15 · outbound

This paper cites RULER: What's the Real Context Size of Your Long-Context Language Models?.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning RULER: What's the Real Context Size of Your Long-Context Language Models?

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.244749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.244749Z digest=sha256:608efb88948b172a3f78065a1a44ac26887c4878a580956fda06d791afe192af

Observation b1c97332-b68b-4b22-9915-9ee9be2efefc · outbound

This paper cites doi: 10.18653/v1/2021.naacl-main.365.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning doi: 10.18653/v1/2021.naacl-main.365

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.772989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.772989Z digest=sha256:21d08dae85b02a367db9462b3a9506505dd9e45718f4e0bd73c2f52a37481d40

Observation a7c6c680-f4cd-4972-a310-bef86399acc0 · outbound

This paper cites Gemma 3 Technical Report.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Gemma 3 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:56.105396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:56.105396Z digest=sha256:9c2eebf7ae1f57d4b6981b189400978c794be3285bb0fc0d4146c047c566139d

Observation ceac9ee5-faf5-4ba3-9d24-290cb87df2b8 · outbound

This paper cites LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.337896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.337896Z digest=sha256:9faa4bff7e5b728f2a440ab6dacf3d4b7a342b6bf5fe4b2bdebf0d20060e3035

Observation 61016f6d-48ac-4a0b-bd86-1562f88466a4 · outbound

This paper cites acl-long.183.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning acl-long.183

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.418414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.418414Z digest=sha256:4835b98f7162e7b75779372a5ea8d07819e158b641d7f91e9bc77e0ddd16416e

Observation f2e7e000-6061-4f37-8cfe-1dcaf8c8331e · outbound

This paper cites Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A.

WildTrace: Benchmarking Natural Evidence Trails in Long-Context Reasoning Pradeep Dasigi, Kyle Lo, Iz Beltagy, Arman Cohan, Noah A

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-02T07:40:55.620497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:40:55.620497Z digest=sha256:c7a67c25aa0b5f9d1b29641fefb0fb43272cac1f72b2b1b4905f53aa7d278ec2

Pith citing papers

No inbound Pith citation observations are available.