Pith. sign in

Paper Citation Record · LEDGER

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks

As of 6 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2608.01684.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01684 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T22:55:34.062935Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact4
  • verified fuzzy3
  • unresolved55
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d836d9d7-2953-4e83-821b-13dccc9e5d08 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:35.060611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.801990Z digest=sha256:f826b8cea40ba8ce8174d06e6507491c7ab60281ccc13e2d15cacb590e78e014

Observation f6df4890-713d-4185-af91-6d43e8e75450 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:35.046025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.807382Z digest=sha256:cc87a483f9fb8f1d2d28585a172f0a39760cd22deb447afdb8a17839a93c98ca

Observation 99b16aa8-1a28-49fd-8fd0-b658780b368e · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:35.032621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.812221Z digest=sha256:f261d0f64f42dfa54f600631161fa70de33baa1c232a36ebee34234a632e91e4

Observation e6390801-0f0a-4f3f-bb2e-cf9fe507a75d · outbound

This paper cites LLaGA: Large Language and Graph Assistant.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks LLaGA: Large Language and Graph Assistant

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.817111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.817111Z digest=sha256:f40f4420e38abe42a940f5556949d8e502efb29252adfe786bbc534d1d40c29c

Observation f66ddb8a-ec3c-4cd3-8525-876758250860 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:35.019312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.821860Z digest=sha256:ee3ae79f786dfae4083e5c82a6bd3518463e73afc07a171440e7968d07416bd1

Observation b3bb4989-ac69-4ff3-b76e-bdbd23e661e8 · outbound

This paper cites Text-space Graph Foundation Models: Comprehensive Benchmarks and New Insights.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Text-space Graph Foundation Models: Comprehensive Benchmarks and New Insights

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.826338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.826338Z digest=sha256:28e494406604fd8ce2e54cfa3f5ab8b7e8eaf50551413c3a9f339c6cac3c28b6

Observation 8dcb1651-f874-4bf9-a855-486533303958 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:35.005597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.831378Z digest=sha256:f31743ea88dd6ec380f884a8776d584f5e60c13e26b795513120a8dec017067a

Observation f0a4c428-9af8-4302-b6be-dffb5c124c65 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.835471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.835471Z digest=sha256:359813869cb4646fbce02671f14cbdd5b8110f30bbe70f8f1738bfc2d6b61148

Observation 8e62699a-8dd4-431d-acdc-c69ffd5d2983 · outbound

This paper cites Say No to the Discrimination: Learning Fair Graph Neural Networks with Limited Sensitive Attribute Information.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Say No to the Discrimination: Learning Fair Graph Neural Networks with Limited Sensitive Attribute Information

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-04T22:55:34.549392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.839694Z digest=sha256:d49564792e0350979c38ee231f1a93fec483c838776fe871107b923a8d5f6e06

Observation 7bfd74b2-a10e-4880-a8f8-b1c89c88b758 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.843988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.843988Z digest=sha256:35eb7e1c7fb1702c0d329facba37db77931116db1b5742565d036817f04bed01

Observation c66fdabe-b608-47dd-b862-000209c685fb · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.991876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.848151Z digest=sha256:c6a8062f3ccc6816c2314e44939ab7bd2b81ed924314ee9895532c052dfcd1f5

Observation 276babb0-3bc4-4f33-824b-142cf4c3df7d · outbound

This paper cites WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.852209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.852209Z digest=sha256:b595f18006d43554b4bd1608cac2a2156d241047eec6319f9e6ec58bbc0b3ff9

Observation c49d4c05-295d-4109-a896-0d666d2a8d4a · outbound

This paper cites 2010.Networks, crowds, and markets: Reasoning about a highly connected world.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks 2010.Networks, crowds, and markets: Reasoning about a highly connected world

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T22:55:34.978345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.856966Z digest=sha256:c59ce670f69aaa628f8b99979a8ae782c6147083402a2d67b638dbd7c73a2560

Observation f9a9cd88-582f-4050-b498-a0b62898cda1 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.861288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.861288Z digest=sha256:26358dffcd38caade1bde49d24f7c635623d636b6a75598416070540383fd5a2

Observation cf2ba001-ef57-49f4-a488-8c0b1e32e85e · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Gemini: A Family of Highly Capable Multimodal Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.865363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.865363Z digest=sha256:43b14b40bf81dbbb211018315fd8a489f628b9b1c4084921b99510d3d6dc89e5

Observation 36f91382-984a-4040-a8a4-da19308afa37 · outbound

This paper cites GLM-5: from Vibe Coding to Agentic Engineering.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks GLM-5: from Vibe Coding to Agentic Engineering

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.869658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.869658Z digest=sha256:102dd622e5c67d4bd968ea286ac5e37dfbdb32d18c07c1feaea59bd75449ef46

Observation dbaa1d4c-6d55-41d7-8aa0-facdc66c79f0 · outbound

This paper cites 2018.Graph theory and its applications.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks 2018.Graph theory and its applications

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T22:55:34.963017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.873902Z digest=sha256:13ffbfdd26f7a072bc64972f3fc126c421ef36f9ad3283f76fd926f42dc9651c

Observation 40a70425-30de-44c8-830f-ca6e46bfc2b9 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.947781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.877854Z digest=sha256:6f1ca1e568d4e1836fdf4705ba89a56317d5c8c95031d20a60022551b332d0b1

Observation f054a222-f859-41be-a318-2cde90f44b7e · outbound

This paper cites GPT4Graph: Can Large Language Models Understand Graph Structured Data ? An Empirical Evaluation and Benchmarking.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks GPT4Graph: Can Large Language Models Understand Graph Structured Data ? An Empirical Evaluation and Benchmarking

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.882180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.882180Z digest=sha256:6797696862ab1e1ff647ddcbdfcac68e7479d190c83f5655e4c30f14ceb147eb

Observation 458ab85c-aefa-49a2-90be-f344b8374335 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.933099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.886421Z digest=sha256:f0c4389ea26702b3cadf6556d184a3f7c61b11221d01f9b33dcc32b22f32313a

Observation 60e18aef-62ab-4cb9-ad65-0d5d5b3fe03b · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.919593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.890239Z digest=sha256:3cac7946e6c5f347cba77f4c38981505af4ba5c9419d76ea24e66e1aed0d1e1c

Observation ba18d34b-cd20-4972-9e1e-2d36bc2c709c · outbound

This paper cites InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks InfiAgent-DABench: Evaluating Agents on Data Analysis Tasks

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.894195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.894195Z digest=sha256:c672fb066a4d9ebf556f8bc6cb8c11dee15c76e5028726891c1e081629c3dae5

Observation 0903fece-2407-4191-bb4a-65778fc31ff4 · outbound

This paper cites DGraph: A Large-Scale Financial Dataset for Graph Anomaly Detection.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks DGraph: A Large-Scale Financial Dataset for Graph Anomaly Detection

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-04T22:55:34.416687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.898335Z digest=sha256:136a8e2ca11b25cdbfb5f11ce138ff35a68e5c8fcaf5f545a9c5f9f16ce29a87

Observation 1cd9c0de-505c-4048-bd78-34d9da68b43c · outbound

This paper cites SWE-bench: Can Language Models Resolve Real-World GitHub Issues?.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks SWE-bench: Can Language Models Resolve Real-World GitHub Issues?

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.902405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.902405Z digest=sha256:bc1813efe6e6da5fcf591f19c9b95dbd056d14b14271b8231959687c57f63e2d

Observation c098a2c3-485e-4bc0-bb43-17d5a2cc4d6a · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.905160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.906434Z digest=sha256:e021bf3f545e9d65cacf534b6b8eae21c4992e432361e7285c352f4c28d5c68b

Observation 5f2567c9-1aa7-4e31-9b75-1477ae6e9972 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.891358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.910525Z digest=sha256:5bec1aba61480616131f259d6564ea14217ba8bf04b2f06deeca32e561f46897

Observation 820fa263-c23f-4c2c-9bae-e967eb081fd3 · outbound

This paper cites GLBench: A Comprehensive Benchmark for Graph with Large Language Models.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks GLBench: A Comprehensive Benchmark for Graph with Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.914644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.914644Z digest=sha256:e6bef6b16a6f7b7ed9dfa4adfb84035edd9728b0be813f022c88333e900838ad

Observation b3b137bb-c544-4b3a-9e23-b1b40c52da70 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.918613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.918613Z digest=sha256:2bcabc68ff5485d405446afd6d53994bfc099977625778e90c9d2912c36d2b8f

Observation c999678e-a75f-498f-827a-374e871093bd · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.868589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.922552Z digest=sha256:80fec156e1c15939ef91ebe506f244e599a0034c794b2c38ace8e2d293d7f269

Observation 4d6b0241-2117-400e-a236-cdc8433cd062 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.855511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.926230Z digest=sha256:848b3bc8b4d9fbe69ad5860f620315c6ff4861963a2f2987c86b757ee0944cea

Observation c4fa36b6-13b8-4280-a455-2d9301b708bf · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.929961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.929961Z digest=sha256:a1e188ccdb9b662358177ec893a5b3e6f210edeb288f4de9b880b6ae5066ec32

Observation 19e07122-acca-43cd-b924-d091cfe0469e · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.842161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.934025Z digest=sha256:3bb9374784ffeb49afac73e46f9473a6d317d2e0ca451750d88697fa11e1d59a

Observation bcec49f7-ac1a-435c-a4a4-ec8572b51161 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-08-04T22:55:34.346610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.937763Z digest=sha256:f8759de45531c89461054b30c5a825a98af5d85a777bf238c87b607b2c0b54e6

Observation fc463f84-d24e-430b-9e6b-2437e1da6901 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.941488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.941488Z digest=sha256:0376b0599e810ebce56193793027d4755d6701557c1c03c79f459e02878d625e

Observation 5d3dcf1d-ef6a-4d14-bf5e-347556ce2f83 · outbound

This paper cites Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Terminal-Bench: Benchmarking Agents on Hard, Realistic Tasks in Command Line Interfaces

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.945747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.945747Z digest=sha256:40068bbb9e7cab804240d6da4600700d4970261f74cf6ca02d72b2f6b8d7b5dc

Observation b2995643-22d2-4ef3-8306-5465d78ff84f · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.828812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.949842Z digest=sha256:c0ed8b30e327715a2cf509c87b55ae0f7aed9f86f8bf9e907f8ea02220ddbfe8

Observation 6d008dc1-2646-4744-b814-24477b429e14 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.815493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.953606Z digest=sha256:567f252f511068799cd0d40ca253e080f015d3da68ec28769f8b3d342688274a

Observation 04a18032-aa63-4bac-b9e4-297559f467de · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.801491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.957401Z digest=sha256:755e39883cc6d75af2e6d3f2c863b437655ba99a06474d46a51c64c7a842712b

Observation bf76f959-b0c6-489e-bafb-c9a64aae59b2 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.787917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.961419Z digest=sha256:57ff921e90875501a92c346b37e1a25a18f10759b2a02e1c35539263fab07a6c

Observation 8c8fb3fb-487b-43b1-bf57-134318c0a0f3 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.773791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.965081Z digest=sha256:44ba09fd385592b902454b90ca7498dcd418636fd9c3810ab1fe19260e232cc2

Observation 4392ff23-6816-4fd7-8f81-e4d3920b265a · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.760691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.968884Z digest=sha256:6aa129ddf83bf78e34c1ce005f5a4ec6c6cc726629d5447143100803ccf39b51

Observation 8b7636d1-0778-4898-bcce-517c4d2aa8c0 · outbound

This paper cites NAS-Bench-Graph: Benchmarking Graph Neural Architecture Search.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks NAS-Bench-Graph: Benchmarking Graph Neural Architecture Search

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-04T22:55:34.291947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.972852Z digest=sha256:8fd570752d126f51acbfb1846d001d4e71f1516f9b79ce66e21b8cc4aea3382a

Observation 3e87baf1-0d51-472a-acda-34ac1fef7de6 · outbound

This paper cites Identifying the Risks of LM Agents with an LM-Emulated Sandbox.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Identifying the Risks of LM Agents with an LM-Emulated Sandbox

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.977101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.977101Z digest=sha256:1f5c166e15f43f62b790c55ea2a8974e0568b220207dc529edcb023ec5881b05

Observation 0c0dc3b5-5472-47fb-9044-e00b21b6dce9 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.747027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.981227Z digest=sha256:91874e75fbe60b758ae9a3a77de2bfcc5709c0539cb2ebe43f6caa4a1233677a

Observation 14526c09-8cfe-495b-a909-2e1955fdd076 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.732674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.985197Z digest=sha256:8083e934169bac3e7ccf180d8feba309c0bd35645d06443716fd29e739d7a700

Observation 7fe10f46-253f-4570-bf96-30e38ab9b671 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.718685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.989084Z digest=sha256:ce755c768743ed9ba85f76bcccb1d97b1edb7330ac2f25b019b3cff89bd4c095

Observation 98547dcf-eb3a-4191-aa38-35cf4adbfef4 · outbound

This paper cites MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:33.992939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:33.992939Z digest=sha256:0499b41f29b188b9c4bd26cbf337f414fb4be20ceb8055b29b8254294f849dd3

Observation 5339d483-f311-46f1-9326-3211bfed6ec1 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.705158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:33.996975Z digest=sha256:8115b1443f68c7fe921bb00dadd1c0bdc163d8941a5b9c745a63c51cbf385f66

Observation 36921498-ecb1-4a0e-91a2-fb8e7ed02af5 · outbound

This paper cites BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks BrowseComp: A Simple Yet Challenging Benchmark for Browsing Agents

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.000943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.000943Z digest=sha256:c3e512517194d6a09ce092765901079289d890cf3bfe4f821e51e1cfdd58fe5e

Observation 724771bf-4c12-4653-8bae-e6c3346ad70e · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.689143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.005709Z digest=sha256:b2dcbfcd6d039e339fa86c2d51395ea7923f1f325fe92c44aa8c40bfc78439c7

Observation 68ad9757-6bbe-4e79-ade1-bc9c04bace33 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.009988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.009988Z digest=sha256:d346e28add0bee22e8cbae7b893d727778742b29e234a23de26e295997b9c1b2

Observation 38416ba2-78e4-47d3-80f4-95207af0287a · outbound

This paper cites MCPWorld: A Unified Benchmarking Testbed for API, GUI, and Hybrid Computer Use Agents.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks MCPWorld: A Unified Benchmarking Testbed for API, GUI, and Hybrid Computer Use Agents

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.013951Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.013951Z digest=sha256:349b919b39cb84118f248388e0c459fa23a2853b1952938848b3cffd98d20619

Observation 59ad121f-0a27-4986-853b-a16fe0ca313d · outbound

This paper cites Qwen3 Technical Report.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Qwen3 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.018314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.018314Z digest=sha256:63a5b66885f454db290c329f9917fc29835149f8d73d5e7cdfe5490639c0b19a

Observation b723d179-826a-4f5f-b1e6-67a4452514f1 · outbound

This paper cites CRAG -- Comprehensive RAG Benchmark.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks CRAG -- Comprehensive RAG Benchmark

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.022856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.022856Z digest=sha256:d2d9843af49cd9424301114a7a1f47f24f0927ba87e3aba45676a276ac87315b

Observation 32f6f36d-2b77-44b7-a92c-7002ac18ede9 · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.673722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.027296Z digest=sha256:d711e7ce80d5c9a9cc335314f7f819ba1bcf0f1c248b65c58132dac0c1915bcd

Observation f2f16652-1457-4f3f-9436-37b43cc7607b · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.659795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.031542Z digest=sha256:d0b1cf1d5e702249ce55450181df2fd0481e29f2925cda2fe2a3cc0de73f6f4e

Observation 157d2c6a-6c36-4464-ad28-3a58bf85ba43 · outbound

This paper cites GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks GraCoRe: Benchmarking Graph Comprehension and Complex Reasoning in Large Language Models

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-04T22:55:34.161644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.035702Z digest=sha256:a527ebde4f42fed6126e714cd8cebce610262507c7c77561ca5d5644f6179597

Observation 74193268-e6cc-460e-9bac-035aac632c61 · outbound

This paper cites GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.040003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.040003Z digest=sha256:795ae35b4bffca90bfe700656934bdfc16b19035a459b33d6694662cb0a7907f

Observation 17f83712-0cfc-4b72-a504-641bfd895c0c · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.044618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.044618Z digest=sha256:6e8d6d22615dd497254dfba27f068d05e6fbf5d39debb4d327c63cc0e24450e0

Observation f8a6f197-db17-4842-abb6-d3bdb53f3acf · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.645340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.049003Z digest=sha256:9772a82780e31091af82326a7b1f78ce408b8146610734a218edf2f2e934fdf2

Observation a47e36eb-e5fc-424e-a678-82bb5e5ae36b · outbound

This paper cites an unresolved cited work.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-04T22:55:34.630620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.053433Z digest=sha256:b684e69d14ec47a115896adef122a8137a7fc9a71e1208594cd2e1adfa487299

Observation 9a3dfbe5-1a45-452a-9e72-65be6d87cc1d · outbound

This paper cites LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.058356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.058356Z digest=sha256:2da21563b7a9cd767a88d2235c0d44dab156259750ce17014ee648e219fc6d00

Observation df18fe85-7603-4509-85ab-3d4359b5ca87 · outbound

This paper cites analyze",.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks analyze",

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T22:55:34.615671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-04T22:55:34.062935Z digest=sha256:6238648c07580db202465035d2b7726b43ab9515fedf61e86b1120846f6b1303

Pith citing papers

No inbound Pith citation observations are available.