Pith. sign in

Paper Citation Record · LEDGER

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.07411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07411 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:03:11.987987Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy2
  • unresolved17
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b29a3ab4-ba31-4f82-86ba-5941050d75d7 · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.859067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.859067Z digest=sha256:0f91b0a1f366171ccad7d5ceed46b1820afcdbb351be90e5e991962bdf2e6377

Observation 1243e9bb-a124-4ac9-9f85-12dba1f655bb · outbound

This paper cites 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.950320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.865139Z digest=sha256:6e77d8d7cd111e852e8552e6cfafa6d346d0a84e48eeee11b68c1a43221bcac5

Observation 0ae1e18d-e2c5-4cb3-9464-d3cc6cdf0ae0 · outbound

This paper cites MS MARCO: A Human Generated MAchine Reading COmprehension Dataset.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.870231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.870231Z digest=sha256:06bf032bb15c403ba934f8df0bb87c20ad93a10690042c9a0d4baf052bc04134

Observation d1f12f62-fd5e-4890-aa91-28c67c241c34 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.935439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.875298Z digest=sha256:419d38387bc774a814df105ccb3fcf18a4e5895c881c172622dfb9561869a2e2

Observation f31abbd8-dc67-471e-bf29-3069bb1c2fd1 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.884565Z digest=sha256:c144347e5b4f9115b8403f7f79f43b06600e03dc870493d2ed2ba55a5bf34893

Observation 898014ab-fabe-47cc-b51d-c82713aab38d · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 6

Resolution
malformed identifier
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.767628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.888820Z digest=sha256:98a372c4d82b02f25d0f8adefd78d93ca80873621e724b3b0f9e71428a98cfd3

Observation 18b06fc2-66e8-44c4-8bd7-8663da8aafe8 · outbound

This paper cites Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev

Reference 7

Resolution
verified exact
doi, observed 2026-08-10T05:03:12.036948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.893227Z digest=sha256:a1fcf570330bb68ff5253e9e4f597d4116ec81cd05a58246a1bed0c140ac4870

Observation 41bfce68-ec68-495a-a591-8bbe64bf0386 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.906588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.898261Z digest=sha256:162cfad9076d0142ccfb876f823099508aa416ab6d34e6ed55052831f752f6ca

Observation b4143ca2-f8a7-42b8-acbf-72e82efc93c6 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-10T05:03:11.902705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.902705Z digest=sha256:ef90bb4f36979a13a848c1ac7cd9d96c5b0cb7a6406c69370063dc52c06c284d

Observation d1ddcc79-2481-45d6-8f98-b7a3d5e64b9a · outbound

This paper cites When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.594227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.906911Z digest=sha256:e6056a03121c5a053bc1e6b3afb28f710e88a8af2dc0062903b1323f2cbd7770

Observation 461d081f-8558-4617-bf5d-40bef1bc4ec3 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.882941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.911376Z digest=sha256:1889bb1364490222b631c2503b3f2f6898c9fa0429ee719bd3373ed1bdc0d2c3

Observation 193be87d-49b2-4449-b6d3-764a7e3b3fce · outbound

This paper cites Location Aware Modular Biencoder for Tourism Question Answering.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Location Aware Modular Biencoder for Tourism Question Answering

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.571961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.916168Z digest=sha256:c7a29ce0aab7dfb3b7ef7a2c1fb4f88edf956c501b0b68384ecd5754d2b559bb

Observation 2e7a74a1-2726-4083-98aa-b16793dbc806 · outbound

This paper cites GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.920484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.920484Z digest=sha256:a13a75375d89fbb3f51c955aad4757a8f3c035bfe68bae9b111357a2535f1777

Observation d3a9abd7-488d-4076-a13c-b9e7f241dc99 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.925122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.925122Z digest=sha256:69fecd9c0d911a7b5adafedd234a257342cc90bdbc0899dd46987fa44c8bd8cc

Observation a34b83a5-0c74-4b30-b27d-1620ab96a38a · outbound

This paper cites MapQA: Open-domain Geospatial Question Answering on Map Data.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MapQA: Open-domain Geospatial Question Answering on Map Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.929941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.929941Z digest=sha256:2b75da08d6b5e4ac5ecbafd247abe6071b13942936f44871be8f697953cd17c1

Observation 80dbe486-ce64-400d-b954-2c3ec0deb3bb · outbound

This paper cites Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.503759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.934820Z digest=sha256:49de398ef08861d55b3307834e6397a4cce8342e4b4dc7e032fcee8c6b30a714

Observation c21d67c9-5d70-4df1-b96e-3d8154952b83 · outbound

This paper cites GeoLLM: Extracting Geospatial Knowledge from Large Language Models.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GeoLLM: Extracting Geospatial Knowledge from Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.939699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.939699Z digest=sha256:7fe09ec8aee914655a2c40bdd936b2cfff311e7909a288e0a8bf70fce0311ff4

Observation 95dcee66-3548-4287-b9a8-b9a604c62c60 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T05:03:12.868619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.944499Z digest=sha256:edd0d1a2bcc312619955e19b74d4eb33fc5e938c470f2f5b2050d903e730972b

Observation 09bc2716-bf8c-4f47-91b4-0ffe1cc5b046 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.949570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.949570Z digest=sha256:cda81ac70ae06c20fba7ad9bfd569f030a9b0ee3d9d76d6cf29af3937cc245c8

Observation e7fcb46f-6059-432e-91fe-444767dbe26e · outbound

This paper cites Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.954622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.954622Z digest=sha256:c1e472aa9b8c3ca1867d5eadcc1f6550f806aeb443ecfacd002e4fab19b5bae8

Observation c9526de0-46b9-4c11-8933-710b74330e2f · outbound

This paper cites Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.959631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.959631Z digest=sha256:e4eb62823458f66985252d017f027ac03d29a79781a3f4b4da4332b8e51999e4

Observation 78f93cbe-31e3-45f4-9e16-1c6ee9a1e4a4 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.964369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.964369Z digest=sha256:b550f3add64198e63170527d6342486103f6f3c3ee21287490f8a4a95f463858

Observation fcda4f81-077c-4bca-aa95-1c4558a5ddb9 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.434184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.973919Z digest=sha256:98eebd80d35dc835dbf387c02be17c5a0b23bc7f10bb3169283290ea6bad8dcb

Observation 8624ffda-299e-49f4-bfac-82692c81a839 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.252810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.978541Z digest=sha256:e286ba0f3c9e13ac3524b52e48621e68acafef107aafa0bce3ab5829f1f2f089

Observation 6c35c5d7-cc75-42eb-97ce-9a8d03fe1a13 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.828565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.983283Z digest=sha256:70600f6a5e2bdeb59facbe2b75da3735c091ff34c0e5034702ceb5b15c3db5e8

Observation 6697a866-c5c8-489a-adca-4d75d1c6d205 · outbound

This paper cites Zelle and Raymond J.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Zelle and Raymond J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.813608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.987987Z digest=sha256:cfad04801ac42ba06480ecb8c408f46f2edbe9a56b9d00659c2fe6790a70fdf2

Observation cd7dade0-df3f-4d1a-adb4-09abed50e8a3 · outbound

This paper cites In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21).

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21)

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.880072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.880072Z digest=sha256:5ab2b4362363c73e59bdf925efd4215314fc0280ccf1cad3d557b2cc81d60b19

Observation c8cc3774-ee36-469c-af82-f0fe177c1202 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.844144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.969114Z digest=sha256:028cfe5ea2e39c93145a11faaf3b18caaaf6397efed0a701b8e3b13a01ec05de

Pith citing papers

No inbound Pith citation observations are available.