Pith. sign in

Paper Citation Record · LEDGER

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.07411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07411 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:03:11.987987Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy2
  • unresolved17
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b29a3ab4-ba31-4f82-86ba-5941050d75d7 · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.859067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.859067Z digest=sha256:950a58086ed097470a918dbe18b23bbb40e9b93678b015f9ae64cbce7ac4e66e

Observation 1243e9bb-a124-4ac9-9f85-12dba1f655bb · outbound

This paper cites 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.950320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.865139Z digest=sha256:69189ca73f3a24f18f46451da8d071af0838df3ef57d779196fbc4b576ff37e2

Observation 0ae1e18d-e2c5-4cb3-9464-d3cc6cdf0ae0 · outbound

This paper cites MS MARCO: A Human Generated MAchine Reading COmprehension Dataset.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.870231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.870231Z digest=sha256:ac21a1ff158d37f7558ab4c5423923ef883b671aa578aa5474fa36de1549e709

Observation d1f12f62-fd5e-4890-aa91-28c67c241c34 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.935439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.875298Z digest=sha256:31899f920ca400283420f1a99019d5f5cb63e50d0a39e1d3c9d0dcc2bbd337c7

Observation f31abbd8-dc67-471e-bf29-3069bb1c2fd1 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.884565Z digest=sha256:c257aee8efdc9777b962864dd7843a3a60e912cab6abb98d7d66efd99b734169

Observation 898014ab-fabe-47cc-b51d-c82713aab38d · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 6

Resolution
malformed identifier
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.767628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.888820Z digest=sha256:3242f253e25a3d880ac692fe10a859722a078d196d83c644eb2fc998f07f85d7

Observation 18b06fc2-66e8-44c4-8bd7-8663da8aafe8 · outbound

This paper cites Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev

Reference 7

Resolution
verified exact
doi, observed 2026-08-10T05:03:12.036948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.893227Z digest=sha256:0b4958c66ff93d1ca10f536dc039adcd2092a0eb1700f5cffecef44565168fc8

Observation 41bfce68-ec68-495a-a591-8bbe64bf0386 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.906588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.898261Z digest=sha256:3e9e0e5344c56469a822d7e9bf8354c2ce765ed4d16149ca967ecd7f2f8d2d87

Observation b4143ca2-f8a7-42b8-acbf-72e82efc93c6 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-10T05:03:11.902705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.902705Z digest=sha256:9dc0192985231256ceb55cc26830dd143a82628a0dc91eb8f9a6e2506d502fa6

Observation d1ddcc79-2481-45d6-8f98-b7a3d5e64b9a · outbound

This paper cites When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.594227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.906911Z digest=sha256:d2424bb18c2f4941bbb4efded938fd38b22401274273616492dd7598c2a59b79

Observation 461d081f-8558-4617-bf5d-40bef1bc4ec3 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.882941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.911376Z digest=sha256:97d71d9243cdb3ae04bc46220c27f5ce62659b9bce02907e6877d4afed8222b2

Observation 193be87d-49b2-4449-b6d3-764a7e3b3fce · outbound

This paper cites Location Aware Modular Biencoder for Tourism Question Answering.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Location Aware Modular Biencoder for Tourism Question Answering

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.571961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.916168Z digest=sha256:97e9b53166b847f55862d9c51a6ff4ecd11e1d4094b2566ab236e9ae690634e0

Observation 2e7a74a1-2726-4083-98aa-b16793dbc806 · outbound

This paper cites GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.920484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.920484Z digest=sha256:9c3081e30bedc5c7dbfa1c1eff1d8b6f243ad36c300fda37ec5f32ea0934a48e

Observation d3a9abd7-488d-4076-a13c-b9e7f241dc99 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.925122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.925122Z digest=sha256:0291201f16eb479ed8f6450b1db257119d8e1bd17aec93c2eb564e5f7616231b

Observation a34b83a5-0c74-4b30-b27d-1620ab96a38a · outbound

This paper cites MapQA: Open-domain Geospatial Question Answering on Map Data.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MapQA: Open-domain Geospatial Question Answering on Map Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.929941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.929941Z digest=sha256:9d7385f2b5cefa6e142c0e3b22f33d11c4cceb1a052bb37bc53d6ccd3f594a2a

Observation 80dbe486-ce64-400d-b954-2c3ec0deb3bb · outbound

This paper cites Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.503759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.934820Z digest=sha256:c5125f0e3e809316776a85b172dd6d1527cb0ab8efd3f05aa10bf85b69b01bae

Observation c21d67c9-5d70-4df1-b96e-3d8154952b83 · outbound

This paper cites GeoLLM: Extracting Geospatial Knowledge from Large Language Models.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GeoLLM: Extracting Geospatial Knowledge from Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.939699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.939699Z digest=sha256:bd1291b84073adc246deacfbad20fe903332fff0f02e7cde1fc24ef177261fdd

Observation 95dcee66-3548-4287-b9a8-b9a604c62c60 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T05:03:12.868619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.944499Z digest=sha256:37908f228e56cf72bdbee3f524ae238a3fcee48c3d4b5534f1343a4082f56713

Observation 09bc2716-bf8c-4f47-91b4-0ffe1cc5b046 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.949570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.949570Z digest=sha256:c8059ba4d5e3a43b8c867f9d830f9e383594dd25b0c68f12e187c7ef3b649711

Observation e7fcb46f-6059-432e-91fe-444767dbe26e · outbound

This paper cites Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.954622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.954622Z digest=sha256:4ec5986917a9bc15ab633214ee4a0ab4ba1bd8ea32f67cd44a4dec1440291311

Observation c9526de0-46b9-4c11-8933-710b74330e2f · outbound

This paper cites Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.959631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.959631Z digest=sha256:108f14ee3579211c10845fac5d7e5652db1430f9884918d7077ede12d54b2239

Observation 78f93cbe-31e3-45f4-9e16-1c6ee9a1e4a4 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.964369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.964369Z digest=sha256:6936786e6d71ab0f3448504f35a9fb3867501f34149381ffa377f35447360492

Observation fcda4f81-077c-4bca-aa95-1c4558a5ddb9 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.434184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.973919Z digest=sha256:c8d6eafb06539e21ce6fb8195edc9b9488e3313121a009a3ed973625c744b419

Observation 8624ffda-299e-49f4-bfac-82692c81a839 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.252810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.978541Z digest=sha256:a3905db7823abc7c361c8ca1459272ba149de29224afad5834448d3c527c86c2

Observation 6c35c5d7-cc75-42eb-97ce-9a8d03fe1a13 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.828565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.983283Z digest=sha256:e26cb12e28f0d71ddc1253266a12f616ecc21aff0f114d32f8c1e81f3b67ca5f

Observation 6697a866-c5c8-489a-adca-4d75d1c6d205 · outbound

This paper cites Zelle and Raymond J.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Zelle and Raymond J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.813608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.987987Z digest=sha256:9795111e8211fffd6a632007a38a9303c745f144750500660c1cb2bb69c3085f

Observation cd7dade0-df3f-4d1a-adb4-09abed50e8a3 · outbound

This paper cites In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21).

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21)

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.880072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.880072Z digest=sha256:49c767078c711395f11b35b70486878332c9c81bf994e3c0a12cb10346d82eec

Observation c8cc3774-ee36-469c-af82-f0fe177c1202 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.844144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-08-10T05:03:11.969114Z digest=sha256:e1ab08723f81953084be91955f277622f2a526ea893dc5fd567c685db783d4b2

Pith citing papers

No inbound Pith citation observations are available.