Pith. sign in

Paper Citation Record · LEDGER

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks

As of 11 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.07411.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.07411 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T05:03:11.987987Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact6
  • verified fuzzy2
  • unresolved17
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b29a3ab4-ba31-4f82-86ba-5941050d75d7 · outbound

This paper cites Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.859067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.859067Z digest=sha256:4a69af63e5fea1d9a337dfa5e8ec2b0f0c095c15cd54a1250c4a84e411ae1afc

Observation 1243e9bb-a124-4ac9-9f85-12dba1f655bb · outbound

This paper cites 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks 2026.GeoBenchmark: Probing Large Language Models for Geo-Spatial Knowledge

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.950320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.865139Z digest=sha256:3b9ad7b8231aef8676e0fec41224be187b219cc0be38a12a0fa79b7559bad8bb

Observation 0ae1e18d-e2c5-4cb3-9464-d3cc6cdf0ae0 · outbound

This paper cites MS MARCO: A Human Generated MAchine Reading COmprehension Dataset.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MS MARCO: A Human Generated MAchine Reading COmprehension Dataset

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.870231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.870231Z digest=sha256:2229634e7db8180cdb1e4056003c75440eedb01ae0bf9d34209c9535ac9392f8

Observation d1f12f62-fd5e-4890-aa91-28c67c241c34 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.935439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.875298Z digest=sha256:070e4c3cf856413cf9c43df6a1342e80fa99072fda51e010e71c1a255edfbb5e

Observation f31abbd8-dc67-471e-bf29-3069bb1c2fd1 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.920975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.884565Z digest=sha256:8124aba23347afd5868dd1f7cf7c85d3f0e32494e314ba1dd245e6a27540a1e6

Observation 898014ab-fabe-47cc-b51d-c82713aab38d · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 6

Resolution
malformed identifier
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.767628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.888820Z digest=sha256:bafeabc8927eb21d5642bb4c8ef11609c9c2ceff7ec19b2a22c266e690f6bf3f

Observation 18b06fc2-66e8-44c4-8bd7-8663da8aafe8 · outbound

This paper cites Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Kummerfeld, Li Zhang, Karthik Ra- manathan, Sesh Sadasivam, Rui Zhang, and Dragomir Radev

Reference 7

Resolution
verified exact
doi, observed 2026-08-10T05:03:12.036948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.893227Z digest=sha256:aa6336536324ea7c9c4bd84e5e3087aad52f5e39d1843e7ff0aa42315069cbc7

Observation 41bfce68-ec68-495a-a591-8bbe64bf0386 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.906588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.898261Z digest=sha256:578ef52b7de9d11bee4854b5edbc9ce560027f45a5ed954786974dac7911e964

Observation b4143ca2-f8a7-42b8-acbf-72e82efc93c6 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-10T05:03:11.902705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.902705Z digest=sha256:280b47ee41d067daf51cdc9540b9b4f23a59349f149fb43e8a334dd7b441ea15

Observation d1ddcc79-2481-45d6-8f98-b7a3d5e64b9a · outbound

This paper cites When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks When Retriever-Reader Meets Scenario-Based Multiple-Choice Questions

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.594227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.906911Z digest=sha256:e8f36e14ba31fada47b9c20ce51e7334f91577575715a7585ab3d2cb2fe09b1e

Observation 461d081f-8558-4617-bf5d-40bef1bc4ec3 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.882941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.911376Z digest=sha256:4399c82f611cbb7cd20a72848a24f9f503e24e291bf1f1e490922549c3fc980c

Observation 193be87d-49b2-4449-b6d3-764a7e3b3fce · outbound

This paper cites Location Aware Modular Biencoder for Tourism Question Answering.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Location Aware Modular Biencoder for Tourism Question Answering

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.571961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.916168Z digest=sha256:4be37aad6f6189ec6af8b3036a0333806389de3c654d4938dd5b3b27ed596df8

Observation 2e7a74a1-2726-4083-98aa-b16793dbc806 · outbound

This paper cites GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GridRoute: A Benchmark for LLM-Based Route Planning with Cardinal Movement in Grid Environments

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.920484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.920484Z digest=sha256:6418d0f097706b70b49e2cc468a7717da13880070a9fc99fbcd6f7eb61e0b78f

Observation d3a9abd7-488d-4076-a13c-b9e7f241dc99 · outbound

This paper cites STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks STBench: Assessing the Ability of Large Language Models in Spatio-Temporal Analysis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.925122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.925122Z digest=sha256:48455f71b54f025710bed0483d614c56ad137ff8393011d948a727899d4162e2

Observation a34b83a5-0c74-4b30-b27d-1620ab96a38a · outbound

This paper cites MapQA: Open-domain Geospatial Question Answering on Map Data.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks MapQA: Open-domain Geospatial Question Answering on Map Data

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.929941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.929941Z digest=sha256:1ff534f7445fd2db2b10a89cc38b5f0c4bd7eafc68983896516780a5705408db

Observation 80dbe486-ce64-400d-b954-2c3ec0deb3bb · outbound

This paper cites Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Geographic Question Answering: Challenges, Uniqueness, Classification, and Future Directions

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-10T05:03:12.503759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.934820Z digest=sha256:2de749d619e29af37ad72e0dde548e91d641632b612be30b4b0b9b22270250f6

Observation c21d67c9-5d70-4df1-b96e-3d8154952b83 · outbound

This paper cites GeoLLM: Extracting Geospatial Knowledge from Large Language Models.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks GeoLLM: Extracting Geospatial Knowledge from Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.939699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.939699Z digest=sha256:442506195f7ee1b298ae3e250386c9971155c47fb0696b412a6a63a8ca078cb5

Observation 95dcee66-3548-4287-b9a8-b9a604c62c60 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 18

Resolution
malformed identifier
raw_fallback, observed 2026-08-10T05:03:12.868619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.944499Z digest=sha256:bc41f03a2ff7525bf6c2f2ab633f4d07f952d8244b4249abee690802188db4a4

Observation 09bc2716-bf8c-4f47-91b4-0ffe1cc5b046 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.949570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.949570Z digest=sha256:583342c58626e114ea0c53ee0f31336f73031bcdf4448992b93f7f53aa9ca194

Observation e7fcb46f-6059-432e-91fe-444767dbe26e · outbound

This paper cites Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Towards AI-Complete Question Answering: A Set of Prerequisite Toy Tasks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.954622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.954622Z digest=sha256:e0dcf71a25ce9b7c304c2917b870b998bfb0301be074384176fbd5f4bc3b5e53

Observation c9526de0-46b9-4c11-8933-710b74330e2f · outbound

This paper cites Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Evaluating Large Language Models on Spatial Tasks: A Multi-Task Benchmarking Study

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.959631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.959631Z digest=sha256:34e9f26d83aa0b36e06a88fc3cbabcbf4e40e785ebaa7076e03b4984e13707ff

Observation 78f93cbe-31e3-45f4-9e16-1c6ee9a1e4a4 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.964369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.964369Z digest=sha256:3069f84e0e68ec63d705b2de5b2752625d488267c2d9a623d2f755994bd624fc

Observation fcda4f81-077c-4bca-aa95-1c4558a5ddb9 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 23

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.434184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.973919Z digest=sha256:836c395a58ca24175e2cc81852b144e43679a1cd5fd546c0230930d73d9c3cce

Observation 8624ffda-299e-49f4-bfac-82692c81a839 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 24

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-10T05:03:12.252810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.978541Z digest=sha256:6faf85dd7540b1d1fe181f215262b8f09c96fe302787afcdc1b84098dc2c0a99

Observation 6c35c5d7-cc75-42eb-97ce-9a8d03fe1a13 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.828565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.983283Z digest=sha256:1e5dfa612f740bbed33f7012f29ffaa5323e43b4816011e32aadfb1b19211068

Observation 6697a866-c5c8-489a-adca-4d75d1c6d205 · outbound

This paper cites Zelle and Raymond J.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Zelle and Raymond J

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T05:03:12.813608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.987987Z digest=sha256:a4ec6def53c381eae001f8e692ac21b68e96a9f672798356ce564b24fa79f20c

Observation cd7dade0-df3f-4d1a-adb4-09abed50e8a3 · outbound

This paper cites In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21).

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks In Proceedings of the 30th ACM International Conference on Information & Knowledge Management(Virtual Event, Queensland, Australia)(CIKM ’21)

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-10T05:03:11.880072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:03:11.880072Z digest=sha256:98867a0c1cfdc07c6f804a96373e42a4baddaf9c6c24f9217809c68717413c73

Observation c8cc3774-ee36-469c-af82-f0fe177c1202 · outbound

This paper cites an unresolved cited work.

GeoBenchLLM: A Comprehensive Benchmark for Evaluating LLMs on Geo-Related Tasks Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-10T05:03:12.844144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-10T05:03:11.969114Z digest=sha256:ee0817c619fceb2c99e5985768d5a52ef34eff80c28516683c2050970f15618f

Pith citing papers

No inbound Pith citation observations are available.