Pith. sign in

Paper Citation Record · LEDGER

Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2304.11164.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2304.11164 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:56:32.368266Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:28:59.707323Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6ba9befe-17fb-4703-8e8c-ccf3433b87fe · inbound

Can Large Language Models Reason about the Region Connection Calculus? cites this paper.

Can Large Language Models Reason about the Region Connection Calculus? Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T06:06:07.564890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T06:06:07.564890Z digest=sha256:7567b2b232487592361e6af04cbaaf40c695eefb84e952d34be9e65dc09a9eff

Observation efc64125-888e-4f3f-a9e2-08e2b4294b78 · inbound

Geospatial Mechanistic Interpretability of Large Language Models cites this paper.

Geospatial Mechanistic Interpretability of Large Language Models Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T23:56:32.368266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:56:32.368266Z digest=sha256:bff12a0331a802d4774f363b2516d32e0513aa58bc1911f561b308736a970138

Observation 3c39ff05-dc27-4795-82cd-3a5d395a57ef · inbound

Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations cites this paper.

Foundation Models for Geospatial Reasoning: Assessing Capabilities of Large Language Models in Understanding Geometries and Topological Spatial Relations Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:08:08.025968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:08:08.025968Z digest=sha256:246f7c7676a8f875844830aa650a60ac5170d3bb68f9ab82116dcaa663963904

Observation be08a52a-dfe6-45da-bd06-bf7d3d97ef18 · inbound

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark cites this paper.

From Reasoning to Generalization: Knowledge-Augmented LLMs for ARC Benchmark Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:50:51.103242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:50:51.103242Z digest=sha256:75ebc3fceb0cef307cf10444c506207747feee5f0407539167ab8a9d70c79a38

Observation 17eaa702-14d6-46c0-b172-226e24bd879a · inbound

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models cites this paper.

MazeEval: A Benchmark for Testing Sequential Decision-Making in Language Models Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T13:40:01.425301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:40:01.425301Z digest=sha256:e48c66674494954d4ff6a3ed60b44e269f4eeb797189bc91fcd7f3eca84900e5

Observation 6ad35b72-3ee1-4c6d-b81f-bbdcb287468a · inbound

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi cites this paper.

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:18:13.594209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T11:18:08.326304Z digest=sha256:57cd0888f62479b10cb852c1c6689f7acaee00c4b3978cb59fff87de21f230dc

Observation 09b2c60a-4f3c-41ca-9161-cba22ef04fd3 · inbound

GS-QA: A Benchmark for Geospatial Question Answering cites this paper.

GS-QA: A Benchmark for Geospatial Question Answering Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:34:31.671810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T02:33:47.304356Z digest=sha256:7fd56641e2d1c923caf976f0f8ef0da7a1bd9fb293214d3e3acc944d335b20b2

Observation 7119106f-d003-40e8-92ad-4debd76a80a3 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.484476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:0424035bdf89f7569c9f2a3d26f77473f026fe7d8273c8b5b5c0da15be85591b

Observation 828d2b03-da5c-487d-91ec-f89c92fb5c1b · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation Dialectical language model evaluation: An initial appraisal of the commonsense spatial reasoning abilities of LLMs

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.709846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:eb31d44dfdcb13041e137f750d59384396cad861ae001ea370e0fccc8aebae8b