Pith. sign in

Paper Citation Record · LEDGER

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

As of 1 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2501.10074.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10074 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-01T06:32:01.292127+00:00

measured 22 of 22 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T09:09:24.863958Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:28:59.692534Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 71090c80-2619-4d3f-9187-59ef4726c803 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.227502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:62c38452e562e8a403ba7e336cd137d63a84672cc8e12081192ed4ccaed33705

Observation 5dad4a7b-f9b2-4a70-b6c6-d4b53cb38a56 · inbound

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models cites this paper.

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:10:48.842964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-05-18T03:09:09.713822Z digest=sha256:7e60b91591733b3a6af22552ae3cf8a53603361085c173e2c4dd391e6c44977b

Observation b30b3fe8-a60c-4285-afbf-d40364a3d9eb · inbound

SCP: Spatial Causal Prediction in Video cites this paper.

SCP: Spatial Causal Prediction in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:11.146675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-15T16:47:44.523606Z digest=sha256:ad4069555efba0990c33c51405e078386d4c55d68870b6d19800998a97462201

Observation 4de075f5-7088-49b8-9dab-da2a933f0d85 · inbound

Token Warping Helps MLLMs Look from Nearby Viewpoints cites this paper.

Token Warping Helps MLLMs Look from Nearby Viewpoints SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:08:17.361639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T21:07:55.062113Z digest=sha256:05b106670337c32ca27aa65319b1801b25752f37f333424ef617e977b1fd6214

Observation 5779c4ba-d08c-4616-b328-18e48d2f23df · inbound

Spatio-Temporal Grounding of Large Language Models from Perception Streams cites this paper.

Spatio-Temporal Grounding of Large Language Models from Perception Streams SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:02.684607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-10T17:10:45.837684Z digest=sha256:e637832082073dba0bb45bfb5826eefad94e1e472bdd50990c5e10d092e5cd39

Observation b6d757a7-7c85-4524-a156-a2e9aadef4ea · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:11.747234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-08T14:20:08.404090Z digest=sha256:859b1ccb946bb76c89ec44e81710acf4c002448fed1bad4638a928187224083c

Observation 735994be-fcfe-4f8e-b488-ed43f427ee0b · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:16:39.949775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-25T06:15:33.062980Z digest=sha256:0c5ed643507529441ac04419d0853eef2271df75e83b5e15dc2924689a66cc9b

Observation 7d5b6bdc-d4c0-4847-b0dc-db113d0d9d6d · inbound

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment cites this paper.

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:59.892965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-11T01:55:39.721560Z digest=sha256:6bb5086b61215f8f50a0333dce2023d66bd921cc4d5d5402f1db75e87009cdca

Observation 6f9e0018-747c-4b0d-b457-958542a9d008 · inbound

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images cites this paper.

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:01.622856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-13T01:26:47.052025Z digest=sha256:b9c0a14b14516d8100106dc0543c3f0d92bde7e2cb36c23a65d993c72dc9c767

Observation e69a6f8e-d685-4842-8d40-009b644c3d13 · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:53:13.184532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-20T10:52:22.778489Z digest=sha256:f945b32616426c5e54a7e8f45032386e9e05de519fc3d31dabaffef9826a8174

Observation 89df9b66-4673-47fa-82f8-04cce5bdfade · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:05:47.169697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T18:25:17.831116Z digest=sha256:dfdfd891fb042acdad558cb0661ddfb3b845eb7e274d8b92447638f575c97817

Observation 30082d11-5f43-49ce-a533-2931f22751ee · inbound

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches cites this paper.

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:03:58.061704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-21T05:00:11.848526Z digest=sha256:e86ed2036f7f59f43835398e095287ec8f128330e2a654d33a160c24ff5bfae2

Observation f0d63128-776f-4dc8-92fb-3379e8744763 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:15:22.732097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-05-25T05:10:32.522453Z digest=sha256:26998ad399516762d047c64085a4d35de47ffd1f00fef123adbcc6b8ed919cf3

Observation c178bf64-6215-4858-8932-b616a2f27228 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:44:56.062222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-30T16:40:22.441025Z digest=sha256:275c2fe5969d8582458c9db1681dce95f9eb254dd6a740cd806fdbad7bc3e642

Observation f2fb586f-e6fd-49cc-a50b-6a6d5ebf95c2 · inbound

Grounded 3D-Aware Spatial Vision-Language Modeling cites this paper.

Grounded 3D-Aware Spatial Vision-Language Modeling SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.342690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-29T08:08:36.012761Z digest=sha256:71f2b4db29d82c18a15f861f03f534281822b855b5c73dc53997002056b18e62

Observation 8cc528be-b179-4200-b688-23ea88a620e4 · inbound

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks cites this paper.

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.607933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-06-27T16:35:14.099586Z digest=sha256:70caaf5af914e879580f5a01a55c2225d9710e8989320b7152bfac9afc41f142

Observation 32232175-98d9-4877-ba72-d02745cd166f · inbound

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models cites this paper.

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:55.601454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-06-27T01:42:30.005911Z digest=sha256:43d44e443a19e0affe0e2e3efeeb048012c1110c4adf3f927caf2c6ed386995f

Observation 4344ecef-9e55-4a5e-9e19-5f42073828f2 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.479409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:0167672fc580a13b1c8c5c6d9af40e1b46de1a2c5c0267d3bc556072a224618d

Observation f481d8e6-21b1-4613-9024-233ef425327c · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.694996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:ca21f4d38e7706e2836e2baf9d792f4a35bd0f9adb0a7656aa9d19eb2356f618

Observation 5317129b-a50e-4e53-882f-79221021b8ed · inbound

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping cites this paper.

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-02T14:17:02.420247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-02T14:16:39.649823Z digest=sha256:49040a7034c7c79463cfce4016503ab0eaf255daa4052b5ee99760fa2efbf464

Observation 490f06ad-ce8e-40d3-a95f-7ae53961e6b1 · inbound

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video cites this paper.

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:18:37.311510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-01T06:32:01.292127+00:00.

source=pdf_text observed=2026-07-03T16:16:41.412451Z digest=sha256:891a8a4678f2fef8acd4e6cda1e60c3f16af0068e892717fc7189a779ddf117a

Observation ddc2f67e-6117-4bc4-8b82-5fb77958e9ed · inbound

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? cites this paper.

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T09:09:24.863958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:09:24.863958Z digest=sha256:ff7556a5fd1130f5775c29aab43c9ae13cf6568e2ec42594084137581d684a3e