Pith. sign in

Paper Citation Record · LEDGER

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

As of 22 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 34 inbound Pith citation observations for arXiv:2501.10074.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10074 v3

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T19:27:41.318182Z

measured 71 of 71 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:02:43.415017Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:28:59.692534Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy16
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 927de391-c499-411a-86ee-264c46115d73 · outbound

This paper cites Llama 3.2-vision 11b.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Llama 3.2-vision 11b

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.980862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.141628Z digest=sha256:86a77ea9a09400e889740be771f4bfbb361f31f2cfa7d89fa10edf0342d63e06

Observation 2cf7cd93-3c66-4d17-817b-ad1d86a00336 · outbound

This paper cites Graph of thoughts: Solving elaborate problems with large language models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Graph of thoughts: Solving elaborate problems with large language models

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.967467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.148080Z digest=sha256:79e0faa72af4942a492746e751249d344c38b6b13ab2cf6fefa9b7ba187618f4

Observation 3cca13cc-2ab9-48d2-b96f-b35498f97bfd · outbound

This paper cites Spa- tialvlm: Endowing vision-language models with spa- tial reasoning capabilities.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Spa- tialvlm: Endowing vision-language models with spa- tial reasoning capabilities

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.953787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.153336Z digest=sha256:f9329ca18c561af6ce859afd2faaa52f5f928dc67ceb841b781426f15be54f6b

Observation 54c3d290-1983-4d2a-b8da-c92a795f7c8e · outbound

This paper cites Decision transformer: Reinforcement learning via sequence modeling.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Decision transformer: Reinforcement learning via sequence modeling

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.937778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.158358Z digest=sha256:7f85db37bafbea6b6452fe87ef18bff040c895652a46e78f598f6609f5874148

Observation a7d83a61-6c28-4aea-8018-158965d44b22 · outbound

This paper cites PaLI: A Jointly-Scaled Multilingual Language-Image Model.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning PaLI: A Jointly-Scaled Multilingual Language-Image Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.162719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.162719Z digest=sha256:ff0d991c2daaffd71ac8d95cfe06f7e5647179618b94b44964e7f2f55746cb78

Observation 89f02065-ef9a-4bc1-b2ec-9e5d976aa0b8 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.167762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.167762Z digest=sha256:02f8e88d6b7449544ef07adbbe5ba728eff04990b6a3a341691e8223069f44df

Observation 7e067972-00c5-4e9e-b842-868d8a23d66e · outbound

This paper cites Diffusion policy: Visuomotor policy learning via action diffusion.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Diffusion policy: Visuomotor policy learning via action diffusion

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.921220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.172797Z digest=sha256:8ed89fda128e855e37090832e9bab355fc9dbbf4e1dabb491160dad0f4d99364

Observation 7d4dec4f-11b3-4fd0-98f2-456a5634e555 · outbound

This paper cites Vicuna: An open-source chatbot impressing gpt-4 with 90% chatgpt quality.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Vicuna: An open-source chatbot impressing gpt-4 with 90% chatgpt quality

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.906044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.177123Z digest=sha256:e116c36d8819de0a210e8d08e03ef4c3780b5acd418ac33a01824cad69e55ffd

Observation 43a90f14-1e44-4877-b27c-d70f37426f97 · outbound

This paper cites Blender - a 3D modelling and rendering package.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Blender - a 3D modelling and rendering package

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.890339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.182143Z digest=sha256:d60031d2f09db24c31a809e0de6fa2e74d26092c948d8834bbb021c48fa94ffb

Observation 7139171f-eb7f-47ad-a53d-b07ab2219d38 · outbound

This paper cites PaLM-E: An Embodied Multimodal Language Model.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning PaLM-E: An Embodied Multimodal Language Model

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.186928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.186928Z digest=sha256:1f22573f8a9d1d464e87c27438c07a2b9a4b27b9bfe88450944560fffe5ef24c

Observation c7468055-8085-408b-8a51-351f731a8958 · outbound

This paper cites Scaling up and distilling down: Language-guided robot skill ac- quisition.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Scaling up and distilling down: Language-guided robot skill ac- quisition

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.876410Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.192698Z digest=sha256:7b1fd3c41fb0bc57d8448cf938f1898262acd2637bd4e9c9de82e6f64c1a6d26

Observation f28008bc-47aa-4e60-8ad5-037773ee2cbf · outbound

This paper cites Embosr: Embodied spatial reasoning for en- hanced situated question answering in 3d scenes.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Embosr: Embodied spatial reasoning for en- hanced situated question answering in 3d scenes

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.863679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.197857Z digest=sha256:e0097f46da804503c985061415694980e4f5b1b0a251026dff1c19d2da6b30b2

Observation be9e595a-7846-412e-a880-b55f6e4ecaf6 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning LoRA: Low-Rank Adaptation of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.203151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.203151Z digest=sha256:60ce4310a4371dda94409329ec8cc1bfe7d11d51141dfdc9fcdefac00ccfe1c4

Observation 17382584-7dd3-4ab2-9e2c-22e54235d5af · outbound

This paper cites Inner Monologue: Embodied Reasoning through Planning with Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Inner Monologue: Embodied Reasoning through Planning with Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.209278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.209278Z digest=sha256:43013a34717650d974f1c143df858a51c41443bdd5d58f3f94c966957b062c3a

Observation ea34b9da-9c5a-4551-b327-a3eb17fdb33f · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and com- positional question answering.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Gqa: A new dataset for real-world visual reasoning and com- positional question answering

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.215126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.215126Z digest=sha256:cf45bc57692bd11d0607e3a8b4b2664ce0d192974a0c40c2baa0acdef4f0c459

Observation b4dd435e-7536-471f-9547-e51413e5cfc6 · outbound

This paper cites Habitat synthetic scenes dataset (hssd- 200): An analysis of 3d scene scale and realism trade- offs for objectgoal navigation.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Habitat synthetic scenes dataset (hssd- 200): An analysis of 3d scene scale and realism trade- offs for objectgoal navigation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.841262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.220436Z digest=sha256:a4592e822a579036141b81477f5b0d48851d9e94790f4aa23cdea709642fa013

Observation 5bdbd71b-5ece-4a1d-8f5d-32f713de8ebb · outbound

This paper cites Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large lan- guage models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large lan- guage models

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.826129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.226595Z digest=sha256:d64e9b5f0cef60a6e0d5f10e320d0c71a0525be3daf9eec1cf16846b5eb3d9d7

Observation d7dbbbf8-a4a9-49fd-8cce-f69f7b9ccdbc · outbound

This paper cites Code as policies: Language model programs for embodied control.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Code as policies: Language model programs for embodied control

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.810858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.231258Z digest=sha256:2b9572d3835ded097ca8f561cc3b893d9d4a12ee6ba2a85174c3c046b2678fc6

Observation 4deae3aa-6d87-43ad-9bc4-8eb13625e369 · outbound

This paper cites A Comprehensive Overview of Large Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning A Comprehensive Overview of Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.236015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.236015Z digest=sha256:2d250912d7f7abd153f51e52c82925fe2a5e0022306a8befeacbe59d96b0f56c

Observation e8e5eb32-b0fb-4b8f-a832-9b2e748f8797 · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.241491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.241491Z digest=sha256:8a00d7fa93eb05b9f5583250521e87bb9274694d45341fa4ae51f046dbc8aa56

Observation 570406d9-3bd2-4e39-b981-cd05125fa5c7 · outbound

This paper cites Habitat 3.0: A Co-Habitat for Humans, Avatars and Robots.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Habitat 3.0: A Co-Habitat for Humans, Avatars and Robots

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.245823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.245823Z digest=sha256:f4d3d696555aed53636aee80980d4901da85986cee8d73e5ec3441a0ed642a5f

Observation 551ddee2-c12c-4dcf-944c-881b6079f446 · outbound

This paper cites GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.250052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.250052Z digest=sha256:09086a92ef59f174d2776885b2cedd557e8d3f50d9f369359fbee05896a2594a

Observation 1fecb30a-d1a3-4a46-8e4b-1d4462bd8ea4 · outbound

This paper cites Robospa- tial: Teaching spatial understanding to 2d and 3d vision-language models for robotics.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Robospa- tial: Teaching spatial understanding to 2d and 3d vision-language models for robotics

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.254487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.254487Z digest=sha256:83ad5ce2f17f4eeb95478cb0cc3ab265f1034c3c4cb8ae0eb43eadd852bfbae7

Observation 9ef58d20-b83e-455a-8933-721507a843fc · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Chain-of-thought prompting elicits reasoning in large language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.795480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.258784Z digest=sha256:a994e3e483c423f85770f05de3f08a62c3c61d0de9263405d1f05a3a3727483b

Observation 0f8b76e2-bc52-4ec0-8e55-7f1438181e21 · outbound

This paper cites NExT-GPT: Any-to-Any Multimodal LLM.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning NExT-GPT: Any-to-Any Multimodal LLM

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.262686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.262686Z digest=sha256:85e53b6e943ecfa911357287f4ae0d819129ce070b5a9d1738b9af279009647f

Observation 7b05e6c2-9bfb-4110-8eda-ac495284bec4 · outbound

This paper cites Sapien: A simulated part- based interactive environment.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Sapien: A simulated part- based interactive environment

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.778842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.266717Z digest=sha256:cd421c2dc437abae3c31d7fbcfcd5ecd5c15ab1163a9eb87815f10c41dbabdfe

Observation 8e69d494-0637-41eb-9e09-0b30c6cba36b · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.271083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.271083Z digest=sha256:7f6c49b62f2b9def7ee11ee4472e6017602185b916c72319ad89cd42883bac34

Observation 7f8628d6-d69a-4534-bb10-61813cca9605 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning ReAct: Synergizing Reasoning and Acting in Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.275181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.275181Z digest=sha256:a6ea7a17935c180ba4dfe176b2ea7e9b2c29e713ae80f50e67b5e56eb2731d4a

Observation 3f638032-c87b-4a3d-8575-885288daa94d · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Tree of thoughts: Deliberate problem solving with large language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.764115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.279879Z digest=sha256:133914179a4b81dfa645865f580bc7f7358100cc3bdf4cf032fd12b1b89f085e

Observation a1efe0ae-50aa-423a-9d52-4c2359f9e48f · outbound

This paper cites CLEVRER: CoLlision Events for Video REpresentation and Reasoning.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning CLEVRER: CoLlision Events for Video REpresentation and Reasoning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.284166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.284166Z digest=sha256:6f20a69c56beae984f9cd567683bdb74e693860487c895487b20052a09c6be49

Observation 57a632c2-0a25-495a-9a35-fead3b46b510 · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.288730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.288730Z digest=sha256:53ff75ab8deefc880d2072dcd3a81335e6a3f4e1a39ae9ad63aa463f3f1855d8

Observation e33a095a-8a57-4fc1-a87a-f4faff51c886 · outbound

This paper cites Robotic Control via Embodied Chain-of-Thought Reasoning.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Robotic Control via Embodied Chain-of-Thought Reasoning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.293505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.293505Z digest=sha256:10257f0dc6a46377f01072b4ae52f42860d65e9ead2184f04ccadd23faa86569

Observation 9f8ee0dc-573e-45ae-9860-6096f69eb10c · outbound

This paper cites Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.298418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.298418Z digest=sha256:fb82fe7c29ed16513008e8744a0902089b7b0a573b8cae91cf8de973ce90e2ad

Observation 2acfc58e-0a2c-4c4c-a655-e68b82535588 · outbound

This paper cites MM-LLMs: Recent Advances in MultiModal Large Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning MM-LLMs: Recent Advances in MultiModal Large Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.303344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.303344Z digest=sha256:18dd75b5bc9d2bbf6befe59a711992b4b06d0426aabf8c0c6385672b72301cf8

Observation 2a7ce8f3-f546-4005-b000-f2b1fbc8a8d1 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.308358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.308358Z digest=sha256:0b33cd65188b1ae1069eddbafcb36bb3c4becc4177cd974530f1ed9158577d7b

Observation 45c250f7-8c85-433c-9e03-4ab1ec7c2059 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T19:27:41.313318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:27:41.313318Z digest=sha256:3351a42d9fa7081c0aaf6718f96d593ca3472e85dd774afd2c980bef79fac3ff

Observation b1cb25b6-0f88-4ddd-a203-aa7685851ea2 · outbound

This paper cites yes” or “no.

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning yes” or “no

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T19:27:41.748201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-10T19:27:41.318182Z digest=sha256:45df3ac0749c11831dc1f1dd2a4986785c758bd01b7597c067f58a1e3c9ab2ae

Pith citing papers

Observation 18877e83-5800-48e1-ae17-c318030e4a98 · inbound

Retrieval-Based Interleaved Visual Chain-of-Thought in Real-World Driving Scenarios cites this paper.

Retrieval-Based Interleaved Visual Chain-of-Thought in Real-World Driving Scenarios SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T21:30:36.398003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:30:36.398003Z digest=sha256:264a56bc125f5779bfd7fdb1ce086d72fc45d5c70e402786be74794b207c81e4

Observation 71090c80-2619-4d3f-9187-59ef4726c803 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.227502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:3fa7b2173f9d858679475b2700a3b1d8d0e0a1dcd7d17b8ffe228d0e82bdba2e

Observation 074227e4-1792-45d2-a42e-e185fa74d942 · inbound

Generative AI Act II: Test Time Scaling Drives Cognition Engineering cites this paper.

Generative AI Act II: Test Time Scaling Drives Cognition Engineering SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 204

Resolution
unresolved
no resolver link, observed 2026-08-16T12:02:43.415017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T12:02:43.415017Z digest=sha256:cd0c48b9293226b65af47d2e9f64e510c516c8723d785c997a3f85c8af81b8b5

Observation 279f78b2-a9b3-4f1e-a622-532cf3036841 · inbound

Scaling and Beyond: Advancing Spatial Reasoning in MLLMs Requires New Recipes cites this paper.

Scaling and Beyond: Advancing Spatial Reasoning in MLLMs Requires New Recipes SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-16T11:37:03.549855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:37:03.549855Z digest=sha256:6630b601b2e3be6aa69f6c984cff28dd4ac49c47c1e83e1b3aac4b7d334a9476

Observation c761a9a6-a5c9-4e1c-81c7-8091cf7a59b2 · inbound

Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation cites this paper.

Perspective-Aware Reasoning in Vision-Language Models via Mental Imagery Simulation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T10:50:43.926669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:50:43.926669Z digest=sha256:fa52cd0b92c80fac3e1832624c28a65b43b715be015c34679871b3824df0d321

Observation 9710cac0-0ab5-4b56-982d-4fd0edfbc6aa · inbound

Nature's Insight: A Novel Framework and Comprehensive Analysis of Agentic Reasoning Through the Lens of Neuroscience cites this paper.

Nature's Insight: A Novel Framework and Comprehensive Analysis of Agentic Reasoning Through the Lens of Neuroscience SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-15T23:31:12.340805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:31:12.340805Z digest=sha256:97e839ef0e70aae71bb30a8da44333437510f45d836fe65a29fa714ab1375cc7

Observation b022d3f3-a60b-435e-99dc-d710ad8ea032 · inbound

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models cites this paper.

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:27.158065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:22:27.158065Z digest=sha256:d02a601af144aa8cd242e462140b26a4ccb01d6f8a17ef5c0ed37f57879d2239

Observation 6b9a3432-c1f9-4b28-913c-24ea6a71de4d · inbound

AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning cites this paper.

AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:14.640225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:56:14.640225Z digest=sha256:93ae409cecd20c556c4d85c1d79ab427f8d067740ba82337e6abc035aecfff03

Observation 474139ab-99b5-46cc-801e-99b140d37b0b · inbound

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning cites this paper.

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:48.113277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:48.113277Z digest=sha256:b09d143a5ad52db6267cb71427e648cccdf5ecbbfe61f1adc1f7d945fcb2044c

Observation fb2d0e45-7ae2-4b1d-b81b-91325487b923 · inbound

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline cites this paper.

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T23:56:04.037430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:56:04.037430Z digest=sha256:894d498cd0bbc75c5b57c8452674198ca20ce5accb758eb8bf2f42863fb94f8d

Observation d7e59e26-d885-4a83-9c37-ab0c44ded2ff · inbound

PySeizure: A single machine learning classifier framework to detect seizures in diverse datasets cites this paper.

PySeizure: A single machine learning classifier framework to detect seizures in diverse datasets SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:05.094176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:05.094176Z digest=sha256:f304f138c794b5561d3ed42fe9828b2e89734c535d2a7b7cccb458d15e37fd45

Observation 7ee2fe4f-6421-460a-b40b-9b26b89109b7 · inbound

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture cites this paper.

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:23.101934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:40:23.101934Z digest=sha256:7adfa0242efe0d2b50f33270413fb0f75bbc933348eb6ea7ea2f1943261ed44e

Observation 5dad4a7b-f9b2-4a70-b6c6-d4b53cb38a56 · inbound

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models cites this paper.

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:10:48.842964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-18T03:09:09.713822Z digest=sha256:5585445db3ba306b0b803b2fd22dc158ca04deb9552560b6d8d8a02add6e8f94

Observation 634d84b9-9dcc-413a-b90b-19e7c5a5b5b3 · inbound

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL cites this paper.

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T18:42:12.111148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:42:12.111148Z digest=sha256:b9fbe138e34c18307498a1f2c9d090e142a8322184ce0864d67a95479fa31be3

Observation b30b3fe8-a60c-4285-afbf-d40364a3d9eb · inbound

SCP: Spatial Causal Prediction in Video cites this paper.

SCP: Spatial Causal Prediction in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:11.146675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T16:47:44.523606Z digest=sha256:4aeb6c2d2930deb46e6b72777f4c00670ff8d9f3156768b408ee68cb956d714d

Observation 4de075f5-7088-49b8-9dab-da2a933f0d85 · inbound

Token Warping Helps MLLMs Look from Nearby Viewpoints cites this paper.

Token Warping Helps MLLMs Look from Nearby Viewpoints SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:08:17.361639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T21:07:55.062113Z digest=sha256:0f1aac6eebb1a1c73d5484e6cc6840515c5bec125b2351f82fea5f9eb0582108

Observation 5779c4ba-d08c-4616-b328-18e48d2f23df · inbound

Spatio-Temporal Grounding of Large Language Models from Perception Streams cites this paper.

Spatio-Temporal Grounding of Large Language Models from Perception Streams SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:02.684607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T17:10:45.837684Z digest=sha256:d676552d68c434ee8943a7d12c62793a96eb0f9d76a1a7f6a734135f8fea651f

Observation b6d757a7-7c85-4524-a156-a2e9aadef4ea · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:11.747234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-08T14:20:08.404090Z digest=sha256:0c651f73bd0b46abd4b134025cb0d3eaf06e3249ca1d0eaeb22cac3523f702b2

Observation 735994be-fcfe-4f8e-b488-ed43f427ee0b · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:16:39.949775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T06:15:33.062980Z digest=sha256:656cc5a91f04ff847e1f2addc6275d51fbba138f24e7cb01597faf63bc39aa53

Observation 7d5b6bdc-d4c0-4847-b0dc-db113d0d9d6d · inbound

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment cites this paper.

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:59.892965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:55:39.721560Z digest=sha256:9fbffdd9905412189b854fc1ce190d78765f9694134a0a2f3800c4d0258e6f79

Observation 6f9e0018-747c-4b0d-b457-958542a9d008 · inbound

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images cites this paper.

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:01.622856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-13T01:26:47.052025Z digest=sha256:8bedc83a3206021ffcf160180b4d60623096a581f30d6bc48ae553b6092d009a

Observation e69a6f8e-d685-4842-8d40-009b644c3d13 · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:53:13.184532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T10:52:22.778489Z digest=sha256:7567837b4e65f0fc809f8e420d22207b68a40650c11bebef1b48e31989915acc

Observation 89df9b66-4673-47fa-82f8-04cce5bdfade · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:05:47.169697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T18:25:17.831116Z digest=sha256:7d2005a4180a1fd2dee58e99e510d63ebcf6ccb2e8a4dcd64734ad896bfada56

Observation 30082d11-5f43-49ce-a533-2931f22751ee · inbound

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches cites this paper.

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:03:58.061704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T05:00:11.848526Z digest=sha256:3f5d59c570d4784775309f9cbd8fb8244089d5005699047b74f8d1598d5f1b53

Observation f0d63128-776f-4dc8-92fb-3379e8744763 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:15:22.732097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T05:10:32.522453Z digest=sha256:2c514af2182db6cd8e455949d02fa358507ffc71953a2aadeaa1d34f73334c96

Observation c178bf64-6215-4858-8932-b616a2f27228 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:44:56.062222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-30T16:40:22.441025Z digest=sha256:fd298cdfa3b678abaec9c760ba86981c9a30ecc046cdbd289dd18819610b1207

Observation f2fb586f-e6fd-49cc-a50b-6a6d5ebf95c2 · inbound

Grounded 3D-Aware Spatial Vision-Language Modeling cites this paper.

Grounded 3D-Aware Spatial Vision-Language Modeling SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.342690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T08:08:36.012761Z digest=sha256:ebbff673acf338c8ea3b8701b35e4c1a5baecca6686d4dfef501206244b73f08

Observation 8cc528be-b179-4200-b688-23ea88a620e4 · inbound

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks cites this paper.

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.607933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T16:35:14.099586Z digest=sha256:bc563587782fa0bb4ddc50bb0cb39d7fed10866d5d3ed755f50bf707cb53927a

Observation 32232175-98d9-4877-ba72-d02745cd166f · inbound

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models cites this paper.

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:55.601454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-27T01:42:30.005911Z digest=sha256:cc43b3f4c00b311edda2e47e2958fc216d33c10ecfe3b589f377a6bb95849ace

Observation 4344ecef-9e55-4a5e-9e19-5f42073828f2 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.479409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:2f5557c502699f3360c3770196b970b66b3e13b27886da8ad468c25af159598a

Observation f481d8e6-21b1-4613-9024-233ef425327c · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.694996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:596ea3bee1b108cabdcf65436d1c7ccc23ee6d912698a115bc0831670f136dc8

Observation 5317129b-a50e-4e53-882f-79221021b8ed · inbound

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping cites this paper.

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-02T14:17:02.420247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-02T14:16:39.649823Z digest=sha256:99ac9e25e7bb527e77346017fcdc917279e66de11553f0e116dc98f8fe821947

Observation 490f06ad-ce8e-40d3-a95f-7ae53961e6b1 · inbound

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video cites this paper.

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:18:37.311510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-03T16:16:41.412451Z digest=sha256:5a08d5945d6a42d599336f3a136151f99b9b17360c92dfe352ffcb6ba2211041

Observation ddc2f67e-6117-4bc4-8b82-5fb77958e9ed · inbound

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? cites this paper.

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T09:09:24.863958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:09:24.863958Z digest=sha256:e21319c8f8a9ffa0256628b771259e8637bd7582bf6a0bc6f6f75a910d895f1f