Pith. sign in

Paper Citation Record · LEDGER

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

As of 6 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 14 inbound Pith citation observations for arXiv:2511.07403.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.07403 v2

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T23:08:57.431844Z

measured 75 of 75 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T05:12:59.576553Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-05T01:50:34.827326Z

Reference resolution

61 of 61 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved61
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c89204c0-a66f-433b-9146-90382ced24ee · outbound

This paper cites SpatialBot: Precise Spatial Understanding with Vision Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards SpatialBot: Precise Spatial Understanding with Vision Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.001243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.001243Z digest=sha256:1d8957a5ee9264ef99458b6e0691b81adf4db36855ee5024bc202a63751c1421

Observation 78df0953-2684-4d92-a73a-77feec1f572f · outbound

This paper cites The final set consists of 50% samples from the relation category, and the remaining 50% distributed across the eight other categories.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards The final set consists of 50% samples from the relation category, and the remaining 50% distributed across the eight other categories

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.351341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.351341Z digest=sha256:7f678c05892c49d77e19ce9410c82d5144952ec9608a0c884b404edb60e5ac24

Observation 754c7e6f-e5aa-4259-9b5c-544b0f9bf65d · outbound

This paper cites MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.211667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.211667Z digest=sha256:6470be8e9fcc8d1b3868d6f44d235d5957ed3e562e8f4fe8e31aa9edf75c15f6

Observation b75dce60-90e9-4b68-b252-93ff18ef8dd0 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.319147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.319147Z digest=sha256:ee37832877bb42916d6962d023dea5c4c14d52129232770c32424a522a063619

Observation f4106c95-603e-4526-b91a-3109f8fdc443 · outbound

This paper cites Kimi-VL Technical Report.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Kimi-VL Technical Report

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.425505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.425505Z digest=sha256:2302d9580203c4132fef1db541949e57f81ac88da996db48b547f794e9c55b88

Observation 80da89b1-98fb-43e9-8400-1d37f0472ea6 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.532620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.532620Z digest=sha256:dd856eba975e368ed7d11d917165f4e0e63523ca6b9b9221c4d43d1faef71b3e

Observation 691282a9-ed0d-4b96-9c9b-afde5300dec0 · outbound

This paper cites Xia, Ted Xiao, Jiajun Wu, Brian Ichter, Anirudha Majumdar, and Dorsa Sadigh.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Xia, Ted Xiao, Jiajun Wu, Brian Ichter, Anirudha Majumdar, and Dorsa Sadigh

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.752446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.752446Z digest=sha256:41c109fad12b86fb27584b6104833473db302606a8fb05207387dba5fa19a87d

Observation 3ec25f26-b34a-410d-8b88-9f51e742458a · outbound

This paper cites Tenenbaum, Antonio Torralba, Florian Shkurti, and Liam Paull.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Tenenbaum, Antonio Torralba, Florian Shkurti, and Liam Paull

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.905731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.905731Z digest=sha256:dd798f9c42c5ea6620e43ebe331c2d0553a441caf590a956837a8752d50d7034

Observation f00d6395-04a7-43f8-a724-3c149d321124 · outbound

This paper cites an unresolved cited work.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.006008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.006008Z digest=sha256:9ecbf169b9c9a95d1369f1ccf01ace4bdbde6c9ccf7cb52dd0826e95809d7e2c

Observation 37e990fd-39a4-40b1-a9fa-b3e215801bea · outbound

This paper cites Scene Graph Reasoning for Visual Question Answering.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Scene Graph Reasoning for Visual Question Answering

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.161728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.161728Z digest=sha256:7ca291fcda2afd460a64a27b4111f880fbd64ea1ff3f220365a096dc3f9bd5af

Observation 684ffb9a-8e32-4797-9e20-825ddc225477 · outbound

This paper cites Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Correcting the Mythos of KL-Regularization: Direct Alignment without Overoptimization via Chi-Squared Preference Optimization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.379903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.379903Z digest=sha256:d59c0e019aa82d005aafb9fddbc3881bb7c33ec5f9801ba3d4dbf8c9559c4acd

Observation e5ed267b-4d82-460f-8344-e74390229a56 · outbound

This paper cites Visual language maps for robot navigation.2023 IEEE International Conference on Robotics and Automation (ICRA), pp.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Visual language maps for robot navigation.2023 IEEE International Conference on Robotics and Automation (ICRA), pp

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.569283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.569283Z digest=sha256:476ea44161c86177e1e7c82d739883d08f34cd472b436c117ea04bee6e930fae

Observation 12199c74-cd1d-47a9-8c57-0bec1d2b0098 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.090448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.090448Z digest=sha256:3a6f62eb3b877ae9dbe7b97b3e1fc67f9349e1aa975f5e415eb5de52d66153c3

Observation a74d0c71-6a7e-4dac-b709-2950e1d0698a · outbound

This paper cites What's "up" with vision-language models? Investigating their struggle with spatial reasoning.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards What's "up" with vision-language models? Investigating their struggle with spatial reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.252391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.252391Z digest=sha256:347e8ca231a5d61c1f0fd316872edb36d6a044c22934c2a21e18919de3c932a2

Observation 3b651970-7826-49d1-aa5a-51902c38b2e3 · outbound

This paper cites VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards VR-GPT: Visual Language Model for Intelligent Virtual Reality Applications

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.434311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.434311Z digest=sha256:ac62a5970fdf050b083a2f4d8962dca76120ed6866e2d4bfd87e9d856bc47c17

Observation fc975e27-dd5f-468e-ae0f-9bf2ba40cd60 · outbound

This paper cites Relation-r1: Progressively cognitive chain-of-thought guided reinforcement learning for unified relation comprehension.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Relation-r1: Progressively cognitive chain-of-thought guided reinforcement learning for unified relation comprehension

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.695890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.695890Z digest=sha256:4e05f5465c8efa730fb51266328cf0012357fbfdd8b082329871384b7a92f2c4

Observation 83870d2b-27e4-4637-b5e3-c217f810ec08 · outbound

This paper cites Visual Instruction Tuning.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Visual Instruction Tuning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.834140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.834140Z digest=sha256:3006d381d9a16b356b33def56dae4e299921a474e62a19ef030102c59c7eae58

Observation 0747c495-7def-4a55-b911-31dd296e47ba · outbound

This paper cites Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.041071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.041071Z digest=sha256:994a09ede8be2feef4058994b3f7ba11b91b054f8c2307e31179f4824fb562f2

Observation c891422c-6e37-45f5-bffc-9d02a50b97e0 · outbound

This paper cites SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards SpartQA: : A Textual Question Answering Benchmark for Spatial Reasoning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.208502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.208502Z digest=sha256:32497fb65f1499b3c587d9f3de72e790a3e210c62b0fe995147a753f7f0a99c9

Observation a06ed670-65f6-48a4-991d-add349d902c4 · outbound

This paper cites SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic Data.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards SpaRE: Enhancing Spatial Reasoning in Vision-Language Models with Synthetic Data

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.541313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.541313Z digest=sha256:a4029f83be1e2163d5803ede62302b3f89b63530dd28257e6e9b149cb84d94b1

Observation 2e759782-f9f9-4f57-9c33-1ef2fafb3c75 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.730617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.730617Z digest=sha256:d9f6fbc33fc0b4e552922a7c5765742697de8dd4fa32b32f0c272fd345fcc3ba

Observation f39ec1f3-04ff-4a4c-9c30-c8683155be63 · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.900128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.900128Z digest=sha256:f2ab600a089edd9c050affdba596a0d062e7d64b6b96e220aaecb67a08aa1460

Observation e7353cbc-a74a-4f24-9d51-4d6adfe96c72 · outbound

This paper cites Tuning computer vision models with task rewards.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Tuning computer vision models with task rewards

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.063278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.063278Z digest=sha256:73ab54abb35f9355e510d75e4efc1d2a2770d39c7e1335970a0377e85719ce73

Observation 0b5eb21f-d1d3-4ea5-b9d0-0162255ccabd · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.181240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.181240Z digest=sha256:7c2eb174a9eb170885ed46a34d9f262572459eedf0754caf4c924a9b629875d6

Observation 45e8f9db-be2d-4c7d-aa71-c9aac05fda5d · outbound

This paper cites Satori-r1: Incentivizing multimodal reasoning with spatial grounding and verifiable rewards.ArXiv, abs/2505.19094, 2025a.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Satori-r1: Incentivizing multimodal reasoning with spatial grounding and verifiable rewards.ArXiv, abs/2505.19094, 2025a

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.286554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.286554Z digest=sha256:e89613f3ce1919537199ed6aa95eded7dc811beb1d7ea05c325aff9f6d259ef9

Observation 1b7a79e8-c0f5-47b5-9ddd-417b685a1554 · outbound

This paper cites LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.435983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.435983Z digest=sha256:f0b3da564787cc88c208e6d6876dd3f8e07404b985006be4427e406e90139d99

Observation 18eb29e9-b2f6-43ac-bd47-c68f2c47f2f7 · outbound

This paper cites Gemini Robotics: Bringing AI into the Physical World.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Gemini Robotics: Bringing AI into the Physical World

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.597396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.597396Z digest=sha256:5a3a7684151b31d7c9d476864db9e3aaf8942a4252355d05d07641b320d91374

Observation 3e2c0cf1-8d70-4105-b250-082a05566d2f · outbound

This paper cites Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.784146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.784146Z digest=sha256:43753fcff2fdd580b61bb13c9faa490846cbb873386fdfc52f100ecc5e467c14

Observation 184ba021-4085-4680-adac-263da207cb1b · outbound

This paper cites Learning 3d semantic scene graphs from 3d indoor reconstructions.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Learning 3d semantic scene graphs from 3d indoor reconstructions.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:53.880851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:53.880851Z digest=sha256:3a7541029d7ae57136a0fa49287ad430c85295b912b41567985e9963a3d951a0

Observation 49f1a966-d0d5-4fa1-b6ae-f49e5e29be33 · outbound

This paper cites SVQA-R1: Reinforcing Spatial Reasoning in MLLMs via View-Consistent Reward Optimization.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards SVQA-R1: Reinforcing Spatial Reasoning in MLLMs via View-Consistent Reward Optimization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.014934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.014934Z digest=sha256:4171be6f0cc0bac70766370de280cd2a75bb9ae06bc208e3726977a30e5749f8

Observation 8bcb050c-f830-470d-bf5e-a9fb35ed203d · outbound

This paper cites Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Compositional 4D Dynamic Scenes Understanding with Physics Priors for Video Question Answering

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.150077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.150077Z digest=sha256:3bdcf770b731d9ffa0a2a676122ac0a83bdf25fa8a8fddcfda7a9b2beaa65167

Observation 60ab87b4-c715-4fd4-a794-66a83e21b3ac · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.313086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.313086Z digest=sha256:8b5b724b44817c34eb998dfa93143d9a17b7374410b7e83428414feb1ed1424d

Observation c3ddaedc-c736-4ea3-8d8d-1ca4bb035de5 · outbound

This paper cites V*: Guided visual search as a core mechanism in multimodal llms.2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards V*: Guided visual search as a core mechanism in multimodal llms.2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.436266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.436266Z digest=sha256:8da5c33826d537625efcdefaac4f1a75a9d1b508060504505094e7a9669cb157

Observation eb66f52d-fd70-4f0a-8f51-9b80d47b2c4f · outbound

This paper cites Zang, Peng Gao, Yixuan Li, and Kaiyang Zhou.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Zang, Peng Gao, Yixuan Li, and Kaiyang Zhou

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.579374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.579374Z digest=sha256:0c02da18f8fadf3ed0b94a37a54a6dbdaf8a5329ae29afa8b9000e60a911d0c5

Observation c1e0177b-f4a2-4bee-8e91-f8ccc98bf475 · outbound

This paper cites Advancing multimodal reasoning capabilities of multimodal large language models via visual perception reward.ArXiv, abs/2506.07218,.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Advancing multimodal reasoning capabilities of multimodal large language models via visual perception reward.ArXiv, abs/2506.07218,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.699568Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.699568Z digest=sha256:9d41fdac190aeb39469f13e10a47411cec23e1b87dcff3f5629bbced6c468a05

Observation 2da96a50-0811-4637-a97e-afca4a1a2534 · outbound

This paper cites Evaluating Spatial Understanding of Large Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Evaluating Spatial Understanding of Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.820458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.820458Z digest=sha256:6d63f5a9ce3c0c75ac5ecd48b2f05f3d5ec30ed4574b5171d91c8fb168e9fb78

Observation 097102b2-558c-4e44-b174-d754c5e13c78 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.096964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.096964Z digest=sha256:f064af63438be5f999f7e4dbb02282383af646f1eca693064e81ad19408460ea

Observation 8e378c4c-a10a-4695-a5af-ca05f8850fde · outbound

This paper cites RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.221065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.221065Z digest=sha256:166264d5500fdd3220d8bba5449974f8e98205bf6c855ac7a35ac3ee1fca9699

Observation 872c63a0-cae0-4106-8105-33e54e94d821 · outbound

This paper cites MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.368717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.368717Z digest=sha256:bc5d643db49f3c216dcfb370754329457b3766e9a77663422804faf3f1074b66

Observation 741f7457-e977-48e8-906b-26677e10b43a · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.516798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.516798Z digest=sha256:c9622fd3b40011171dab6d2b16a5646b41b4026b8b13fada11eb2b8ef3926572

Observation 30ad8c8b-2656-4860-81cc-30be0e6f4a53 · outbound

This paper cites The doubly librating Plutinos.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards The doubly librating Plutinos

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.631675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.631675Z digest=sha256:d2518c5689d607d231ef975876affb5c3d106ff4f6cb91b800f2258c5f76017a

Observation 5d1af005-a726-46b2-94d0-6c2d3323fdb6 · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.751893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.751893Z digest=sha256:84e204d3999a80f35289424ccc44e10d6938dd513049686ff42836863110d15b

Observation 30a3f40e-dfc0-4bbe-9685-9adb8bd9b6bf · outbound

This paper cites Struct2d: A perception-guided framework for spatial reasoning in large multimodal models.ArXiv, abs/2506.04220,.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Struct2d: A perception-guided framework for spatial reasoning in large multimodal models.ArXiv, abs/2506.04220,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:55.911489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:55.911489Z digest=sha256:fcf685f792056ffb679cde0c3d3fde60986bccd0902396700a25cf4ead7e0192

Observation 1a89eb08-1985-4207-a103-1cbe9ea12b18 · outbound

This paper cites an unresolved cited work.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Unresolved cited work

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.091197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.091197Z digest=sha256:801f98c23ced4f4109a2396d4369463f6809a37d6cc9a189a0bcea9d0dd57780

Observation eb9456ff-e1b5-41a9-948e-44ef2f8846f9 · outbound

This paper cites an unresolved cited work.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.235392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.235392Z digest=sha256:2935ea8219efa91493f48375396822642e28e50743a5779f6c040c39a7755c78

Observation 5831bdf6-d86d-49f3-89c5-382b79c9af35 · outbound

This paper cites These serve as upper bounds for spatial generalization under non-public training regimes.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards These serve as upper bounds for spatial generalization under non-public training regimes

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.662635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.662635Z digest=sha256:f6b427746cdfbcd26d1bfb7a5bd22fe127f4f017447ffc1fe11d7afe24c56a34

Observation 01525c57-1a31-4dfa-85e1-b4bd8224ac02 · outbound

This paper cites an unresolved cited work.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Unresolved cited work

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.802456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.802456Z digest=sha256:7ada7657667eae17c6b42b0fe4992a1739f005ec3cf8f4606cb11a4bff9a93fa

Observation f535576d-f97a-44fe-b75b-e468cb08611c · outbound

This paper cites an unresolved cited work.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:57.144009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:57.144009Z digest=sha256:9297cbbcbb37fdc73522c61fa5aab1f325a05028020e7914f193a3a4c1aecff4

Observation 004f16b0-dabe-41d0-a4de-48ae173b8d56 · outbound

This paper cites aha moment.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards aha moment

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:57.290068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:57.290068Z digest=sha256:1bd73cdf4e075120879b02deb2219ff78f822845133a87201d4e55d585d1dc6f

Observation 392b51f9-2491-4a61-94b4-539e7c057893 · outbound

This paper cites In contrast, Huang et al.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards In contrast, Huang et al

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:57.431844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:57.431844Z digest=sha256:8c3dd0da4bcc1cb58ffe924f879d3c3ffbed464fff11960ad01b60dbf283800f

Observation 3cb75062-c408-4e15-b1be-f34b8d972534 · outbound

This paper cites Training time totals around 13 hours for the 3B model and 15 hours for the 7B model.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Training time totals around 13 hours for the 3B model and 15 hours for the 7B model

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.468585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.468585Z digest=sha256:0ff333c3faafa59983f055887363d63bf6b756484387ed1259029c9578204d4f

Observation e4a28644-7eea-4469-ab5b-f78d4e3d5ea8 · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:54.932557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:54.932557Z digest=sha256:afd9e413da27ad445d9740932d48b3662bfa6a2554ba5651ce809b6ac3693bab

Observation 6be8b81c-1b5a-44a9-849b-a13e185874eb · outbound

This paper cites TopViewRS: Vision-Language Models as Top-View Spatial Reasoners.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards TopViewRS: Vision-Language Models as Top-View Spatial Reasoners

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:51.560457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:51.560457Z digest=sha256:71d3784493eac6d30a5b0a87257e0ad28d5e0b2066302bdbce50cfdb40bf0d4b

Observation bc46fea0-5a8f-4cd9-8c99-16b686b37315 · outbound

This paper cites GPT-4o System Card.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards GPT-4o System Card

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.963202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.963202Z digest=sha256:df0835a17fcd1c00fd685b47914536802f97b8f655bfa4c7ae34629fd5e89ada

Observation a184775b-87aa-4aad-b5e2-e31d00093f42 · outbound

This paper cites Scene Graph Generation with Role-Playing Large Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Scene Graph Generation with Role-Playing Large Language Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.079544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.079544Z digest=sha256:d30f88051284bbeffd8ef69b1294e83c6202dee1e215d469d275ed2bfdd0eb62

Observation 2ace252f-c4e0-4843-b1d3-24b59dcc3d43 · outbound

This paper cites PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards PIVOT: Iterative Visual Prompting Elicits Actionable Knowledge for VLMs

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:52.414897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:52.414897Z digest=sha256:ee551a5aa42a0e4b9fc833fb420bcf4aeb479e34ab4cd3738079328d89663db1

Observation 850b4d05-8185-4f29-8c22-664b61f7907f · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:50.746629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:50.746629Z digest=sha256:822f8b5af34aa7ff6093eb1eded899d01c6049e8d9ea50b72de1133e9a433494

Observation 5e1f95ee-15d3-4f39-ab9a-a06880eb7999 · outbound

This paper cites Compile Scene Graphs with Reinforcement Learning.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Compile Scene Graphs with Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.130584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.130584Z digest=sha256:01bca173f387d33aa14d205d42dbacf9110f7b7dfbd9bb360e9743fd7533bff7

Observation 984968e4-4ccf-46c3-bc48-db3fe18d1771 · outbound

This paper cites Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Cognitive Behaviors that Enable Self-Improving Reasoners, or, Four Habits of Highly Effective STaRs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:49.679221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:49.679221Z digest=sha256:d7af93ce39d05863494a17f062d767e8883960d444d04de150d7ad8b44dc5d1f

Observation bba40aad-bee4-4534-b73c-9358d4f9d998 · outbound

This paper cites Qwen2.5-VL Technical Report.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:48.929476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:48.929476Z digest=sha256:1409c89be3e0c7bd61dc0361c08551b4ba4c8e0a6266e500031112acee64e3b5

Observation 3412f57a-a493-4f14-a5ce-f65aa372c0f5 · outbound

This paper cites For models with spe- cific reasoning templates such as VLAA-Thinker, SpaceThinker, and SpaceOm, we utilize their corresponding structured prompts.

SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards For models with spe- cific reasoning templates such as VLAA-Thinker, SpaceThinker, and SpaceOm, we utilize their corresponding structured prompts

Reference 2048

Resolution
unresolved
no resolver link, observed 2026-08-03T23:08:56.979052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:08:56.979052Z digest=sha256:a52bf13385146dcf80144a48407b5b81c30a448b6a9c1cdd37f31a02c1f4102b

Pith citing papers

Observation 6d737bce-47ec-4e2c-8b49-03e8735a432b · inbound

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence cites this paper.

SpatialScore: Towards Comprehensive Evaluation for Spatial Intelligence SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-22T13:07:11.548885Z digest=sha256:44a6f63c4f5defa69d591a859cb6321a5fe302124746f3c6b7267694e0d0ca29

Observation 34996bd8-253e-4ffe-b57e-554946a45837 · inbound

VIEW2SPACE: Studying Multi-View Visual Reasoning from Sparse Observations cites this paper.

VIEW2SPACE: Studying Multi-View Visual Reasoning from Sparse Observations SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T23:42:44.515158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:42:44.515158Z digest=sha256:e91298f498c8939994ddf651cdeb02ce60e284aa2e7dfef7353e0330839305b1

Observation fc44c71c-409f-4383-b157-2cb6a8ac337c · inbound

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning cites this paper.

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-09T22:52:22.203519Z digest=sha256:24b4d1a8d92991b3fe325ae38dd783392ebf5e429c5ed37f30e1e2df4efd71b6

Observation c9708611-744c-4878-a6f2-9b9a02fc20ed · inbound

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning cites this paper.

SpatiO: Adaptive Test-Time Orchestration of Vision-Language Agents for Spatial Reasoning SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-05T01:49:16.845025Z digest=sha256:1ad18b109bec328321f08e1593e7921d4efcf42c92eecdf6c9873ee6a395fe2c

Observation ca4585f3-9bbd-455f-99f5-b88ee5681fbf · inbound

Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models cites this paper.

Uncovering and Shaping the Latent Representation of 3D Scene Topology in Vision-Language Models SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T00:56:06.243890Z digest=sha256:0b562810cf1cd4050c625129d2d6406195976ab853063559d1dce5c10b0e2ab2

Observation 7b3f3a60-4cf2-4ab3-a2cc-181b3a92d60b · inbound

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models cites this paper.

ViSRA: A Video-based Spatial Reasoning Agent for Multi-modal Large Language Models SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T04:00:23.681682Z digest=sha256:60be07f1c2ad96e6e539e116a182aa4049a5297e8965d6b44f32c8922c125eca

Observation b770fcad-f881-40ee-942e-4d520cea9eb5 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-25T05:10:32.522453Z digest=sha256:12001432ed0fd7a3b7d62c923caf19755a0c4b5190ec1bd9befe2e19402636a1

Observation 38dfe81d-c142-4e66-b0af-0bf30c46bbe1 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T16:40:22.441025Z digest=sha256:9f2181f8be9c1deb322688654aa4f74900f962f84d341c74d4ba7ba49ab6d5aa

Observation 0bb8841c-5d38-4287-be56-530b505ae253 · inbound

Rethinking VLM Representation for VLA Initialization cites this paper.

Rethinking VLM Representation for VLA Initialization SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T22:21:29.733181Z digest=sha256:b4cb8dd88f470fa8b306ca3e4526783643f988add92d947162e30e9ec75d9077

Observation c29adb50-0f1a-4caa-b011-fe7216da0b91 · inbound

OneCanvas: 3D Scene Understanding via Panoramic Reprojection cites this paper.

OneCanvas: 3D Scene Understanding via Panoramic Reprojection SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T21:38:15.988253Z digest=sha256:d5c3b6233abad45d1028355a1bcbfd2108ecc4b19e98b35bd32c29db546c0624

Observation 11d137df-2567-4086-9f64-d5816779820b · inbound

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping cites this paper.

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-07-07T01:16:00.407423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-02T14:16:39.649823Z digest=sha256:a4a23f4327cb4645a540ef322fe3965f21de5419aafcec9748dcf49ff4e7de8a

Observation f5472762-4fe3-4c0a-b2ec-3df3a4fd17b6 · inbound

GReFEM: Multimodal LLMs as Zero-Shot Semantic Assistants for Physics-Guided 3D Mesh Refinement cites this paper.

GReFEM: Multimodal LLMs as Zero-Shot Semantic Assistants for Physics-Guided 3D Mesh Refinement SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T06:42:45.558324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T06:42:45.558324Z digest=sha256:30e7cf5ae5b027e0fcdb5aa64c396b0e72a2f7ac21dc677b25056150f5caa844

Observation 9f830fba-1192-4c41-a984-bf514729bb13 · inbound

GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding cites this paper.

GeoAnchor: Collaborative Reasoning via Latent Decomposition for 3D Spatial Understanding SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T05:12:59.576553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:12:59.576553Z digest=sha256:5074446ea1b717fb03e39a30d62806006127cfca8423789c67687ba0eaaa8691

Observation 795b0743-653f-439a-b168-6fc537821dbe · inbound

Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models cites this paper.

Geo3R: Mitigating Spatial Reasoning Hallucination in Multimodal Large Language Models SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T08:34:11.633653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T08:34:11.633653Z digest=sha256:da35cebc523ccf7070ab0e3d8920961f5180ee0efc02e77fb5bb85e32ff05e3d