Pith. sign in

Paper Citation Record · LEDGER

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning

As of 13 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2605.27310.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.27310 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T18:19:40.323661Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-30T21:38:08.608340Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact13
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3dd08ca0-8425-41e0-a0e9-95a3899708ac · outbound

This paper cites Qwen3-VL Technical Report.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Qwen3-VL Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.688850Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:16092037d275ae00b5f9d21d864435ab135a6ddf55764481928667f7dc5f1c4a

Observation 348a80d0-11ff-4c90-aff6-3db1a4d152f6 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:3eb675d4c66ef615644a9340a380a9652bf7fecd0924d37d6dcc378d138975fa

Observation 7f08ef09-1647-4115-9d24-6dbfe995fd19 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:7b9ce9d9a6a04e6fc15b6bc19ce17d3ab1c0476d5464ddf4700a07766554d1a7

Observation 242bab3d-42aa-49d0-849a-45fdb372b92d · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.691296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:b3ca33817cbb17744899e7efd38182b48ca45fdbcc689eee7c27dbf1e93818c4

Observation 2eedf577-e12c-48af-ae79-3b8247e1e6a7 · outbound

This paper cites Think with 3d: Geometric imagination grounded spatial reasoning from limited views.arXiv preprint arXiv:2510.18632, 2025a.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Think with 3d: Geometric imagination grounded spatial reasoning from limited views.arXiv preprint arXiv:2510.18632, 2025a

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.667321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:6cd8dfd585940f50de5c2a62434bf9097a3d39f83d17d951ad77d7ccc727a8e1

Observation 8c5b67d6-ac9e-4c79-b641-4a7fc60a7294 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:50304a392db2ef939e9a1985cdf07329dcec8a8956fd812c6bce0f54efcff258

Observation 0ae30ed4-f04d-4cfb-b719-a72c90bb89a3 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Emerging Properties in Unified Multimodal Pretraining

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.661816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:dcfb7c97fa89051cbd36b09b5be68e5bd1af9a95fcfd7a258f6d29bc316e919e

Observation 78146988-1f12-4e74-b914-132dd2518d34 · outbound

This paper cites SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.686391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:6938abe27f1517339f91c237d707b05fa030c9d8dd88bf6b7b472ec33d4a63ba

Observation a4b2b6e6-7c3f-442c-9d88-65bac298cf8c · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:e0bc4a64ef172c5bd6bb1ca6a287e299707b05f79f4b8358e8c54907b6e753ef

Observation 69a50e7a-a27a-426b-8448-648e11a19133 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:b5ddb545f08ba996e6b46959f15d33b175119b5a59aefda55ee20fc390231354

Observation e45465a3-ecc2-437c-b3b8-39cfdc143f7d · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:9c9bc239551ee7247efb45784b28a14169f0d4a3cca35be5a3440851c0603c67

Observation fd83d925-ee9d-41af-9ded-f73f6b3a706b · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:7733afdb9c38559f8654488118d159549519c90f2dec43c609cbcc5424ca291f

Observation eabba675-f5fe-4290-97aa-aa6886716286 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:cf36fe3860b261d16bb3f4acbf968884df4b0f858cf0de172488ae0036e9e50c

Observation a8737088-92dd-4d1e-bd3d-f4dc09fc2394 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:e31b145c0a3e62129464592854722a8feb9bd0d3f09d4d479925c708fa846d7d

Observation 4bfbd459-9464-431b-9e35-5019b30701de · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:736ed8e126276465e74efd1594817255208aec784391812dc69a0ea4a23d0025

Observation 44fdedd9-251c-4262-a198-b6f0b9458c61 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:00268f1bc7e668059f93dac6e0a524f31d18d83f796d3ab3d46e42c387ea31d6

Observation 8ba89a69-4dea-4536-a222-9954741a6f6c · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:53c4574fd1ca4a8a77d515f39c46db16e16e59610098a6d137731b53e5647861

Observation aaae50c0-09d6-4294-91d0-c03f773b6f88 · outbound

This paper cites Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.680893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:83df36665d77a24d11be41d9ee9cda61ac121778faafa45d38ab29eae3cca287

Observation b19625ec-1ce9-4422-b2ba-bc1869d8ca6f · outbound

This paper cites Tuna: Taming unified visual representations for native unified multimodal models.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Tuna: Taming unified visual representations for native unified multimodal models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.672677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:846aa85f9b8a9d26a13fd82addc9a6be01ad46c124476495d3f35921a6fc0e36

Observation de90a6bd-cdd4-46d2-9451-94e40f4ab370 · outbound

This paper cites On the faithfulness of visual thinking: Measurement and enhancement.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning On the faithfulness of visual thinking: Measurement and enhancement

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.675174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:cf2a66376444bebc5267f91abc7fa2d24a00ad375631d664c31adafae06895ea

Observation d4477653-4d6a-4fd8-9643-6576f37facb9 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:9c51ff9656a3326f69540e524eaed3eea595b54d156ed99e679d0a8d6c37214f

Observation d320b631-8431-490b-bb25-79c4b7d283dc · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.683511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:9ccb13bf0c1f12bbcada058046f664c3888dddafbad251c7419eba9a88cb8483

Observation 8de9581b-c8bb-4c7b-a79a-4d47e2a7520e · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:58f980b4d95cd6ea51bc2bb970f85f18500b767f10bb38eb00c8eddbf84fe132

Observation 45797a99-c0b8-4573-8d41-c7e65b249fc0 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:107bfc6bf94c8b5b75c52d78fc25382ad6c0e286ce0e77f1c4ebd92f6f77231a

Observation 549e4634-a0c4-4941-8f20-a0b7a8a22075 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:d3b81343b4f0ca72e7cd09e62ff0ae149347b08db41dc4fa6161061d31a2acec

Observation fecc4a5a-f9d3-4d1d-a18d-3c133eaf00b4 · outbound

This paper cites Towards cross-view point correspondence in vision- language models.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Towards cross-view point correspondence in vision- language models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.669993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:cc60b8d3870315432acf7f46fb8e4c0c13422f8c15cd44ad57065c7f45155de0

Observation 76ececa2-e4ed-44aa-be9f-33c49030e597 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:b40494c3781dc30b7d2b0b2f028a976487a49b433e8d39856e5052582ce34fd2

Observation 5ffac8d2-4516-4e8e-b5a6-9562c2514e9f · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:6da0d956daa87535420aa1f437422b4a9d0babde9ca66f29311a195b9bf1eac8

Observation 179039e1-a077-4cc9-860f-8e96cfea2ede · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:73f213cd52761ef8e58f01467da6f3f6bde40bec966ccb6d0d2585024413ace7

Observation 5bb566fc-f590-4140-a412-deb0452517a4 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:82c26ca2b315be3e3b17e62f535ed0419679d4f18019cd7469b600767267b278

Observation 30f171c9-5550-477d-b10e-386c8c6a3a32 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:ba266d11e52cc0ca54e64d888a5ec41745e03cda8a6a2603b40f0619c70e01a4

Observation 163df406-cacd-4f23-9923-1dc5599c9bc8 · outbound

This paper cites an unresolved cited work.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:00f5437fc186abf88959a4db68aeb02f0572031fb2e523a56b11d8c2451185f7

Observation 6f6731d0-2cb5-46a0-8379-006a217dc5c9 · outbound

This paper cites When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.664354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:08d9918ab2c61ef6bcbf9907320f6783416d7101307ccbef66b9c0b938a71b09

Observation ec4655f5-24ac-495b-9951-0dc8b407ad74 · outbound

This paper cites From Where Things Are to What They Are For: Benchmarking Spatial-Functional Intelligence in Multimodal LLMs.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning From Where Things Are to What They Are For: Benchmarking Spatial-Functional Intelligence in Multimodal LLMs

Reference 34

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:50.658839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:406b36757888ca2db517bd31c1f35f2b6a66daa40747b1abb04181fca362d3ba

Observation 973e3003-c948-46c9-94b5-f531e1ed3ed5 · outbound

This paper cites Think3d: Thinking with space for spatial reasoning.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning Think3d: Thinking with space for spatial reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:23:50.677720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:90e3b639435828f74e53acd945116ef8209fa19fad8c71504e514cb487a9e9be

Observation 6570a2c6-f5b5-47f8-9291-8d2452c90763 · outbound

This paper cites online" 'onlinestring :=.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning online" 'onlinestring :=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:768635a07c4be061ad424cc2150fe6020e205807b7e97a3199dd3eff77791f50

Observation 1436ab21-ad17-4501-9848-73b8f80678c5 · outbound

This paper cites write newline.

How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning write newline

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-29T18:19:40.323661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-06-29T18:19:40.323661Z digest=sha256:a899273fdf5681e4993e544f14f0758ba0c2dbd37966562bc09b3743a0a997ae

Pith citing papers

Observation 0d11f08f-6272-47f8-8f0b-86454827cea6 · inbound

See2Think: Do Multimodal Models Really Use Intermediate Visual States? cites this paper.

See2Think: Do Multimodal Models Really Use Intermediate Visual States? How and What to Imagine? Visual Thinking in Unified Multimodal Models for Cross-View Spatial Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-30T21:38:08.608340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:38:08.608340Z digest=sha256:4fc9da793cde34c10ef5420b9aa7bea2e8ba3f271f8de33cb4f0e98f74110777