Pith. sign in

Paper Citation Record · LEDGER

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

As of 7 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 5 inbound Pith citation observations for arXiv:2507.07781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07781 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:37:21.425460Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T13:51:30.008232Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:37:29.656433Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact5
  • verified fuzzy27
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9376cfc3-aed9-4cf8-83f9-69393bdf2cb9 · outbound

This paper cites Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.624177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.114837Z digest=sha256:189c8a855fb1f47f075029508f347e58e156b22c732f31e55ea9d3d78a59d96e

Observation d8b9af87-1faa-4dd8-b234-3bcb510b0aef · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.491641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.120880Z digest=sha256:a6ffb845b560f6fd114dc168be399f34b5599d6f4ebfcb7a62ddbaf52cbe8a63

Observation 1b40e861-a928-4ed0-997f-f53026a46edb · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanqa: 3d question answering for spatial scene understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.126873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.126873Z digest=sha256:a7e2d0c2e77c0f9e516a4d70347a0c44bca23c65e4440756158d2b2cd223251b

Observation b7e840de-bc4e-4b8c-901a-a6f8dc94ff4b · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.133258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.133258Z digest=sha256:6f5a710c59cd6364cbd019d3f5efdf08557bf7d0b0e8610747bc434d305b51a7

Observation 543dca29-9e36-4ef3-9fa9-ac2c3d0652e9 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.323272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.139122Z digest=sha256:27b77b0ccb32ebaf84aa63aef42aaa90c76bd91859bc08195c0fb672d5f784ff

Observation 7aa50b6c-d6c4-4909-84da-bdd7a04867a1 · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:26.181776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.145147Z digest=sha256:ff8dd971c16c92cee7c5f29fc64bb5d56a77de32603deaa9ffcec760c9cf746d

Observation a49de642-bb8a-4a85-ab87-769d11004151 · outbound

This paper cites Towards label-free scene understanding by vision foundation models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Towards label-free scene understanding by vision foundation models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.042482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.151006Z digest=sha256:79ae8fef7eba9261eac8e577278ff0eb9e8241d74671942da529b455ac7938e5

Observation bf533800-795b-499b-937b-7d9eb46bb329 · outbound

This paper cites Clip2scene: Towards label-efficient 3d scene understanding by clip.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Clip2scene: Towards label-efficient 3d scene understanding by clip

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.903944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.156005Z digest=sha256:ebc8765ffb96edb91140bd792d4d3f885bffefab21409b4aff6fbe711193ccc4

Observation 4148b6cc-5b0e-4a17-aed6-75c5f3173d47 · outbound

This paper cites OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.160848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.160848Z digest=sha256:ed78bfb6cc4975a003e23e036bb83177906fadda16e2a874d8f4ce1e22506845

Observation c6b1ba35-5ff8-42c0-bb3c-8e907b4ee5c9 · outbound

This paper cites Zero-shot point cloud segmentation by transferring geometric primitives.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Zero-shot point cloud segmentation by transferring geometric primitives

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.363226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.166231Z digest=sha256:c4196abe82b81cc77632d070dd0a3bb8c556d9512003409cad65e0d48c1fbfc1

Observation 241bf559-752a-4323-ac7b-128ab4e7f507 · outbound

This paper cites Bridging language and geometric primitives for zero-shot point cloud segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Bridging language and geometric primitives for zero-shot point cloud segmentation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.784832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.171663Z digest=sha256:2a83e1f9160736480de45f3823ec971a714b145e074e39f2fc2ba09440c93278

Observation b452f0b8-dea8-4990-ad3f-496a7d70925b · outbound

This paper cites Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.684857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.176775Z digest=sha256:4f508d36f73f8f709324efa794147850bf9d5dc07c06fee993f40a937c148633

Observation b292152c-adad-48cc-b2e8-98a6979bbdb0 · outbound

This paper cites Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.181823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.181823Z digest=sha256:4dfab700137a86d72be7dd607e9c3dc2835fb7afdc27031c9c3eacfc993cfe62

Observation 93c3ecf0-d3c8-466b-af6a-67e62075dabd · outbound

This paper cites Grounded 3D-LLM with Referent Tokens.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Grounded 3D-LLM with Referent Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.187570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.187570Z digest=sha256:b3de864468a64bf92643c2eed95f48777ad8ec316ae90297f0481b4a491b0e79

Observation 3d085221-29a1-4e6c-a68a-63ec232906cd · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:25.545431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.193512Z digest=sha256:f149cea36978ba3b96160c59fd6f47fc9d5fbb2191afed219eae18c70fc75490

Observation b5a4291c-4900-4b9b-9bad-50c7e71bef18 · outbound

This paper cites Segpoint: Segment any point cloud via large language model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Segpoint: Segment any point cloud via large language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.369542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.199440Z digest=sha256:f1fafe9947dcfc947d8e76cd42483cb8b95c987b415cf6c9ccf05f42f72f355b

Observation 1c0ae59c-f287-4a9d-b403-2c10bf660c0b · outbound

This paper cites 3d concept learning and reasoning from multi-view images.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d concept learning and reasoning from multi-view images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.223082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.205237Z digest=sha256:5db92fd95a3ab6853b3aea39990693da14ebe50bcffa74b83edd00bc18a6bf9a

Observation 88a24fde-1d8c-443a-8169-e0f0c819eac2 · outbound

This paper cites 3d-llm: injecting the 3d world into large language models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-llm: injecting the 3d world into large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.076517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.211899Z digest=sha256:377d809af46560cdd8d0efc08af2b8dbd37b911378d1f5e21b3d1b5656cda834

Observation 38fc496f-7a4c-4614-af2a-f2a6eef7d609 · outbound

This paper cites Chat-scene: Bridging 3d scene and large language models with object identifiers.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-scene: Bridging 3d scene and large language models with object identifiers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.930825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.217983Z digest=sha256:404d8041c0cdd7d6fa9e06ecea91686d4931065070f5e55286a5db2da42c7436

Observation f7ee941a-0953-47fe-822d-78787c2cc680 · outbound

This paper cites An embodied generalist agent in 3d world.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes An embodied generalist agent in 3d world

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.808362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.224453Z digest=sha256:a03405999375aab17c8bfd78289ad083e43842a3ee8a735044a7a91c16c0374a

Observation 66773a7b-6149-40a1-9c5a-b97c8fc6ddeb · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.231214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.231214Z digest=sha256:62f9940db7cad1e5ec3fc494584d1f1796b29f2d53fe5e51549890e02c478f9c

Observation 7600b532-ee8d-4d1b-8f5d-8ae1bade0159 · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.674946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.237038Z digest=sha256:c61e8fcd3b37741af309ea8d4fca71b3a3ecfaeabb84120eb4aeaddb472f549e

Observation 042d3286-fb21-4c1f-ab19-3971b06480f7 · outbound

This paper cites Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.046731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.243530Z digest=sha256:397b9137be6484b1c185947cab25ac12e3ab32ab476d936522087ba4bc9942a0

Observation 2d6570b7-2cda-4a48-95c0-ecf1afc66378 · outbound

This paper cites Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.529985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.255021Z digest=sha256:7b3a7bfeaad0173ce341399355a7e2ee0c6519546502f6df859673005034be21

Observation d9ac1b62-ec57-42e0-8d96-4246995d0ff7 · outbound

This paper cites Dense object grounding in 3d scenes.ACM Multimedia, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Dense object grounding in 3d scenes.ACM Multimedia, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.383244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.261876Z digest=sha256:c0811068f3918e7d9ac72f1a4268425f8c8120707bb6dafbfa2d089c7dbe1440

Observation 4471163f-cb6b-4db4-89ab-5737e3b369ce · outbound

This paper cites Sceneverse: Scaling 3d vision-language learning for grounded scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sceneverse: Scaling 3d vision-language learning for grounded scene understanding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.255373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.267560Z digest=sha256:96e052c14b4fbccdf968b6b6ff5fd84d7fdb8132eea09b730a80f10e9e978cb7

Observation c9d89ce4-7ef4-476e-a7ca-79ec368e6539 · outbound

This paper cites Multimodal 3D Reasoning Segmentation with Complex Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multimodal 3D Reasoning Segmentation with Complex Scenes

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.860973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.273496Z digest=sha256:57f9b545abf5addbc33a2a55483af25164c9389edd18f8dcab9b28a65fe61138

Observation 6c0c30db-6c45-4dbf-bf87-63ee7a7f99dd · outbound

This paper cites Intent3d: 3d object detection in rgb-d scans based on human intention.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Intent3d: 3d object detection in rgb-d scans based on human intention

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.091965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.280081Z digest=sha256:c53c177c7d29a8977ea367f1da09c4b147346c8c0e959702c29228ebdd304b42

Observation 83c48785-f887-4b67-be35-c9b7c3c962ab · outbound

This paper cites M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.955399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.288173Z digest=sha256:1f01a3fdcc832f5f6fa447fce32c3b9005bd80248dd713d6bc0a172160a76619

Observation 362f372f-b1d3-4e5f-a3aa-0d1015d2618b · outbound

This paper cites 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.293763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.293763Z digest=sha256:98cb1c6efd8e26dc52241a94516cc64f5a290a1d1f8c01d4b3f437fe7c2ce772

Observation b948b5bd-2019-432c-820e-060b18784fe6 · outbound

This paper cites Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.821728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.299608Z digest=sha256:d0ff599185e1ecae3be20f8f686590dbcb023e2266ceccfd69c15a0cd0e89890

Observation e141879a-7cc9-41f9-bcd1-4302cbbbcea4 · outbound

This paper cites See more and know more: Zero-shot point cloud segmentation via multi-modal visual data.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes See more and know more: Zero-shot point cloud segmentation via multi-modal visual data

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.645328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.307224Z digest=sha256:cbe9922dc1a05467d94e3a46a002283be9a430915e23ced2c9d14c7dd1db3c2e

Observation 346a79c3-8c57-4270-93cb-b5eba08d6163 · outbound

This paper cites 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.314108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.314108Z digest=sha256:5b91281c24ccf3b120f7bd32dc2f6008663fb3fb4cba0f56baf41263eb6f8bdc

Observation 99404787-2c76-42ee-98e0-c7d6561e17b2 · outbound

This paper cites Sqa3d: Situated question answering in 3d scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sqa3d: Situated question answering in 3d scenes

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.319192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.319192Z digest=sha256:21061f2a55cde6e45708cc68983cc810ae61ed9df3c5253ae253510f80926e49

Observation fd250474-66e0-431b-bd7d-d4b83e2ddde7 · outbound

This paper cites X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.395769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.324456Z digest=sha256:e1077175eabe21cb58080bbb6b549d48a67621b757cfb4508299801913dce94a

Observation 181c2de8-7ad0-41ba-9f83-b112b78c4865 · outbound

This paper cites Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.282816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.330745Z digest=sha256:a2996f8f463e085dc9d060c0196624a9b4c967c55195857795600048449ab0ec

Observation 644d44f1-f96e-423b-b35c-a0403c355991 · outbound

This paper cites Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.182958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.337102Z digest=sha256:47b3d0b2394cda060d7050886204401782603c8fe51b2342b7f14bebfef0a7c0

Observation 3770243f-952a-4a6c-94dd-05f355ed0d06 · outbound

This paper cites Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.344604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.344604Z digest=sha256:04323bf9343cbb13f9a4243a2eea8794bd1796a6be226acc9cd15b3ad5f54b5d

Observation b23ebd52-1067-4957-95b2-552b4f7e0a31 · outbound

This paper cites 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.052030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.352680Z digest=sha256:434021facc0f744196803cfa89b956f05a7adbdaf0970d89ed916f986a1f65f1

Observation eac10f64-5bb7-472e-a662-ba3269557217 · outbound

This paper cites Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.908862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.360014Z digest=sha256:fccebbb7daf007d086a4080bb0173efd213fcf7846636c5ca1646183861eaecd

Observation c2c71b5d-d935-4927-bd90-1dc2685e2feb · outbound

This paper cites Fouhey, and Joyce Chai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fouhey, and Joyce Chai

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.814955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.367267Z digest=sha256:ba83820444478df40ee834fa411912dee907e4e1a2a5c9fb8f91594be0bc9533

Observation e9854bb2-90b6-42a3-b153-64c907f4714c · outbound

This paper cites Scannet++: A high-fidelity dataset of 3d indoor scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scannet++: A high-fidelity dataset of 3d indoor scenes

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.375086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.375086Z digest=sha256:d6c08daeed139ec04c4e35bc2c3feea67ef3bf063d77b32c2378cd9fb5b0b843

Observation 30d36a3c-73ef-4ec8-8854-6d3fbf77c3c3 · outbound

This paper cites ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.594527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.381545Z digest=sha256:d57c616e4ad082e6a51ed7bc37cbc1b7550b2ce012ad42f6692aafdb1e9989a1

Observation 6a899754-8dde-43f0-912c-07f2f9f8c2bb · outbound

This paper cites Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.542408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.389400Z digest=sha256:562bb2633b224552510d45516b5b1f9149232448e642f577fca2ef52de779824

Observation 4253dfeb-ddb6-4e97-9581-e8379f3a1fde · outbound

This paper cites VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.399482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.399482Z digest=sha256:859d45b4de568aef4abd0811d0c5f2eb5e98f16de3808bebb5d0b6a87cb7a9b4

Observation 869ee334-947e-452a-a4cf-0c9bcf0a93f8 · outbound

This paper cites ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.413309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.413309Z digest=sha256:7c9c53e091bef785c9c99ee7589f3ee6bd7065914e8af8f3125e55c10bf85143

Observation 773e9acc-2f00-4620-961d-89669b1f0225 · outbound

This paper cites 3d-vista: Pre-trained transformer for 3d vision and text alignment.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-vista: Pre-trained transformer for 3d vision and text alignment

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.621098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T18:37:21.425460Z digest=sha256:c738ef1ca37284eae7f5dc881a4efc1c11707c9ac371c924e8dc288c4c5a2bf2

Pith citing papers

Observation 6a07e2e5-fa18-4df6-a88c-1e53ebdf9428 · inbound

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing cites this paper.

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:33:44.661273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T00:33:29.282349Z digest=sha256:12cd4c7bf97f8a4f867e8acc360200cb82601c24a08ad2b37827022b3498bf88

Observation f83cbea5-9f4e-44d1-bdd0-89ab94a8b873 · inbound

What if? Emulative Simulation with World Models for Situated Reasoning cites this paper.

What if? Emulative Simulation with World Models for Situated Reasoning SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T13:51:30.008232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:51:30.008232Z digest=sha256:3157d13478525493e7b89adee1d18b17e08293c7a56ef1a47db4c3552f7c3541

Observation bba1b6d4-3609-4ea9-a515-fc025b6adf68 · inbound

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models cites this paper.

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:29.657958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T17:08:22.609095Z digest=sha256:d40a22696d659476ec2f2e32797890608bc0038094d9ffaa6ba1e864600a61b8

Observation e147e5dc-eb09-4b09-b902-16ed734977d4 · inbound

Holo-Captioning: Toward the Text Equivalent of 3D Scenes cites this paper.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:3a680f1efaeba4680c9169b5bdda1c4a17605a42b24faa1c2062340ddc59b403

Observation 01684f2e-885d-4528-b5d0-182573a6ee2e · inbound

G$^2$TAM: Geometry Grounded Track Anything Model cites this paper.

G$^2$TAM: Geometry Grounded Track Anything Model SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T23:56:52.009530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:56:52.009530Z digest=sha256:1dedd6c6caa516eaafa0dd2ac8d9c0605a767bf954474d000b425ddf02538e49