Pith. sign in

Paper Citation Record · LEDGER

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

As of 18 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 5 inbound Pith citation observations for arXiv:2507.07781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07781 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:37:21.425460Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-15T13:51:30.008232Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:37:29.656433Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact5
  • verified fuzzy27
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9376cfc3-aed9-4cf8-83f9-69393bdf2cb9 · outbound

This paper cites Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanents3d: Exploiting phrase-to-3d-object correspondences for improved visio-linguistic models in 3d scenes.WACV, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.624177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.114837Z digest=sha256:ac082392f133ac109de979ca935fe4cf5fe4c3838a39db7f148550e5ce399ef0

Observation d8b9af87-1faa-4dd8-b234-3bcb510b0aef · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.491641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.120880Z digest=sha256:7ed51f7c8518bc7b4a869cd48f1bcfbb5fd87d289755d84f825568aa99904f5f

Observation 1b40e861-a928-4ed0-997f-f53026a46edb · outbound

This paper cites Scanqa: 3d question answering for spatial scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanqa: 3d question answering for spatial scene understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.126873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.126873Z digest=sha256:ae8a524fc8620da5f16a6ea0351d10572a2303e52f17c74cbeaa28c4a610c7de

Observation b7e840de-bc4e-4b8c-901a-a6f8dc94ff4b · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.133258Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.133258Z digest=sha256:3fb445d4bbdd7a0332efb2b582006d72dae3cdd3b6bfda3ede1d0beea302f7ad

Observation 543dca29-9e36-4ef3-9fa9-ac2c3d0652e9 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.323272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.139122Z digest=sha256:220cd98efa063b619f6cc7445e68f9537eb99f7185a241ee7c6b762b6d585cc3

Observation 7aa50b6c-d6c4-4909-84da-bdd7a04867a1 · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:26.181776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.145147Z digest=sha256:aa7aee67972913c8a0c89f964b633950966f0b4f8d3f3a51bcb14ef281ef34dc

Observation a49de642-bb8a-4a85-ab87-769d11004151 · outbound

This paper cites Towards label-free scene understanding by vision foundation models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Towards label-free scene understanding by vision foundation models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:26.042482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.151006Z digest=sha256:42705f51ec304cf3cbb1d2ecdeca176538b946a6e63a80a79705fcb947579438

Observation bf533800-795b-499b-937b-7d9eb46bb329 · outbound

This paper cites Clip2scene: Towards label-efficient 3d scene understanding by clip.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Clip2scene: Towards label-efficient 3d scene understanding by clip

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.903944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.156005Z digest=sha256:fd0bf08fb880f786073f51f630c8d6295f11e3c4798269510354349393240fb6

Observation 4148b6cc-5b0e-4a17-aed6-75c5f3173d47 · outbound

This paper cites OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes OVGaussian: Generalizable 3D Gaussian Segmentation with Open Vocabularies

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.160848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.160848Z digest=sha256:8a2804ce3f8da12a9ec241ff78ff564c085d2c3a6d7ae1e79f80e3f30cf271d9

Observation c6b1ba35-5ff8-42c0-bb3c-8e907b4ee5c9 · outbound

This paper cites Zero-shot point cloud segmentation by transferring geometric primitives.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Zero-shot point cloud segmentation by transferring geometric primitives

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.363226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.166231Z digest=sha256:555e0d587c735487e32c0dd682227452532a7da2cff67e67501be0a80662e352

Observation 241bf559-752a-4323-ac7b-128ab4e7f507 · outbound

This paper cites Bridging language and geometric primitives for zero-shot point cloud segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Bridging language and geometric primitives for zero-shot point cloud segmentation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.784832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.171663Z digest=sha256:1dfbcdb2395169320425c32d7687536853cbab0dd5c681d8328a90982fb92452

Observation b452f0b8-dea8-4990-ad3f-496a7d70925b · outbound

This paper cites Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Ll3da: Visual interactive instruction tuning for omni-3d understanding reasoning and planning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.684857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.176775Z digest=sha256:94330499ca2c55eac0da912e637b262cb7e33309bba587cdc1dc483bc6c315b3

Observation b292152c-adad-48cc-b2e8-98a6979bbdb0 · outbound

This paper cites Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reasoning3D -- Grounding and Reasoning in 3D: Fine-Grained Zero-Shot Open-Vocabulary 3D Reasoning Part Segmentation via Large Vision-Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.181823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.181823Z digest=sha256:cc0027be6f9a17948169f84a3498694e4f7807f0b251a62ef26a08a17241db14

Observation 93c3ecf0-d3c8-466b-af6a-67e62075dabd · outbound

This paper cites Grounded 3D-LLM with Referent Tokens.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Grounded 3D-LLM with Referent Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.187570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.187570Z digest=sha256:faebdf11fa760314b9bdd9d0d4ab49ed7ab8838381c79ffb1a845e12267aadc2

Observation 3d085221-29a1-4e6c-a68a-63ec232906cd · outbound

This paper cites an unresolved cited work.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:37:25.545431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.193512Z digest=sha256:9304404df02080b95f3c07d2b2f45ac00fb710452f82d17a440dbce46e4e327d

Observation b5a4291c-4900-4b9b-9bad-50c7e71bef18 · outbound

This paper cites Segpoint: Segment any point cloud via large language model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Segpoint: Segment any point cloud via large language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.369542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.199440Z digest=sha256:031d4e24f81757a626b68f70fda500ea03421db73fa12939c9bc052f736d35a7

Observation 1c0ae59c-f287-4a9d-b403-2c10bf660c0b · outbound

This paper cites 3d concept learning and reasoning from multi-view images.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d concept learning and reasoning from multi-view images

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.223082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.205237Z digest=sha256:7f2cf90de60673e48a39357aa3b6eb62ad608cd5a7785a5ac0ba882fba54b653

Observation 88a24fde-1d8c-443a-8169-e0f0c819eac2 · outbound

This paper cites 3d-llm: injecting the 3d world into large language models.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-llm: injecting the 3d world into large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:25.076517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.211899Z digest=sha256:aec919b7f4c2fb7a31aba92770d0252253d30ec3b1fadd0b8d17ecd3f2331e36

Observation 38fc496f-7a4c-4614-af2a-f2a6eef7d609 · outbound

This paper cites Chat-scene: Bridging 3d scene and large language models with object identifiers.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-scene: Bridging 3d scene and large language models with object identifiers

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.930825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.217983Z digest=sha256:9a85665451baeb763a9c161cb3c1984fe2a4f5c40d955d1dd08abf41c435e6ef

Observation f7ee941a-0953-47fe-822d-78787c2cc680 · outbound

This paper cites An embodied generalist agent in 3d world.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes An embodied generalist agent in 3d world

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.808362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.224453Z digest=sha256:995d2f00819f4b17df28cb86d78c76d9d7dd6f6cfa77b4931f98f840e4508083

Observation 66773a7b-6149-40a1-9c5a-b97c8fc6ddeb · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation.arXiv preprint arXiv:2503.18135, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.231214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.231214Z digest=sha256:47879f81a9cc2d616f053cb3ff0f60741135abb7ad4f403107db4d67358f6669

Observation 7600b532-ee8d-4d1b-8f5d-8ae1bade0159 · outbound

This paper cites Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Mllm-for3d: Adapting multimodal large language model for 3d reasoning segmentation, 2025

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.674946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.237038Z digest=sha256:25808438e6612fd51d8d199b8b80fd077937d5052ae1799ded37ffed10c144b4

Observation 042d3286-fb21-4c1f-ab19-3971b06480f7 · outbound

This paper cites Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Reason3D: Searching and Reasoning 3D Segmentation via Large Language Model

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:22.046731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.243530Z digest=sha256:e5917ada6be6c400f0e2983ca2d5bf437278b264dedd6bbbf6abdc4cab7016a2

Observation 2d6570b7-2cda-4a48-95c0-ecf1afc66378 · outbound

This paper cites Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Text-guided graph neural networks for referring 3d instance segmentation.AAAI, 35(2):1610–1618, May 2021

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.529985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.255021Z digest=sha256:0f7f739a4a3c4659272dfda4c1197469045f43df7915b2062c821d122e1108bd

Observation d9ac1b62-ec57-42e0-8d96-4246995d0ff7 · outbound

This paper cites Dense object grounding in 3d scenes.ACM Multimedia, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Dense object grounding in 3d scenes.ACM Multimedia, 2023

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.383244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.261876Z digest=sha256:2b9960e4b44148c48159cd327a477040474b0719cd9401cc7e79c3f826353f75

Observation 4471163f-cb6b-4db4-89ab-5737e3b369ce · outbound

This paper cites Sceneverse: Scaling 3d vision-language learning for grounded scene understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sceneverse: Scaling 3d vision-language learning for grounded scene understanding

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.255373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.267560Z digest=sha256:6eea2384eaf8746ec3ee0d34ff78565ad961687267b75a8173b970e4e580d081

Observation c9d89ce4-7ef4-476e-a7ca-79ec368e6539 · outbound

This paper cites Multimodal 3D Reasoning Segmentation with Complex Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multimodal 3D Reasoning Segmentation with Complex Scenes

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.860973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.273496Z digest=sha256:665af422b8831aa29996ea24b99ae11e86dae40c792c61c3a54be6a6a1ebe229

Observation 6c0c30db-6c45-4dbf-bf87-63ee7a7f99dd · outbound

This paper cites Intent3d: 3d object detection in rgb-d scans based on human intention.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Intent3d: 3d object detection in rgb-d scans based on human intention

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:24.091965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.280081Z digest=sha256:f14cd328bd43d0fb8fdb837045aa61679b7284c44d0544c4b807a3008ace5191

Observation 83c48785-f887-4b67-be35-c9b7c3c962ab · outbound

This paper cites M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes M3dbench: Towards omni 3d assistant with interleaved multi-modal instructions

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.955399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.288173Z digest=sha256:e5fd3e5ded53208c59ebbd9404af0acc02ea7a1a86153ad1d88f6dd4598a74b8

Observation 362f372f-b1d3-4e5f-a3aa-0d1015d2618b · outbound

This paper cites 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3DMIT: 3D Multi-modal Instruction Tuning for Scene Understanding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.293763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.293763Z digest=sha256:243ca14aa9c96fd56d2a528326a39ab2bc50fb801cb7996397c13985db7bcbf0

Observation b948b5bd-2019-432c-820e-060b18784fe6 · outbound

This paper cites Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Multi-modal situated reasoning in 3d scenes.NeurIPS, 2024

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.821728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.299608Z digest=sha256:50a7aaffcc5d02b009e97191592ae57b219a036d7af05947e16f139d2bc94a99

Observation e141879a-7cc9-41f9-bcd1-4302cbbbcea4 · outbound

This paper cites See more and know more: Zero-shot point cloud segmentation via multi-modal visual data.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes See more and know more: Zero-shot point cloud segmentation via multi-modal visual data

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.645328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.307224Z digest=sha256:7821b6d7d337aa10d282d209fd3fd6033b8a5f7d87a7987d3e04afd78c43b476

Observation 346a79c3-8c57-4270-93cb-b5eba08d6163 · outbound

This paper cites 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3dsrbench: A comprehensive 3d spatial reasoning benchmark.arXiv preprint arXiv:2412.07825, 2024

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.314108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.314108Z digest=sha256:d947639fa2738f8be33b36171e44fd3f0d7219b9376e26d3f7e001ea16f97b17

Observation 99404787-2c76-42ee-98e0-c7d6561e17b2 · outbound

This paper cites Sqa3d: Situated question answering in 3d scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Sqa3d: Situated question answering in 3d scenes

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.319192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.319192Z digest=sha256:0db20603bfae3ff3d1b89b94323d6b7ecb00681ff4d7ff8f71f149af8ba5d6bb

Observation fd250474-66e0-431b-bd7d-d4b83e2ddde7 · outbound

This paper cites X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes X-refseg3d: Enhancing referring 3d instance segmentation via structured cross-modal graph neural networks

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.395769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.324456Z digest=sha256:33819ee643aed6eef2da4c02dc6e9a35aa274a3cf3b50cf0b51e60bd60385519

Observation 181c2de8-7ad0-41ba-9f83-b112b78c4865 · outbound

This paper cites Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fully convolutional networks for semantic segmentation.IEEE TPAMI, 39(4):640–651, 2017

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.282816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.330745Z digest=sha256:059e569f7d6293f776c448638823ef4f304e24cd1510e7b339f18ced100764bc

Observation 644d44f1-f96e-423b-b35c-a0403c355991 · outbound

This paper cites Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Embodiedscan: A holistic multi-modal 3d perception suite towards embodied ai

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.182958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.337102Z digest=sha256:fc8461ff618deb90fda83fe29531c99953bd6e2f5aa934970aa861af73486e65

Observation 3770243f-952a-4a6c-94dd-05f355ed0d06 · outbound

This paper cites Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Chat-3D: Data-efficiently Tuning Large Language Model for Universal Dialogue of 3D Scenes

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.344604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.344604Z digest=sha256:38311d82814d077922f2d9571eb77f0469f4c31000c2f394c188a7d35a490723

Observation b23ebd52-1067-4957-95b2-552b4f7e0a31 · outbound

This paper cites 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-stmn: Dependency-driven superpoint-text matching network for end-to-end 3d referring expression segmentation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:23.052030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.352680Z digest=sha256:abbec2a9695df480f4caa09ffca548313d18664934b64279071b9662d69b8dea

Observation eac10f64-5bb7-472e-a662-ba3269557217 · outbound

This paper cites Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Com- prehensive visual question answering on point clouds through compositional scene manipulation, 2023

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.908862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.360014Z digest=sha256:9a1a452117ab6f1b53f5a12dd80b70aac9cf18d4fe937430ece6fd356661477d

Observation c2c71b5d-d935-4927-bd90-1dc2685e2feb · outbound

This paper cites Fouhey, and Joyce Chai.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Fouhey, and Joyce Chai

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.814955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.367267Z digest=sha256:1249b4d3676013202f3308d6d45df826e16aebe18087a3e354b77523bf8ef032

Observation e9854bb2-90b6-42a3-b153-64c907f4714c · outbound

This paper cites Scannet++: A high-fidelity dataset of 3d indoor scenes.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Scannet++: A high-fidelity dataset of 3d indoor scenes

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.375086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.375086Z digest=sha256:97495c677379cb7d1b37200e5574505bbb1f2f4db10946954ba7b0eb9f0dbdc2

Observation 30d36a3c-73ef-4ec8-8854-6d3fbf77c3c3 · outbound

This paper cites ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ExCap3D: Expressive 3D Scene Understanding via Object Captioning with Varying Detail

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.594527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.381545Z digest=sha256:6417a29ad1bdcc2d95a1107d80054fb3d5c14202e3e16421ae542f5516ffc349

Observation 6a899754-8dde-43f0-912c-07f2f9f8c2bb · outbound

This paper cites Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes Toward Explainable and Fine-Grained 3D Grounding through Referring Textual Phrases

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:37:21.542408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.389400Z digest=sha256:e0424f1ca76749c2cb8c76c8318ba8af6aecc135e1b3f51cb6ac92cee97fdfb0

Observation 4253dfeb-ddb6-4e97-9581-e8379f3a1fde · outbound

This paper cites VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes VLA-3D: A Dataset for 3D Semantic Scene Understanding and Navigation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.399482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.399482Z digest=sha256:b48e9ac52821c62ec8ee4420a56245a84980aa795783276bc2102a9f30093cf9

Observation 869ee334-947e-452a-a4cf-0c9bcf0a93f8 · outbound

This paper cites ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes ScanReason: Empowering 3D Visual Grounding with Reasoning Capabilities

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:37:21.413309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:37:21.413309Z digest=sha256:c3ef9158a13d8971cd1a481a4d0f4a0db5f9cc0310ed913f7f3a23dbc3d7449a

Observation 773e9acc-2f00-4620-961d-89669b1f0225 · outbound

This paper cites 3d-vista: Pre-trained transformer for 3d vision and text alignment.

SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes 3d-vista: Pre-trained transformer for 3d vision and text alignment

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:37:22.621098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-06T18:37:21.425460Z digest=sha256:e0092c20ee40abf9c71ea469402e1b3a342ef71a9a29cc0039f40d7a8133a342

Pith citing papers

Observation 6a07e2e5-fa18-4df6-a88c-1e53ebdf9428 · inbound

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing cites this paper.

RDSplat: Robust Watermarking for 3D Gaussian Splatting Against 2D and 3D Diffusion Editing SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:33:44.661273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T00:33:29.282349Z digest=sha256:a4d629271ba57d22e6614f62282547a4b401a67b993e156a89dee0488a272f58

Observation f83cbea5-9f4e-44d1-bdd0-89ab94a8b873 · inbound

What if? Emulative Simulation with World Models for Situated Reasoning cites this paper.

What if? Emulative Simulation with World Models for Situated Reasoning SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-15T13:51:30.008232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:51:30.008232Z digest=sha256:5bc6e5faa5acbd2dd799b9dd2731e083535137b47271cd1112a2a230946f0548

Observation bba1b6d4-3609-4ea9-a515-fc025b6adf68 · inbound

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models cites this paper.

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:29.657958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T17:08:22.609095Z digest=sha256:2a9707274606dd29b695f5ed3202bde782f9e19f8ae5714cc19d9c6d2b7bcd16

Observation e147e5dc-eb09-4b09-b902-16ed734977d4 · inbound

Holo-Captioning: Toward the Text Equivalent of 3D Scenes cites this paper.

Holo-Captioning: Toward the Text Equivalent of 3D Scenes SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T06:12:48.722467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:12:48.722467Z digest=sha256:18f44ec7b55dee27f075b5112d6cac953317e924e8ec260a5e2ba47d74670a2e

Observation 01684f2e-885d-4528-b5d0-182573a6ee2e · inbound

G$^2$TAM: Geometry Grounded Track Anything Model cites this paper.

G$^2$TAM: Geometry Grounded Track Anything Model SURPRISE3D: A Dataset for Spatial Understanding and Reasoning in Complex 3D Scenes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-11T23:56:52.009530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:56:52.009530Z digest=sha256:3558810f7e737eeadfd999ec110100e2779c2f5916033867d9c8b9eb0e4d5a7b