Pith. sign in

Paper Citation Record · LEDGER

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA)

As of 17 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2508.08781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.08781 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:25:17.677115Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact0
  • verified fuzzy49
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 908425ca-f249-4141-8a1c-91594e33ddf0 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:28.854476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.161460Z digest=sha256:058135788dc7aa87fb79639081a621529ed0642097263b0f442cc0ec1100ddc2

Observation c6ee75ca-2bce-45d9-9358-348bfe27eefc · outbound

This paper cites Matterport3d: Learning from rgb-d data in indoor environments.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Matterport3d: Learning from rgb-d data in indoor environments

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:28.675005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.322846Z digest=sha256:096ead9416c9154ca23866f50c421451f93605f52656e472051a1e373b9a65c4

Observation 9c8f2c4d-223c-43cb-aa5c-408581fe01a0 · outbound

This paper cites 3d-future: 3d furniture shape with texture.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) 3d-future: 3d furniture shape with texture

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:28.512740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.416523Z digest=sha256:1e12dbf0c8ef7cbe3a3e199071f8b160ed48b807b7583aba7939e5f8df5485a3

Observation 54a1b864-86e7-48d2-a345-580ccbfe57a2 · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Referit3d: Neural listeners for fine-grained 3d object identification in real-world scenes

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:28.357745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.497984Z digest=sha256:aa06d66cbd1c4b1bfe0a03ae7b031f2cded21d964018a394277299307c5a06e5

Observation 0d543063-92a7-4896-afd9-ef32c2ba2958 · outbound

This paper cites Scanrefer: 3d object localization in rgb-d scans using natural language.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Scanrefer: 3d object localization in rgb-d scans using natural language

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:28.156284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.607607Z digest=sha256:f3fe0e5a1b6c15c0243db18b39ff36dbd1435a74e04b25eef5a667653b2cbdd6

Observation 98ecac41-2724-4359-a618-c53dc2ad1830 · outbound

This paper cites Vigil3d: A linguistically diverse dataset for 3d visual grounding.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Vigil3d: A linguistically diverse dataset for 3d visual grounding

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:27.948099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.724139Z digest=sha256:38e376707fc6790b502c159f7ae5b57fcb7e0195a43ab38beecbf9c0ced8e015

Observation 7d414922-00a9-4ed6-b0fc-a1b8dbd4161e · outbound

This paper cites 2D Image-Based 3D Scene Retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) 2D Image-Based 3D Scene Retrieval

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:27.710291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.803739Z digest=sha256:0da4bc33e464f2ec5bb39adfbce234cd896dd6363606b1c780d3679d88fa3a9c

Observation eac076f4-b0a8-45bc-9b04-ae676ec8d509 · outbound

This paper cites SHREC’19 track: Extended 2d scene image-based 3D scene retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) SHREC’19 track: Extended 2d scene image-based 3D scene retrieval

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:27.508957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.874240Z digest=sha256:ba3d30e281028649d0ad081f931d97c7ef2372cce1e68a85d3105d7f96ed93ee

Observation 4ce0ebc7-7e0e-4519-9837-9d31f55e9d93 · outbound

This paper cites SHREC 2019- monocular image based 3D model retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) SHREC 2019- monocular image based 3D model retrieval

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:27.329756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:13.984373Z digest=sha256:3e9ee95cbbe4576af7fd980f5d592928166865505b64cc847065d3f99cdbe7b6

Observation 41029d4c-0c40-4def-a64d-b4c590d4be2d · outbound

This paper cites SHREC 2020 track: extended monocular image based 3d model retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) SHREC 2020 track: extended monocular image based 3d model retrieval

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:27.115793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.062921Z digest=sha256:bbdf2d6c8a62556a6a176b83f1f025dd8bc139b4a088b0016a30051ab594120d

Observation c02271db-68dc-4c24-8ccf-f8d73acad5ee · outbound

This paper cites SHREC’22 track: Open-set 3D object retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) SHREC’22 track: Open-set 3D object retrieval

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:26.885801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.175559Z digest=sha256:2356e7f17f4974ba4b88d457f132fc1111af4d144cb020fab4dfeca6e4377326

Observation 727da715-548e-4968-b5c1-f02e45b39242 · outbound

This paper cites Shrec’22 track: Sketch-based 3D shape retrieval in the wild.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Shrec’22 track: Sketch-based 3D shape retrieval in the wild

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:26.702204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.293566Z digest=sha256:db2d8ca020471fd90b951842b8183b06911c761557920993d73f0b9a608d603d

Observation aabbc13d-32c0-491f-a087-50efad8448dd · outbound

This paper cites Sketchanimar: Sketch-based 3d animal fine-grained retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Sketchanimar: Sketch-based 3d animal fine-grained retrieval

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:26.459662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.406030Z digest=sha256:05296850a607e198b48c4aa5f6d130b6e9e971e551ab7c8d3320f665973841df

Observation 8cad2707-da89-4c5c-9a44-2a88304525fc · outbound

This paper cites Textanimar: text-based 3d animal fine-grained retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Textanimar: text-based 3d animal fine-grained retrieval

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:26.270129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.520462Z digest=sha256:96cb82c8e0fa82a46366e01e1bf086b9ab864444fd61ab388e91c93a60297047

Observation 5fa9aee1-d699-4a35-aa2c-aed2553f0fb3 · outbound

This paper cites Multi3drefer: Grounding text descrip- tion to multiple 3d objects.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Multi3drefer: Grounding text descrip- tion to multiple 3d objects

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:25.996119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.600140Z digest=sha256:4608d1d07bbd586b2fc49d026d91f919537e790c2101c8d629ff2c23eac8fe18

Observation b42c0903-5264-4d00-97df-fe311a848d0d · outbound

This paper cites Referitgame: Re- ferring to objects in photographs of natural scenes.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Referitgame: Re- ferring to objects in photographs of natural scenes

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:25.738564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.672220Z digest=sha256:846fa7db5f4d21dee41ae89487b97f4fe48a3aa013bd43dccc4d36b8408b7cec

Observation b049efca-5fc8-48ba-8501-15e632697b7b · outbound

This paper cites Generation and comprehension of unambiguous object descriptions.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Generation and comprehension of unambiguous object descriptions

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:25.413430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.746469Z digest=sha256:6e8e6b33c52dc2f4978c92ebd547ec95368428274e61ad9aa99dcfb7fb97928e

Observation 3add20de-8449-4727-81f4-469bb3ed41c8 · outbound

This paper cites Microsoft coco: Common objects in context.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Microsoft coco: Common objects in context

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:25.151296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.857670Z digest=sha256:57b75f33026e085de75f285cd652f832e77455fd37252627b678f9b66c09be74

Observation 1b206f2a-a1ce-4895-ae23-a073fa2f75e2 · outbound

This paper cites Mat- tnet: Modular attention network for referring expression comprehension.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Mat- tnet: Modular attention network for referring expression comprehension

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:24.894486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:14.934659Z digest=sha256:91568d676c6b95ef9911c56091051996d49380a4d07432dbbbcf2b176752a691

Observation 9906e3ea-37b6-4bd5-8722-ff06bf744d3b · outbound

This paper cites Lavt: Language-aware vision transformer for referring image segmentation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Lavt: Language-aware vision transformer for referring image segmentation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:24.569899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.052393Z digest=sha256:119101a96e98f32d6d521b2ac8110dc4ff4b4ade0a255f0a516b4dd2f625bb66

Observation eb583db1-cb6f-4a8f-bd89-3c6a2df6aa74 · outbound

This paper cites Language as queries for referring video object segmentation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Language as queries for referring video object segmentation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:24.232623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.170560Z digest=sha256:6ff53fc352c88e122f151d76e2f973a88976396ce7e67f5d770b76ec65c79a1a

Observation bde0f82c-a628-4a16-b3ff-a9bd6a4a39db · outbound

This paper cites Masked- attention mask transformer for universal image segmentation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Masked- attention mask transformer for universal image segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:23.913268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.280592Z digest=sha256:2ed158eac0eaa51e7bf0658d6863b41a8e063fad3781863e5449362fefaa6668

Observation 71f5c14f-4d87-436c-964e-7f6e5effa588 · outbound

This paper cites Partnet: A large-scale benchmark for fine-grained and hierarchical part-level 3d object understanding.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Partnet: A large-scale benchmark for fine-grained and hierarchical part-level 3d object understanding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:23.637999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.390197Z digest=sha256:cdf20324025794c2434ab4882f45395b20bdcf30a93a97ff38821da866239197

Observation 79a5bf6e-3ed6-492b-9f36-107425363d91 · outbound

This paper cites Point2cad: Re- verse engineering cad models from 3d point clouds.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Point2cad: Re- verse engineering cad models from 3d point clouds

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:23.351673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.575342Z digest=sha256:636f58e3edba4012368efe5aec564723cd6f32400f7d0a422c17d08d9a410c9d

Observation 7448b06e-1ca2-457f-b2ec-9ee1bb59d801 · outbound

This paper cites Mask3d: Mask transformer for 3d semantic instance segmentation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Mask3d: Mask transformer for 3d semantic instance segmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:23.149860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.684977Z digest=sha256:e6396a5bfc548dd038eecbda2e26d4339885e7aa972b3109272c8bf9f7ad2712

Observation 21e17340-e1f2-48d5-a2d2-4084029708d2 · outbound

This paper cites Refmask3d: Language-guided transformer for 3d re- ferring segmentation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Refmask3d: Language-guided transformer for 3d re- ferring segmentation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:22.956126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.791327Z digest=sha256:db8cf155304610638cb8348965796df411420d0d2c13729eea94a3c95ec1abe7

Observation 00fde7c4-0c89-4834-85aa-f3f3c8dc2396 · outbound

This paper cites Segment any 3d object with language.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Segment any 3d object with language

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:22.768006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.865239Z digest=sha256:d7188009aa4bc34d42418cb3a91d08841dd409cdb3bda70cae45b53e1c558fa3

Observation 57ab3f5b-7194-4d3f-b3ba-a992e7751266 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Learning transferable visual models from natural language supervi- sion

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:22.558827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:15.953163Z digest=sha256:7bf3a829a88dc1270dc17223e874a67e893640bc2b87cbf92fa58142c2408f4a

Observation c5f68a0b-cf18-45ed-b49a-0114f1839a43 · outbound

This paper cites Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:22.360633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.083636Z digest=sha256:b26636ffaee156d867022da38334fe868267d2cf0ed2a9cca32ccce5f2357e65

Observation 80557ae1-996b-4b35-8aab-db600304fd46 · outbound

This paper cites Ulip-2: Towards scalable multimodal pre-training for 3d under- standing.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Ulip-2: Towards scalable multimodal pre-training for 3d under- standing

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:22.093300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.220337Z digest=sha256:5cc32fc2acc3b0dc25b786f219c921e33fe275cde10134c5647e4334eff70acd

Observation 1811c936-cfdf-40dd-8b4b-44bd13e1fc07 · outbound

This paper cites Objaverse-xl: A universe of 10m + 3d objects.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Objaverse-xl: A universe of 10m + 3d objects

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:21.859885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.268945Z digest=sha256:5689ce209f09d541df4a97d3722b0323e8c204ade4b6892f1c34a7f4e09b5540

Observation 5038030e-26b9-4dbb-aadc-9c0b6ed5c010 · outbound

This paper cites 3ur-llm: An end-to- end multimodal large language model for 3d scene understanding.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) 3ur-llm: An end-to- end multimodal large language model for 3d scene understanding

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:21.671040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.293856Z digest=sha256:bdbad16a13e8c346c1163f1bde4d6f403eb1dbcf5262019354541c67e02ded28

Observation e5cb46bc-f9f4-485e-8324-2fd973cf68c3 · outbound

This paper cites Shrec’14 track: Extended large scale sketch-based 3D shape retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Shrec’14 track: Extended large scale sketch-based 3D shape retrieval

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:21.477746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.313808Z digest=sha256:c2021f636479f565ad5d7341269ac66069b8343f779d36d26bacc18e9f30f9ff

Observation be49a04d-cd1c-4147-bf70-ecbaabfb8ece · outbound

This paper cites Shrec’19 track: Extended 2d scene sketch-based 3D scene retrieval.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Shrec’19 track: Extended 2d scene sketch-based 3D scene retrieval

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:21.243865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.346368Z digest=sha256:a5125dc64746918597e05f0cc4cd6283b3f60dc8a5b2fa1099c3de7c20df8a72

Observation c0b136c7-ae59-46d3-b211-208454d868df · outbound

This paper cites Llama: Open and e fficient foundation language models.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Llama: Open and e fficient foundation language models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:21.052840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.397587Z digest=sha256:926472bcd1af5abe8ec795514da06f962a585541aa164a731eebcc07982a1cd2

Observation 5e4e45af-8777-49da-a876-d63683d32337 · outbound

This paper cites Sigmoid loss for lan- guage image pre-training.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Sigmoid loss for lan- guage image pre-training

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:20.863625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.452044Z digest=sha256:5a391a5dd35b4a87d5c4f7601bdb64b8f1000d2a231270471af2c1ecb5209967

Observation 62f3a624-0279-4398-9612-2c5f00bbdd7f · outbound

This paper cites Scaling instruction-finetuned language models.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Scaling instruction-finetuned language models

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:20.575769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.553225Z digest=sha256:94456278570611f3c93cf8213170d47c2f387ed94c31ccc67d8624998919b57f

Observation 1c43f6dd-b8a1-4864-ab99-3a7bd87b6d7b · outbound

This paper cites C-pack: Packed resources for general chinese embeddings.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) C-pack: Packed resources for general chinese embeddings

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:20.252637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.718367Z digest=sha256:4a6161509ccff699e6bfaeb870d881d203b2dfb9972dcde216cd8bbf028cccd3

Observation 8d8cacc6-f7a5-4bc8-a672-20da75f6b208 · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:19.938726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.783471Z digest=sha256:41f1eccf6cd58670e9599747e3a4dacf000a893253bd8be91d020ea2f8603124

Observation 0566c228-841f-482f-9f83-e0ed6531b04b · outbound

This paper cites Point-bert: Pre- training 3d point cloud transformers with masked point modeling.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Point-bert: Pre- training 3d point cloud transformers with masked point modeling

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:19.660909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.860381Z digest=sha256:75ef4f09a19ee1f9fff2a30c90663ad9306e5c60b44b1434ef174632937e7e08

Observation 6c895884-222c-478a-bc7a-2c228b6cb2d2 · outbound

This paper cites Openshape: Scaling up 3d shape representation towards open-world understand- ing.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Openshape: Scaling up 3d shape representation towards open-world understand- ing

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:19.378025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:16.924365Z digest=sha256:5b9b47a61d1e7c98f0c750d5272e37994449ded66ecd4381c4f0e9a00e657e70

Observation 9f94caa9-3555-41c8-949d-6c19c09cd5db · outbound

This paper cites Panoformer: panorama transformer for indoor 360 ◦ depth estimation.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Panoformer: panorama transformer for indoor 360 ◦ depth estimation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:19.171635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.000637Z digest=sha256:8bcc4ceb26063c7ad091293cdc0946d1fe3fcb8197cda0d4c661dee075310765

Observation b7d45613-40e5-4150-8254-f0e1574ff875 · outbound

This paper cites Florence-2: Advancing a unified representation for a variety of vision tasks.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Florence-2: Advancing a unified representation for a variety of vision tasks

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:19.009548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.064923Z digest=sha256:bca47d2fbeb207440cef42325444f4466bdf249afa06b06dea6e50cccfef5c77

Observation 5a958535-9815-456b-af78-624a2f4cdc32 · outbound

This paper cites Yolov11: An overview of the key architectural enhancements.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Yolov11: An overview of the key architectural enhancements

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:18.784683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.226238Z digest=sha256:00e3a7b2fc99f21b017c8f975894ca16ed774bb8156719305bf7d885800fc432

Observation b4184087-79fa-4ec8-b4a7-bbb284299f3e · outbound

This paper cites Exaone 3.0 7.8 b instruction tuned language model.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Exaone 3.0 7.8 b instruction tuned language model

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:18.608849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.329602Z digest=sha256:ed27a023e2f3de36e22c21586173f9d50d3b12fec705e94c2176087bbb30a1bc

Observation 74cff093-7cf5-42f4-90f7-a474f86c80e9 · outbound

This paper cites Zero- painter: Training-free layout control for text-to-image synthesis.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Zero- painter: Training-free layout control for text-to-image synthesis

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:18.410207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.439074Z digest=sha256:4525b02b8763f310856f5fd51d514ccf0fd4443c87d9f2c92c354670dabb6c14

Observation 9b39378d-6b1d-409e-a051-c6aa4e8cfa39 · outbound

This paper cites Enhanc- ing the reasoning ability of multimodal large language models via mixed preference optimization.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Enhanc- ing the reasoning ability of multimodal large language models via mixed preference optimization

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:18.217754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.505731Z digest=sha256:ad5adff40450886f4c8237a85398dc100444cdc4e5608210d8dde862e1cbd218

Observation 2a23ca6e-6f77-4f49-a04d-302b39384ee2 · outbound

This paper cites Reproducible scaling laws for contrastive language-image learning.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Reproducible scaling laws for contrastive language-image learning

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:18.042020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.579956Z digest=sha256:8f0d1772c3de8d0d13e46355df68bd4b8bbb528e0c2cd55506cdd6bf151e74c9

Observation fc105b44-ad2b-4b5d-a186-507712f81c5b · outbound

This paper cites Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling.

SHREC 2025: Retrieval of Optimal Objects for Multi-modal Enhanced Language and Spatial Assistance (ROOMELSA) Expanding performance boundaries of open-source multimodal models with model, data, and test-time scaling

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:25:17.869062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-05T21:25:17.677115Z digest=sha256:b30a4d108f45dd92b395706e39d7360f3966dd06c0a53371d5c51e629fec6459

Pith citing papers

No inbound Pith citation observations are available.