Pith. sign in

Paper Citation Record · LEDGER

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model

As of 8 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 3 inbound Pith citation observations for arXiv:2506.04837.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.04837 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:38:52.289632Z

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T22:00:13.802598Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T10:31:25.365391Z

Reference resolution

16 of 16 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 938b140a-ae7c-4a69-8125-7e1cae4193f3 · outbound

This paper cites Henriques, Andrew Zisser- man, and Andrea Vedaldi.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Henriques, Andrew Zisser- man, and Andrea Vedaldi

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:54.616520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.250747Z digest=sha256:5919f8f52f8ea2eb701f132f234922dcf3b2a27810e625e46a7b7018e9c9d1d2

Observation 8dc444f3-b219-45cf-9fb4-dcf35f7b2d9e · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:54.436610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.327719Z digest=sha256:5ac5631761f64ecf6de274067e2565b2826f82b1e4ec8c3e8e6a3154207690b7

Observation e5d3c5f8-78b3-4bb6-83ae-68ecd6c7a3d6 · outbound

This paper cites Image Token Acc@0.25 Acc@0.50 mIoU ✗ 51.4 35.5 36.7 ✓ 55.1 40.2 40.3 (b) 2D Features for Segmentation.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Image Token Acc@0.25 Acc@0.50 mIoU ✗ 51.4 35.5 36.7 ✓ 55.1 40.2 40.3 (b) 2D Features for Segmentation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:54.296577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.442838Z digest=sha256:935db2256999efaf8339c2fc10c676b1bdef4b0e58996c624ce8c8acc74d132e

Observation a982400b-1dc2-4216-8e16-98deb750aeff · outbound

This paper cites Openins3d: Snap and lookup for 3d open-vocabulary instance segmentation.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Openins3d: Snap and lookup for 3d open-vocabulary instance segmentation

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:54.143351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.614827Z digest=sha256:98947036daaf54eb556771babba75391c19a183208fe56c2e56f869470595ea3

Observation 5f6d4232-9cb1-4c9b-9aaf-790f9c9b5bf0 · outbound

This paper cites Lerf: Language embedded radiance fields.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Lerf: Language embedded radiance fields

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.997837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.724330Z digest=sha256:6d0ab49f3ced806ac0493ad5af96aa5c3a260cf5f0395d65769971c9d32f37d4

Observation ed87ee76-174d-415f-995f-af7a93b8c9ed · outbound

This paper cites Decomposing nerf for editing via feature field dis- tillation.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Decomposing nerf for editing via feature field dis- tillation

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.851299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.822720Z digest=sha256:519109e8e697227b8683e75ed732a48f5518d0be285aef3db523b2bd1160a45c

Observation 72cbf191-f244-4f98-81e8-5a3eb246a52c · outbound

This paper cites Weakly supervised 3d open- vocabulary segmentation.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Weakly supervised 3d open- vocabulary segmentation

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.714392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:50.935217Z digest=sha256:24be19b5599a3054b3e807d1f0763205ba1c76c41459eceb3aa2b65539f6e841

Observation 2fb5fc72-94f4-4212-986c-122e471e3611 · outbound

This paper cites Ovir-3d: Open-vocabulary 3d in- stance retrieval without training on 3d data.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Ovir-3d: Open-vocabulary 3d in- stance retrieval without training on 3d data

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.490044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:51.039850Z digest=sha256:1a81a283d478137f9933566c9a0c8ee1848bdf60ddad0cfe118ecc38a344c617

Observation c30d6fbc-d8c7-4272-bbad-818e7f978504 · outbound

This paper cites Openscene: 3d scene understanding with open vocabularies.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Openscene: 3d scene understanding with open vocabularies

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.230575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:51.137020Z digest=sha256:08a3a886250391d85d3f3cbca4953e6aa3ffb5d196bc4f05b616d7e90a45cc39

Observation 2b807271-8275-475a-b801-63742fdc18a8 · outbound

This paper cites LangSplat: 3D Language Gaussian Splatting.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model LangSplat: 3D Language Gaussian Splatting

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:51.322064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:38:51.322064Z digest=sha256:d3550e08fcfc635c2f6f635e67191741f5364c1163a9221e585c57530dbdae99

Observation b4b4e6a2-546a-46a9-bff2-2d96dd175d5a · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Learning transferable visual models from natural language supervi- sion

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:51.420204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:38:51.420204Z digest=sha256:a01440ba04b6d015ea1a6d64257b68831bc25afb591984f8e570bdf274ba3817

Observation 31366615-60c0-44ce-9604-9bfc4052eef5 · outbound

This paper cites Sumner, Marc Pollefeys, Federico Tombari, and Francis Engelmann.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Sumner, Marc Pollefeys, Federico Tombari, and Francis Engelmann

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:53.066921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:51.542471Z digest=sha256:3c3054c7ac73cbfbcc79a5aeac66a47534dc033a56607ee01583f08d8519064e

Observation 166a6fdf-be22-4ab2-9254-7f4ec7a554a4 · outbound

This paper cites VL-Fields: Towards Language-Grounded Neural Implicit Spatial Representations.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model VL-Fields: Towards Language-Grounded Neural Implicit Spatial Representations

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:38:52.698058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:51.702038Z digest=sha256:129605b94d8391e472b7cbb22ba9711bf35af05f4e11acb248b38533c4a8e019

Observation ba7513b3-55e8-412c-8168-c208336dd54f · outbound

This paper cites Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Open-Fusion: Real-time Open-Vocabulary 3D Mapping and Queryable Scene Representation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:51.869988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:38:51.869988Z digest=sha256:f046b9c3750869922a4dd499ba63b1059c4547680650dbb8d78050b29980281e

Observation e51fb179-e8d9-4678-b368-8d8c863744ab · outbound

This paper cites RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model RegionPLC: Regional Point-Language Contrastive Learning for Open-World 3D Scene Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:38:52.071377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:38:52.071377Z digest=sha256:e2be6e0b96657b0bbad14c68f0e02630d251366a5df693e3d312f60647a86dfe

Observation 0568b568-42fb-49ea-af52-aa847b8874db · outbound

This paper cites Clip-fo3d: Learning free open-world 3d scene representations from 2d dense clip.

OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model Clip-fo3d: Learning free open-world 3d scene representations from 2d dense clip

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:38:52.867422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:38:52.289632Z digest=sha256:e6a4167921e0d2329a83457bc28bde68280d15febb4b35e891449399f32d0d90

Pith citing papers

Observation e8f1562a-0e30-4ebd-8319-587ecf1edfff · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:16:34.740516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:12:46.385646Z digest=sha256:1dc46296004a4ac6592660c3ec74455f91e5acfd0cb3009c7f49390b3d3c14c5

Observation a95306e0-ac56-4181-be9d-bea87b5a387f · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-22T10:31:25.368092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T10:30:06.829915Z digest=sha256:32a4938e591a900343986627411643ed703ff60aea32c3431ef906bcb3ad04da

Observation 6e32a0b0-e437-430c-b73f-cb9f38343f12 · inbound

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs cites this paper.

MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs OpenMaskDINO3D : Reasoning 3D Segmentation via Large Language Model

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-02T22:00:13.802598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:00:13.802598Z digest=sha256:eccafb0a40b97c2be26ed77bab7f5948ad48e7388099469c93769797f984b448