Pith. sign in

Paper Citation Record · LEDGER

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System

As of 19 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2504.19266.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.19266 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T06:01:20.008438Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact2
  • verified fuzzy9
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 211a10ca-b3a2-4fb7-9e80-5b2890714ac7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Learning transferable visual models from natural language supervision,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.884012Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.884012Z digest=sha256:20fd1a8e4a5f87ee4050127fc8e5cab61c76a5d7fba6a585848f2c7698e0f238

Observation ba96b6eb-c730-4041-940b-ce05dfa59553 · outbound

This paper cites Towards Real-Time Open-Vocabulary Video Instance Segmentation.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Towards Real-Time Open-Vocabulary Video Instance Segmentation

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:01:20.339228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.889564Z digest=sha256:04c09b665305bb6acf2b175fab4c57133a6d9af33567bb3148ea7de96141e444

Observation 8e1b0f0f-0f9e-4af3-b1a3-d4618f754737 · outbound

This paper cites Segment anything,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Segment anything,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.894678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.894678Z digest=sha256:a7ad5955a207e5545ec2f98d610473c99d50c1b7567e538171abfba5aecd463a

Observation 0b1f0870-de5a-4075-aef6-f77843ef9e90 · outbound

This paper cites Open-fusion: Real-time open-vocabulary 3d mapping and queryable scene representation,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Open-fusion: Real-time open-vocabulary 3d mapping and queryable scene representation,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.545563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.899601Z digest=sha256:88f5786ae0fc338da8c2ee08163d9160a4a1ba23a561f0b05eefd606cbe5f64b

Observation aaa9dcf7-3b82-47ac-987b-4107aba062b8 · outbound

This paper cites Segment everything everywhere all at once,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Segment everything everywhere all at once,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.528518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.904379Z digest=sha256:78a6c90f851c31919a239065049b2ec9e0076a9a815f63c42b46d9ad6bf9cb38

Observation 48d4f719-6b1d-4ce2-a721-cfd3df2940a1 · outbound

This paper cites Emerging properties in self-supervised vision trans- formers,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Emerging properties in self-supervised vision trans- formers,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.909434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.909434Z digest=sha256:4e98ef3b2e39ef1c76c43bac183d026b678691be2afe5cd7e649e884aebe384f

Observation f4c9d245-538a-4a60-bfeb-7d5ffb57cf0d · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System DINOv2: Learning Robust Visual Features without Supervision

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.914668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.914668Z digest=sha256:5942f935191322331fd4eaad620de13663de79eca7d0bba9e95583e2eb5c084b

Observation 6d5890fa-2329-46ef-b887-9ea8c5bf8d60 · outbound

This paper cites Masked autoencoders are scalable vision learners,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Masked autoencoders are scalable vision learners,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.919561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.919561Z digest=sha256:5d0d288164e0c81436ccbe5a3d4d852af647879d3d569957fe172705c373ed7f

Observation 78c7aaaf-2ae8-4f83-8065-b15d0b4dc6cf · outbound

This paper cites Groupvit: Semantic segmentation emerges from text su- pervision,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Groupvit: Semantic segmentation emerges from text su- pervision,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.491569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.924146Z digest=sha256:a7059c6d73d86e3eb3dc4ac21c90e64f194c42314da46d03ca059f72f866b04d

Observation 78867969-f314-4e6b-bffa-7ed5313e4cc2 · outbound

This paper cites Segclip: Patch aggregation with learnable centers for open-vocabulary semantic segmentation,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Segclip: Patch aggregation with learnable centers for open-vocabulary semantic segmentation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.474991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.928719Z digest=sha256:0959611f8d778477e9104df0c39abf618087d050c18a51b179f863e9ea0dc31e

Observation 80d092ab-a67f-4dbd-9b95-e47135bd7a3e · outbound

This paper cites ConceptFusion: Open-set Multimodal 3D Mapping.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System ConceptFusion: Open-set Multimodal 3D Mapping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.933372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.933372Z digest=sha256:26a0b3b57b29f8a4fecd5afcf90196a7f46d6c750f61533ef238877ff40dfc34

Observation e4d06987-2b26-49da-8241-5494471f59a6 · outbound

This paper cites OpenMask3D: Open-Vocabulary 3D Instance Segmentation.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System OpenMask3D: Open-Vocabulary 3D Instance Segmentation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.938518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.938518Z digest=sha256:d96a626f138959f53a71cbd419412d063ef9020f3762767f0ae60d5d60b865ea

Observation 6c54ed42-9483-43ac-a740-76096e5b125c · outbound

This paper cites OpenSU3D: Open World 3D Scene Understanding using Foundation Models.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System OpenSU3D: Open World 3D Scene Understanding using Foundation Models

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-16T06:01:20.265661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.943725Z digest=sha256:5e76a9c0a145df4178ce7844b25b09b0706b88aa2282bdebd620104292d11f09

Observation e513ed1c-06c7-4855-9105-b31b1197a44e · outbound

This paper cites Open3dis: Open-vocabulary 3d instance segmentation with 2d mask guidance,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Open3dis: Open-vocabulary 3d instance segmentation with 2d mask guidance,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.461007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.948573Z digest=sha256:8d3f8b9019ffbf1fead3fbdf5e79921aada970f9af2b0bea2d3989761129fd06

Observation c6d641b0-8571-402b-8a56-9cd7c8fec23a · outbound

This paper cites Openins3d: Snap and lookup for 3d open-vocabulary instance seg- mentation,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Openins3d: Snap and lookup for 3d open-vocabulary instance seg- mentation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.445353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.953359Z digest=sha256:bdf43857775ae8674fd2cfc9d1d792e19e9ee924414b5cdf6d682fcdfdc45a33

Observation ef83b21e-3546-4ea2-8714-fd4293978937 · outbound

This paper cites Openscene: 3d scene understanding with open vocabularies,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Openscene: 3d scene understanding with open vocabularies,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.957962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.957962Z digest=sha256:2178360363b0f10bf582a36a9e55e7f39a5ac3246545a3a9d8549784de3ff6f2

Observation 3f78897c-78ed-4d60-8899-83c5098246be · outbound

This paper cites SAM3D: Segment Anything in 3D Scenes.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System SAM3D: Segment Anything in 3D Scenes

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.962484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.962484Z digest=sha256:ba1ca39b556fac22a9aaa4690c0d3ac1630cb1c86918179c830460ed7768b8e5

Observation 301e25d7-72d8-4695-bad8-47271c290d2f · outbound

This paper cites Maskclustering: View consensus based mask graph clustering for open-vocabulary 3d in- stance segmentation,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Maskclustering: View consensus based mask graph clustering for open-vocabulary 3d in- stance segmentation,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.418923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.967223Z digest=sha256:77e3026bf8b907605ab40936e734f4b027a5b8d6943e51c4fa21e2fda82c5f6a

Observation 73ca5631-1d76-481f-bb24-95b6dd1cb2b5 · outbound

This paper cites EmbodiedSAM: Online Segment Any 3D Thing in Real Time.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System EmbodiedSAM: Online Segment Any 3D Thing in Real Time

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.971626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.971626Z digest=sha256:fde373f0972b990af0f99a497433ef362f1ff3c9e28de17703a5433e02106410

Observation a2e383e4-5220-4a32-ba2b-201533813b99 · outbound

This paper cites Fast Segment Anything.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Fast Segment Anything

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.976378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.976378Z digest=sha256:b5aa4d7b1a1e7665e5bc6f49695ce40919b238aacd4f0506bc70207589463693

Observation 14d6a0eb-a576-445e-b201-d9f559e5bf3e · outbound

This paper cites PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System PanoSLAM: Panoptic 3D Scene Reconstruction via Gaussian SLAM

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.980414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.980414Z digest=sha256:b1f7f5ac57abddd0eb0e3846b491649da9ed494fe48d6054b7952301d6764fd9

Observation a444c903-3256-4a98-8f5d-6a79b65147fd · outbound

This paper cites 3d gaussian splatting for real-time radiance field rendering.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System 3d gaussian splatting for real-time radiance field rendering

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.984489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.984489Z digest=sha256:b5e2b03550c6c901835e5c6edf3277189c772b7d873228cd88e53aa9c0ef1cf5

Observation cc05642a-3518-478c-9f91-21d2b644ba70 · outbound

This paper cites Ovo-slam: Open- vocabulary online simultaneous localization and mapping,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Ovo-slam: Open- vocabulary online simultaneous localization and mapping,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.988243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.988243Z digest=sha256:b4fb4e1bc8bc206e0dee6a5475a70af9a4a547c66a4f30cd6ba6d087fd469553

Observation 6b04ccc2-b69f-4f5e-a606-861d050d0271 · outbound

This paper cites Alpha-clip: A clip model focusing on wherever you want,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Alpha-clip: A clip model focusing on wherever you want,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.393256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.992078Z digest=sha256:d25f42c4b8891b6cb1b9fec2d6932c5ab7700d3bbd1e9d8dd3cb548c17047fcd

Observation e89a8251-3c5c-4a60-a70e-315bda75d025 · outbound

This paper cites A benchmark for rgb-d visual odometry, 3d reconstruction and slam,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System A benchmark for rgb-d visual odometry, 3d reconstruction and slam,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T06:01:20.377365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-16T06:01:19.995911Z digest=sha256:2dd8c895f97c767554a349c72bee541e8a11922e9043e9b6667f9cfa9a8c7cf4

Observation 014c1a8b-09b6-42b5-ba89-65535de154eb · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Scannet: Richly-annotated 3d reconstructions of indoor scenes,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:19.999615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:19.999615Z digest=sha256:98ca7784e22e7ce9fe9bd9920630eab6f20d5324948a80fad31edd83f69c1ee5

Observation c5647868-c814-4cf8-8953-8095915f61b7 · outbound

This paper cites Scannet++: A high- fidelity dataset of 3d indoor scenes,.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System Scannet++: A high- fidelity dataset of 3d indoor scenes,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:20.003349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:20.003349Z digest=sha256:02d87faefbe6640844d906e84f327286e150e657ace2af49cd563b821d4a7815

Observation 4dd3d3ca-0aad-430f-b188-43c28522d0eb · outbound

This paper cites The Replica Dataset: A Digital Replica of Indoor Spaces.

OpenFusion++: An Open-vocabulary Real-time Scene Understanding System The Replica Dataset: A Digital Replica of Indoor Spaces

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T06:01:20.008438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T06:01:20.008438Z digest=sha256:686a81f4be5f8e2f8aa90b60dc03025ed918440ed6e991fc5f90d0ca9e68b106

Pith citing papers

No inbound Pith citation observations are available.