Pith. sign in

Paper Citation Record · LEDGER

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion

As of 10 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2506.15610.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15610 v3

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:57:23.112112Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved16
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1810e1c3-0f9a-48e4-825b-965a552770ad · outbound

This paper cites Omni3d: A large benchmark and model for 3d object detection in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Omni3d: A large benchmark and model for 3d object detection in the wild

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:29.098750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:19.710060Z digest=sha256:07f71beb057ba6d87b949a6bdcbe1ede6325cc6d25a3c96fc98282e0224c822d

Observation 2c0833e2-f403-4540-b463-5e4af2c0c3db · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.953346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:19.773113Z digest=sha256:99dbde1b12bce5a3a2fb7eb2e31811f616f6ce82d8f4afb54589921d493afde2

Observation 0b8c31d0-66db-44f9-9ef4-7319f9e8b31b · outbound

This paper cites CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:19.841794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:19.841794Z digest=sha256:3452390370d2e2345e1b4144fe76a5299c582c7fdcd07d328c96a57f62435703

Observation 388626aa-3df2-4a49-9063-f2f5c4043019 · outbound

This paper cites Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.757700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:19.882914Z digest=sha256:626972cc1e12ea188274b2055595bc2ee764f65499803ed3be1e1890c80a69ac

Observation 74be73ef-6495-406f-9033-f653c0842a43 · outbound

This paper cites A hierarchical graph network for 3d object detection on point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion A hierarchical graph network for 3d object detection on point clouds

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.627155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:19.973335Z digest=sha256:9ea856e669a59340be4c3262577e3b9e4272a8fc3defdc34f344a5ddc4f84246

Observation 075ef2eb-c0fd-4efd-ad99-59efe427d5b2 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.052396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.052396Z digest=sha256:4bb99fb786676d5380d7a49fabe5902883ed8590439c6d76c1f03d5d9f7ed9b3

Observation 38ade343-5ea4-4662-bfd7-b99dbd518673 · outbound

This paper cites Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.354338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.088537Z digest=sha256:edb1259dff252b645a3c0ee7fa480fe5d498965e4464cad70cc8fbb5715ba238

Observation d1050601-59d1-432a-8b23-85b38cba320d · outbound

This paper cites Disarm: Displacement aware relation mod- ule for 3d detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Disarm: Displacement aware relation mod- ule for 3d detection

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.112593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.164335Z digest=sha256:56749bae9ea339a4297868ed760ccd9c97155079633c2a6ae7f0983301d7fcef

Observation d8b0e6e4-1366-422a-ad17-2a5387f33182 · outbound

This paper cites 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.230116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.230116Z digest=sha256:980060a077ea71f208d6d073d35ba467a514a2c13ca213280ac080c8b12c750d

Observation 31c53f9e-0276-4e75-b58f-926c262a9e92 · outbound

This paper cites Generic objects as pose probes for few- shot view synthesis.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Generic objects as pose probes for few- shot view synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.266963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.266963Z digest=sha256:985b3ada09ef127410ff89e2ebaf09c3dd1c4a0473014d5f32a02d0d9c989679

Observation 0d60c4cf-3e4f-4aa1-be89-2bee002f3ca6 · outbound

This paper cites Training an open-vocabulary monocular 3d detection model without 3d data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Training an open-vocabulary monocular 3d detection model without 3d data

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.903048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.342896Z digest=sha256:443d9a1d32d0df7df9ecbdf4a915ce282ebfeb9528ff219ac2dcd7631f79da37

Observation f90b80f4-833b-4283-954a-ebe59cff9416 · outbound

This paper cites Particle filter with swarm move for optimiza- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Particle filter with swarm move for optimiza- tion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.648559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.428794Z digest=sha256:95dfdc8562f55c4e0db8903792c4687ea99a330569d5678cf5cacfcd9ff49dee

Observation deac8d6b-e5d5-4cd7-bbc0-6472fae19635 · outbound

This paper cites Open-vocabulary 3d semantic segmentation with foundation models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary 3d semantic segmentation with foundation models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.458507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.468149Z digest=sha256:121d9971d6f5fd22e631040a450bc14d575274e53f1acec7e7be0ebb6202eda0

Observation c84d6efc-034c-45c1-9efa-fc92731b8523 · outbound

This paper cites Segment any- thing.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Segment any- thing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.519699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.519699Z digest=sha256:cc4f6874bc8fcb4e717179c16c41617a673972337653e638ba79714e9e2aaad6

Observation b7fbde06-5851-4147-b246-6cbd89a7e6b2 · outbound

This paper cites Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.208995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.613522Z digest=sha256:1aeb12ee5b6f0e4114696ccf13c6c9e5ed1d00de2c44f9f31acd588e0086e98b

Observation d59def1e-fea6-4475-af13-29d4b993c969 · outbound

This paper cites Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.039717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.681004Z digest=sha256:6c442aeaaadcd1762f559317ec3b1c0028f234671f031bd950695b9c90c84524

Observation 3227a5b0-11eb-43fa-9c32-8ba76137b8f6 · outbound

This paper cites Arm3d: Attention-based re- lation module for indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Arm3d: Attention-based re- lation module for indoor 3d object detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.883501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.729580Z digest=sha256:c3669bf21dfb7b0befb3179281cc178e09b4bdaedfe734f06cb4dc3c4aadebc1

Observation dee36243-c969-4f07-99cd-a68810ab5ec5 · outbound

This paper cites Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction

Reference 18

Resolution
verified exact
raw_fallback, observed 2026-08-06T23:57:23.622737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.762678Z digest=sha256:2b2be7138bbe7efe77025d36370508fd2cdfdbf3cc62c25df33a2bfe8b9ed6ba

Observation c3d0f184-e88b-462e-b7ed-749b937e064d · outbound

This paper cites Cubify Anything: Scaling Indoor 3D Object Detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Cubify Anything: Scaling Indoor 3D Object Detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.836056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.836056Z digest=sha256:a9ad174333222b5ccdb5a499eb84a31eb46bd2306d7ad6f890236a8b230c6349

Observation ec19e30d-dcdb-4eec-bffa-9f5e596b0550 · outbound

This paper cites Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.706827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.923788Z digest=sha256:db9053fed716c3336df33247f2d6fe24eda022556bc9aaacf7d0156dad52e074

Observation 910dc8e6-aeb9-4fc5-a6a5-3799fbe6260d · outbound

This paper cites Ground- ing image matching in 3d with mast3r.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ground- ing image matching in 3d with mast3r

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.588898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:20.959777Z digest=sha256:fa5b96f295aae5c7af49445fa4a7642ae0425367c1b889b9d31a865624d40889

Observation 59b8dedd-56c5-4851-8b47-19d2c55907c0 · outbound

This paper cites Grass: Generative recursive autoencoders for shape structures.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grass: Generative recursive autoencoders for shape structures

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.435530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.029161Z digest=sha256:175be8229a2041f9681f9b201217432141371c44ebf240babd3469cb303b7ed7

Observation 54325d4d-4b34-4b92-a097-d8f71b7b38e0 · outbound

This paper cites Prompting depth anything for 4k resolution accurate metric depth estimation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Prompting depth anything for 4k resolution accurate metric depth estimation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.106713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.106713Z digest=sha256:65b268f04f5662d402e7c45888ffd072b1005a31c5d1a6dab9f6cd2f7600c007

Observation 9575c0d0-f9d0-40b1-b110-fd3ae562ec36 · outbound

This paper cites Microsoft coco: Common objects in context.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Microsoft coco: Common objects in context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.156947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.156947Z digest=sha256:0590f9a59e416acc144c17fc62d2ef29d0c8d95f4ddd629b4cb85cb4af6bfb0a

Observation 780674c6-9d6d-4cff-8789-faf511f73e8a · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.283502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.186493Z digest=sha256:b7296fafc082f75926a450665d88030950f22b73ab84a7414750939b05f82a78

Observation dd589e9a-127b-47bc-907e-1b2799c5740b · outbound

This paper cites Group-free 3d object detection via transformers.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Group-free 3d object detection via transformers

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.161915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.272610Z digest=sha256:14835239e9571f39255c64086cec7c42ea26024b677bdead2eb88dc37acf8db2

Observation 91b030f3-9a99-4eec-a10d-918adbd6017f · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary point-cloud object detection without 3d an- notation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.006363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.345759Z digest=sha256:749b7d52b63154fb2f9729cf742b5a67e9cf69d4ed38562a9ccb2ea607e6905b

Observation 2a316920-c355-4120-94c4-cefac417f91f · outbound

This paper cites Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.853536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.395450Z digest=sha256:8e0b21df0c7954284da7caacb0ce4b35f5385f9b9f7d20d9cbe34ddaaa0cf213

Observation a546f684-a63e-49ea-95b2-b8102b33067a · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.482945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.482945Z digest=sha256:58b145447893554d82eda08b753d090b07f0261f498f6e6f79b298d626a88b9d

Observation 4839aa72-f402-41b5-ae08-ff2ffea84526 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Deep hough voting for 3d object detection in point clouds

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.752058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.554743Z digest=sha256:cec8210bfcb811fc1c18c7c7bb2cd1e4fef94fdc0e5ac3a3cc135fbad5d516b1

Observation b82ca3dc-b805-459e-a51a-130f286f6df0 · outbound

This paper cites Imvotenet: Boosting 3d object detection in point clouds with image votes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Imvotenet: Boosting 3d object detection in point clouds with image votes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.631090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.596405Z digest=sha256:08c131384752f28cd03ff0e68a347ac724de8323be08257c41d272c318a2671d

Observation ef7a074c-2554-4528-ad20-99f58a209b56 · outbound

This paper cites High quality entity segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion High quality entity segmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.480306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.698243Z digest=sha256:f893a35deb785743919c3bf3967bc0d51f70319de46d2dc858d0a99dbe93c402

Observation 6b2d430e-9b97-48aa-96ac-120d0f782967 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.761942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.761942Z digest=sha256:cd9ba4ee6ff02e223dbccf36113ba57c61f33db9e3a9b3107db5816b0cb0af82

Observation df1d47f0-249e-4836-be9c-390e09e8170c · outbound

This paper cites Fcaf3d: Fully convolutional anchor-free 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Fcaf3d: Fully convolutional anchor-free 3d object detection

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.372401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.814853Z digest=sha256:35e2ecf661098757be8ca6365e583e83486e6b3dafd15ce4c31ad9929928fd65

Observation 1c1b4d4e-87f8-4e03-8a56-88ab71b2c584 · outbound

This paper cites Tr3d: Towards real-time indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Tr3d: Towards real-time indoor 3d object detection

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.175588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.930156Z digest=sha256:ba7218da57e4f5eab35e12990249e158e4acf2e9aeba6c98537ac7384b34f3ed

Observation f2a6349b-9088-4ba2-8369-76754884eeb7 · outbound

This paper cites CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.984804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.984804Z digest=sha256:ede89fe47daaa45a4565129948bafb772452deea739e8bfa9949613a22e0b41f

Observation 71bfd0e5-8d45-4aee-b09c-70204a46798a · outbound

This paper cites Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.058640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.095652Z digest=sha256:5c96661811759ac45230987166424332f5e5ffb8e9eb84517bd444eb6a0d24dd

Observation afcaf3b6-a840-4370-ad8d-9f1204a89f5b · outbound

This paper cites OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:57:23.380699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.137587Z digest=sha256:e50dcf9e36fbfc778229cd8b4b4e91b47c7e187526667391f713cf6dba9f1ef1

Observation b51d861a-0292-4570-92da-364bf9ed4fc3 · outbound

This paper cites Spatiallm: Large language model for spatial understanding.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Spatiallm: Large language model for spatial understanding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.873173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.277083Z digest=sha256:722432610bb68ead259ab7073dd777bca3d0d8196c18c86f804c7849ff1e2b08

Observation 80edae4e-0db0-48cd-ac43-a86dd91cdb9f · outbound

This paper cites Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.755358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.318089Z digest=sha256:9dde1abaf513ae3d8c2e6e854e3a6f67af250bcc863168cccf047f15624c7414

Observation 2c411c6a-c3fb-4ff7-ad39-c8b88103abb5 · outbound

This paper cites Dust3r: Geometric 3d vi- sion made easy.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Dust3r: Geometric 3d vi- sion made easy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.649214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.370207Z digest=sha256:8fe48f63032ca992a1ceb9d438c3c6d106b282da0763fe3b5aaaf6e7dcdb6e47

Observation f1c669fc-2e85-4f4a-9ec6-ca30275c7c6c · outbound

This paper cites Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.553960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.452928Z digest=sha256:d7e624d675378b140a424598e54b44e8066df289317d52e43650022e4b884863

Observation 5a1a0568-8def-4c35-b0eb-3b748886b4fb · outbound

This paper cites Mlcvnet: Multi-level con- text votenet for 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mlcvnet: Multi-level con- text votenet for 3d object detection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.451083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.497715Z digest=sha256:48bb6a251e782e6157d2c204917020dbf36c5b0c53a830c8b2c5a4d25e8d06c1

Observation 1ac36b7a-40cb-4bff-b7a9-79acb6ce0495 · outbound

This paper cites EmbodiedSAM: Online Segment Any 3D Thing in Real Time.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion EmbodiedSAM: Online Segment Any 3D Thing in Real Time

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.527722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.527722Z digest=sha256:b7889c160598a5aa60e8cb4fd0d9d462994cfde5594bb531caddbf44fbaa1a4b

Observation ff68512c-d9b6-40c6-b356-a34588984168 · outbound

This paper cites Memory-based adapters for online 3d scene perception.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Memory-based adapters for online 3d scene perception

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.343165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.557788Z digest=sha256:12fbdf4a6708b67dd2ee2a41aa59ff5ad7d1e7d6505962199d008e103ef688ff

Observation 6cf57c01-6a2e-455c-8279-9dd6305f69c8 · outbound

This paper cites M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.240521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.628208Z digest=sha256:c74d349da4e0aff62951f72e728338fca638d232f87f88bb9f06397b37d96317

Observation 74bc0535-9c10-4fa6-9e84-dc79c263eb7d · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.146333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.685438Z digest=sha256:28bcc86f4dd9892db460dd0cc73c9b67afcb870e9c933324fdc5d0c24077a47e

Observation 015464ae-5d32-4ab9-9b2c-29a2ac39231c · outbound

This paper cites SAM3D: Segment Anything in 3D Scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion SAM3D: Segment Anything in 3D Scenes

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.728460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.728460Z digest=sha256:5baaf8755dc7a1a003e7c1e5fd6631ceea4d7246779726323ce1da1f01ee5d19

Observation 552dd05e-4ae1-4825-bcb0-e26422f24d7e · outbound

This paper cites Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.026578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.763335Z digest=sha256:4dc4ae3453e763dcfda08f9332c19b3603a34a264c13b17b97d53abf7cc96018

Observation 924d9691-d87c-4949-824a-7fdcc3e8ba7a · outbound

This paper cites Detect anything 3d in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Detect anything 3d in the wild

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.826709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.826709Z digest=sha256:efe915749fec3de2df7bf1d2e308cdde420555aa0718019ce6d200a984fd0c11

Observation 4e9fc568-6ea5-47c4-b2a4-90ae0ef3af85 · outbound

This paper cites Rosefusion: random optimization for online dense recon- struction under fast camera motion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Rosefusion: random optimization for online dense recon- struction under fast camera motion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.943526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.878121Z digest=sha256:ead74cd38d1d65201ea2bfbc3d3eabe558faa6ca241209574173f1c4ab68eb18

Observation 55298ec4-c903-45d7-a8f6-00eb2c24a56b · outbound

This paper cites Asro- dio: Active subspace random optimization based depth iner- tial odometry.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Asro- dio: Active subspace random optimization based depth iner- tial odometry

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.835827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.925803Z digest=sha256:2f43ff8efaec5da88e3ad80518d8a93c2b9f61d840d11b38e8038629ec9ffc42

Observation 9ecbafac-1468-433c-84d3-d9f093d7f8cc · outbound

This paper cites Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.010201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.010201Z digest=sha256:08d783b8e1e60c46472e427fcf6c1ac4275d0109d498f15e87534f64316527eb

Observation a182b4fd-c592-49a9-9db0-a63b974bf72d · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.071247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.071247Z digest=sha256:58f60284537e765665402c8ec6cf96b827a895a85612e9df0dd381057bba53da

Observation a8379172-f2ab-4ea4-856e-f298b4209031 · outbound

This paper cites V oxelnet: End-to-end learning for point cloud based 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion V oxelnet: End-to-end learning for point cloud based 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.706522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:23.112112Z digest=sha256:4df4063101cf22c064d1fdc84e7345fbfcc83a1bbba337ea8d8db0261f341155

Observation 01a24053-0f17-464d-ac33-dffac7eb64fb · outbound

This paper cites 1, 2, 6, 7, 8.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 1, 2, 6, 7, 8

Reference 493

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.260929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:21.877808Z digest=sha256:281dd4a33522fa49e40422070f83c4ab7252a2730f0c42c3bf9b14d0bc0ed325

Observation e659c0d4-049c-4405-9e6f-930cac7aa061 · outbound

This paper cites an unresolved cited work.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Unresolved cited work

Reference 2025

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T23:57:24.976651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T23:57:22.189342Z digest=sha256:d19e57749188b2e80b2fc79a727ff3335c1bb10223ea8d0462d268887ebc02b8

Pith citing papers

No inbound Pith citation observations are available.