Pith. sign in

Paper Citation Record · LEDGER

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion

As of 8 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2506.15610.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15610 v3

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:57:23.112112Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact2
  • verified fuzzy38
  • unresolved16
  • parse uncertain1
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1810e1c3-0f9a-48e4-825b-965a552770ad · outbound

This paper cites Omni3d: A large benchmark and model for 3d object detection in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Omni3d: A large benchmark and model for 3d object detection in the wild

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:29.098750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:19.710060Z digest=sha256:aa7e3d418fe2eef717f2982c1561f9544b0b9abb874d73124b0bf76985a31a95

Observation 2c0833e2-f403-4540-b463-5e4af2c0c3db · outbound

This paper cites Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Coda: Col- laborative novel box discovery and cross-modal alignment for open-vocabulary 3d object detection

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.953346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:19.773113Z digest=sha256:ca93782f9d0d38dff26084057e7ed11c241968b8c02d9ff6e9ca364b50304e36

Observation 0b8c31d0-66db-44f9-9ef4-7319f9e8b31b · outbound

This paper cites CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CogNav: Cognitive Process Modeling for Object Goal Navigation with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:19.841794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:19.841794Z digest=sha256:9536bd368fe942a04deabcd6933931a77c020c106df3bef942ff2ae96172df47

Observation 388626aa-3df2-4a49-9063-f2f5c4043019 · outbound

This paper cites Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Collabo- rative novel object discovery and box-guided cross-modal alignment for open-vocabulary 3d object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.757700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:19.882914Z digest=sha256:760fc427de1560b48de0a481276e99d9356e13deb7320fc3af7447aeb76ad62c

Observation 74be73ef-6495-406f-9033-f653c0842a43 · outbound

This paper cites A hierarchical graph network for 3d object detection on point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion A hierarchical graph network for 3d object detection on point clouds

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.627155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:19.973335Z digest=sha256:07f598b5b2ac76e0779673565d9e91d640b565e51477cd362a7746921412dea9

Observation 075ef2eb-c0fd-4efd-ad99-59efe427d5b2 · outbound

This paper cites Scannet: Richly-annotated 3d reconstructions of indoor scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Scannet: Richly-annotated 3d reconstructions of indoor scenes

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.052396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.052396Z digest=sha256:f8b577df47ebdef4dd19a055e07ddbe366933b90e53b83d8bf2a7d1c171b4a74

Observation 38ade343-5ea4-4662-bfd7-b99dbd518673 · outbound

This paper cites Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Graph-to-3d: End-to-end generation and ma- nipulation of 3d scenes using scene graphs

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.354338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.088537Z digest=sha256:e46ed473c56b55d91d02db7a922187931b8c6e058f9374075fe02bd553a2e98b

Observation d1050601-59d1-432a-8b23-85b38cba320d · outbound

This paper cites Disarm: Displacement aware relation mod- ule for 3d detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Disarm: Displacement aware relation mod- ule for 3d detection

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:28.112593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.164335Z digest=sha256:fde5d8f79052fabccefb8729e46bdfeaae8a2ac8a7a456475e795c615585c39e

Observation d8b0e6e4-1366-422a-ad17-2a5387f33182 · outbound

This paper cites 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 3d-mpa: Multi-proposal ag- gregation for 3d semantic instance segmentation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.230116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.230116Z digest=sha256:148a57e5c82d1e66e056bc1e106d84dd7135b2ad33fe14db68c5da110a8b4493

Observation 31c53f9e-0276-4e75-b58f-926c262a9e92 · outbound

This paper cites Generic objects as pose probes for few- shot view synthesis.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Generic objects as pose probes for few- shot view synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.266963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.266963Z digest=sha256:05de6ecd90c38050ef36e4b13679d14b42e3e5528e3a9e54e71f8be87206093d

Observation 0d60c4cf-3e4f-4aa1-be89-2bee002f3ca6 · outbound

This paper cites Training an open-vocabulary monocular 3d detection model without 3d data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Training an open-vocabulary monocular 3d detection model without 3d data

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.903048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.342896Z digest=sha256:886ac5920e27d859e68c5a26c804175c837699dbb33408fc91c7e3859616c800

Observation f90b80f4-833b-4283-954a-ebe59cff9416 · outbound

This paper cites Particle filter with swarm move for optimiza- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Particle filter with swarm move for optimiza- tion

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.648559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.428794Z digest=sha256:db8818e2f079ff24d9da28241aec0753793e532b6123e1270b1a908acd58b15d

Observation deac8d6b-e5d5-4cd7-bbc0-6472fae19635 · outbound

This paper cites Open-vocabulary 3d semantic segmentation with foundation models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary 3d semantic segmentation with foundation models

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.458507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.468149Z digest=sha256:0e91f43c1821faa0266444fd92390dc1fa0f96b26f3e8da70cb3cbfb8573d757

Observation c84d6efc-034c-45c1-9efa-fc92731b8523 · outbound

This paper cites Segment any- thing.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Segment any- thing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.519699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.519699Z digest=sha256:95163de1b996add74cc4720bfb824ce86ff0a325cff1d361f223c50f22018d64

Observation b7fbde06-5851-4147-b246-6cbd89a7e6b2 · outbound

This paper cites Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pycuda and pyopencl: A scripting-based approach to gpu run-time code generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.208995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.613522Z digest=sha256:5697b6c64e9e206d4ed5f24f0119a813b973aa1818077dd683533c3f9a86a734

Observation d59def1e-fea6-4475-af13-29d4b993c969 · outbound

This paper cites Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open3dsg: Open- vocabulary 3d scene graphs from point clouds with queryable objects and open-set relationships

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:27.039717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.681004Z digest=sha256:bb8b9677d67db71959644a6375f85ca0c23062c95f2c4b35dac972d7006ae3d1

Observation 3227a5b0-11eb-43fa-9c32-8ba76137b8f6 · outbound

This paper cites Arm3d: Attention-based re- lation module for indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Arm3d: Attention-based re- lation module for indoor 3d object detection

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.883501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.729580Z digest=sha256:4404aaf023e7fa340887e328090a481fed8bdbd242304303745114e8f056752d

Observation dee36243-c969-4f07-99cd-a68810ab5ec5 · outbound

This paper cites Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Remixfusion: Residual-based mixed representation for large-scale online rgb-d reconstruction

Reference 18

Resolution
verified exact
raw_fallback, observed 2026-08-06T23:57:23.622737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.762678Z digest=sha256:fcd611b25c4555f5f41524ccfc3cfc46f0f3870c4213d6cf46f50d17564b7084

Observation c3d0f184-e88b-462e-b7ed-749b937e064d · outbound

This paper cites Cubify Anything: Scaling Indoor 3D Object Detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Cubify Anything: Scaling Indoor 3D Object Detection

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:20.836056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:20.836056Z digest=sha256:553ca8a7008a2a7f69b23ec06afe1a85f6278764ccdb25c665ada9f7b8937321

Observation ec19e30d-dcdb-4eec-bffa-9f5e596b0550 · outbound

This paper cites Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Patch- work++: Fast and robust ground segmentation solving par- tial under-segmentation using 3D point cloud

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.706827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.923788Z digest=sha256:f0081dc5009a3f2fb96bc97c4048bcd43dcaccf866a03c5b7ee07b27a0aca7ba

Observation 910dc8e6-aeb9-4fc5-a6a5-3799fbe6260d · outbound

This paper cites Ground- ing image matching in 3d with mast3r.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ground- ing image matching in 3d with mast3r

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.588898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:20.959777Z digest=sha256:663499eaacff64c4edc5a9210f82acb1734c5b6fe90445165f897e5057c880d8

Observation 59b8dedd-56c5-4851-8b47-19d2c55907c0 · outbound

This paper cites Grass: Generative recursive autoencoders for shape structures.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grass: Generative recursive autoencoders for shape structures

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.435530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.029161Z digest=sha256:98aaac51e753220e4f3d3d7de79f01612e5226bc2e1405d17d286e45297ecf1b

Observation 54325d4d-4b34-4b92-a097-d8f71b7b38e0 · outbound

This paper cites Prompting depth anything for 4k resolution accurate metric depth estimation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Prompting depth anything for 4k resolution accurate metric depth estimation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.106713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.106713Z digest=sha256:ba57b4493dd0daa7e52448dc5a798812f6fd50bacdc18f7304a5333e7ea42129

Observation 9575c0d0-f9d0-40b1-b110-fd3ae562ec36 · outbound

This paper cites Microsoft coco: Common objects in context.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Microsoft coco: Common objects in context

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.156947Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.156947Z digest=sha256:1082fb0b210995172c75c60f61d37e5ec255113d2c98c51f9a4e8750fe3e419b

Observation 780674c6-9d6d-4cff-8789-faf511f73e8a · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.283502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.186493Z digest=sha256:5fbc36b63fce53718bcad3e04f4513d9e97c04b8e12a3b8f6a5d31e97c889ad7

Observation dd589e9a-127b-47bc-907e-1b2799c5740b · outbound

This paper cites Group-free 3d object detection via transformers.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Group-free 3d object detection via transformers

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.161915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.272610Z digest=sha256:1aa30bc469de20e6c3c293f99765f94cd7baa1c5530ad6427c5e0968f35e0cf8

Observation 91b030f3-9a99-4eec-a10d-918adbd6017f · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Open-vocabulary point-cloud object detection without 3d an- notation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:26.006363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.345759Z digest=sha256:a4914071af23c94a381d938d46a38188bbffcce41ce448971e865582d721db66

Observation 2a316920-c355-4120-94c4-cefac417f91f · outbound

This paper cites Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Oa-cnns: Omni- adaptive sparse cnns for 3d semantic segmentation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.853536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.395450Z digest=sha256:88e7f4dede84b3467854aeabf6e04abc244b767e7ba2ca71f48f9ee465306133

Observation a546f684-a63e-49ea-95b2-b8102b33067a · outbound

This paper cites Pointnet: Deep learning on point sets for 3d classification and segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Pointnet: Deep learning on point sets for 3d classification and segmentation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.482945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.482945Z digest=sha256:03a098a533470ace657f91f69885eee533afdab2ff3d9e8e68081d9afae9ace3

Observation 4839aa72-f402-41b5-ae08-ff2ffea84526 · outbound

This paper cites Deep hough voting for 3d object detection in point clouds.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Deep hough voting for 3d object detection in point clouds

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.752058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.554743Z digest=sha256:32a82b3f84e374fcca4e464b628221c527ff6e81fd253bcaa8d4a341334a587a

Observation b82ca3dc-b805-459e-a51a-130f286f6df0 · outbound

This paper cites Imvotenet: Boosting 3d object detection in point clouds with image votes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Imvotenet: Boosting 3d object detection in point clouds with image votes

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.631090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.596405Z digest=sha256:c9523ac919f06c3837ef0e1ad8bb5f91b025e66da1352a9e6b482e63097a469e

Observation ef7a074c-2554-4528-ad20-99f58a209b56 · outbound

This paper cites High quality entity segmentation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion High quality entity segmentation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.480306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.698243Z digest=sha256:90963619c3f2fe1c0ea56c39132d2e39c25d1d018d59e785dd058cb4a32d3aa1

Observation 6b2d430e-9b97-48aa-96ac-120d0f782967 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Learning transferable visual models from natural language supervi- sion

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.761942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.761942Z digest=sha256:a511219aa1fb96960ea5cbb2893a50f5c05e4accfe48a1e499346e0b451d78d3

Observation df1d47f0-249e-4836-be9c-390e09e8170c · outbound

This paper cites Fcaf3d: Fully convolutional anchor-free 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Fcaf3d: Fully convolutional anchor-free 3d object detection

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.372401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.814853Z digest=sha256:bf604ec57f6a57377271400dcc2423752b9543894c5647d9c847284e80b54405

Observation 1c1b4d4e-87f8-4e03-8a56-88ab71b2c584 · outbound

This paper cites Tr3d: Towards real-time indoor 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Tr3d: Towards real-time indoor 3d object detection

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.175588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.930156Z digest=sha256:4507e349f13b4f390f61a76b7bf9a45e05c2679c10630a2822ba13dca831a0a6

Observation f2a6349b-9088-4ba2-8369-76754884eeb7 · outbound

This paper cites CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion CLIP-Fields: Weakly Supervised Semantic Fields for Robotic Memory

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:21.984804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:21.984804Z digest=sha256:4e4bdb43f44dc862af0b9c1739f88fb2703f74deb0fe9e116c7b29b248eafc19

Observation 71bfd0e5-8d45-4aee-b09c-70204a46798a · outbound

This paper cites Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mips-fusion: Multi-implicit-submaps for scalable and robust online neural rgb-d reconstruction.ACM Transactions on Graphics (TOG), 42(6):1–16, 2023

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.058640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.095652Z digest=sha256:25e93cc04de1481031cedcfd5f246d42cb7eb5a1ad230814d6f33961e878289d

Observation afcaf3b6-a840-4370-ad8d-9f1204a89f5b · outbound

This paper cites OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion OnlineAnySeg: Online Zero-Shot 3D Segmentation by Visual Foundation Model Guided 2D Mask Merging

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:57:23.380699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.137587Z digest=sha256:c4b94be2d304d469897d4be0ccabcc5b8ecf9c2e03daa35da96aa4aed406675c

Observation b51d861a-0292-4570-92da-364bf9ed4fc3 · outbound

This paper cites Spatiallm: Large language model for spatial understanding.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Spatiallm: Large language model for spatial understanding

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.873173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.277083Z digest=sha256:268a5ec6fa4af37f4c01c5e293ff16eb5e874d77a46c674ced517251804ca3c2

Observation 80edae4e-0db0-48cd-ac43-a86dd91cdb9f · outbound

This paper cites Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Appa- 3d: an autonomous 3d path planning algorithm for uavs in unknown complex environments

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.755358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.318089Z digest=sha256:9d57761466329c0e402b5adf3348837fb3e16238f163f2ee7e0182fa394284d4

Observation 2c411c6a-c3fb-4ff7-ad39-c8b88103abb5 · outbound

This paper cites Dust3r: Geometric 3d vi- sion made easy.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Dust3r: Geometric 3d vi- sion made easy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.649214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.370207Z digest=sha256:ad380ded00064b1b5e0ef48379aa18bf509917bed8aa4346d475a49538b58b2c

Observation f1c669fc-2e85-4f4a-9ec6-ca30275c7c6c · outbound

This paper cites Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Ov-uni3detr: Towards unified open- vocabulary 3d object detection via cycle-modality propaga- tion

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.553960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.452928Z digest=sha256:1a53c5a3c46e3edffd451a31e5b32633520602ccb063051889c24573aecd8712

Observation 5a1a0568-8def-4c35-b0eb-3b748886b4fb · outbound

This paper cites Mlcvnet: Multi-level con- text votenet for 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Mlcvnet: Multi-level con- text votenet for 3d object detection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.451083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.497715Z digest=sha256:7903d9e591af7406598237cae50b7e1fe564b6eb3f248528422614c10e2447fa

Observation 1ac36b7a-40cb-4bff-b7a9-79acb6ce0495 · outbound

This paper cites EmbodiedSAM: Online Segment Any 3D Thing in Real Time.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion EmbodiedSAM: Online Segment Any 3D Thing in Real Time

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.527722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.527722Z digest=sha256:460d02d985c5d9c927352d8bcf715287454c845098f784956ebbedbabbd25896

Observation ff68512c-d9b6-40c6-b356-a34588984168 · outbound

This paper cites Memory-based adapters for online 3d scene perception.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Memory-based adapters for online 3d scene perception

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.343165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.557788Z digest=sha256:3dff702f836c7f0cca5bc1b834b38a4156d30a53213a04bc39eddf4f876fccca

Observation 6cf57c01-6a2e-455c-8279-9dd6305f69c8 · outbound

This paper cites M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion M 2 diffuser: Diffusion-based trajectory optimization for mobile manipulation in 3d scenes

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.240521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.628208Z digest=sha256:9fe866e6ff0886f4e3027018f77e8ae7425cd7292c9e062c9b7215f975619370

Observation 74bc0535-9c10-4fa6-9e84-dc79c263eb7d · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.146333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.685438Z digest=sha256:705ca3fece802e60bfa597549bef4c28e80b459875459670f5e7cfadf462a983

Observation 015464ae-5d32-4ab9-9b2c-29a2ac39231c · outbound

This paper cites SAM3D: Segment Anything in 3D Scenes.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion SAM3D: Segment Anything in 3D Scenes

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.728460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.728460Z digest=sha256:411d3440c951a71b71a20f699bd0b5a23a8e8e28a8fb9a129e5465a70b78e177

Observation 552dd05e-4ae1-4825-bcb0-e26422f24d7e · outbound

This paper cites Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Sg-nav: Online 3d scene graph prompting for llm-based zero-shot object navigation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:24.026578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.763335Z digest=sha256:33740d9a8ebf8e43ace13c5df842e4504971a346927c42aba1a6add5d10ec4a1

Observation 924d9691-d87c-4949-824a-7fdcc3e8ba7a · outbound

This paper cites Detect anything 3d in the wild.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Detect anything 3d in the wild

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:22.826709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:22.826709Z digest=sha256:e2a82a874773bb7e690be412f74db175e55d141c4bf4a3a614a6ac8b0fb49ddc

Observation 4e9fc568-6ea5-47c4-b2a4-90ae0ef3af85 · outbound

This paper cites Rosefusion: random optimization for online dense recon- struction under fast camera motion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Rosefusion: random optimization for online dense recon- struction under fast camera motion

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.943526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.878121Z digest=sha256:c8641342f4d17b34dcf7b5fe7294b67456a97cbd9f3fd85313765e988a21fb56

Observation 55298ec4-c903-45d7-a8f6-00eb2c24a56b · outbound

This paper cites Asro- dio: Active subspace random optimization based depth iner- tial odometry.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Asro- dio: Active subspace random optimization based depth iner- tial odometry

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.835827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.925803Z digest=sha256:962465e50b51e9261fc32099ddb9eb6183c59b2ebb4c8e24d02cdea17937fc09

Observation 9ecbafac-1468-433c-84d3-d9f093d7f8cc · outbound

This paper cites Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Gamma: Graspability-aware mobile manipulation policy learning based on online grasping pose fusion

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.010201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.010201Z digest=sha256:02fbb1f55775e1b94601b4de018ba43a7c2b98089ea2d54e0f18057b2d5de708

Observation a182b4fd-c592-49a9-9db0-a63b974bf72d · outbound

This paper cites Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Navgpt: Explicit reasoning in vision-and-language navigation with large lan- guage models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T23:57:23.071247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:57:23.071247Z digest=sha256:55c831634569ea333b52af870873b4f1d05823827dd6937bd2f9d12a57d3e806

Observation a8379172-f2ab-4ea4-856e-f298b4209031 · outbound

This paper cites V oxelnet: End-to-end learning for point cloud based 3d object detection.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion V oxelnet: End-to-end learning for point cloud based 3d object detection

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:23.706522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:23.112112Z digest=sha256:2633760a92313e280bfa0813dd164041100fd38be38dd61050dbf66046570384

Observation 01a24053-0f17-464d-ac33-dffac7eb64fb · outbound

This paper cites 1, 2, 6, 7, 8.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion 1, 2, 6, 7, 8

Reference 493

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:57:25.260929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:21.877808Z digest=sha256:431e8c481882cc4497be24af6c65bf4da71a43384bf475efb50b48b7fef4a749

Observation e659c0d4-049c-4405-9e6f-930cac7aa061 · outbound

This paper cites an unresolved cited work.

BoxFusion: Reconstruction-Free Open-Vocabulary 3D Object Detection via Real-Time Multi-View Box Fusion Unresolved cited work

Reference 2025

Resolution
parse uncertain
raw_fallback, observed 2026-08-06T23:57:24.976651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T23:57:22.189342Z digest=sha256:47b4bc8970bc8c3c18e34a2d6e75e12065fe4d6126352c0330a6c0a3a3e04987

Pith citing papers

No inbound Pith citation observations are available.