Pith. sign in

Paper Citation Record · LEDGER

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop

As of 19 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.13363.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13363 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:53:12.973793Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy20
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 756b31d3-4b71-470c-848b-5e1f3c97582a · outbound

This paper cites Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Lang, Sourabh V ora, Venice Erin Liong, Qiang Xu, Anush Krishnan, Yu Pan, Gi- ancarlo Baldan, and Oscar Beijbom

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:16.240488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:10.707737Z digest=sha256:a71e9b115f1b3c28fbf94f5013c111b03c26121e224a21b37e3c3760d5fd2791

Observation 651cd614-9a53-44d7-b257-583c8034331e · outbound

This paper cites nuscenes: A multimodal dataset for autonomous driv- ing.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop nuscenes: A multimodal dataset for autonomous driv- ing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.924631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:10.790288Z digest=sha256:5db9d56841ca937bb3e6a30b318361552076672f59f09c156efe611c81dce50f

Observation 341c9cb0-c041-4c79-ac67-60d7606ca020 · outbound

This paper cites Glip: Grounded language-image pre-training.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Glip: Grounded language-image pre-training

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.766867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:10.869876Z digest=sha256:c49276a2def741db45261dbfbd0df5a19ce29decfc71075d57a8f725cd6c5ac5

Observation 2ba90232-de17-4fcc-a3d3-99f9c63682b4 · outbound

This paper cites Are we ready for autonomous driving? the kitti vision benchmark suite.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Are we ready for autonomous driving? the kitti vision benchmark suite

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.655536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:10.943668Z digest=sha256:3677f7d53d65cdeaab13438230d50b26d72b42faa918573a7b7f7e5552f0033b

Observation c7b83058-ef0e-47e8-9aae-fe7ef719323b · outbound

This paper cites Lvis: A dataset for large vocabulary instance segmentation.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Lvis: A dataset for large vocabulary instance segmentation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.562487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.019752Z digest=sha256:9b73181e17c2f10c6c44fc8ed5889e7144a113c0bf752488615e43cf4b1d6b04

Observation 30a69994-e77e-44ed-ac00-dbb812543359 · outbound

This paper cites Segment Anything.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Segment Anything

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:11.097615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:11.097615Z digest=sha256:b509e19b469d99f4155a0422d6f5b35c5301a0da63248cca8fce5b003233fdaf

Observation de5854e5-001b-47d3-913c-9da661c8b3cc · outbound

This paper cites Pointpillars: Fast encoders for object detection from point clouds.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Pointpillars: Fast encoders for object detection from point clouds

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.455622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.227219Z digest=sha256:fd0f666e1a68a991a6ead80d534fa0c51d16db7b7929eeb87c314ce9ffd80ced

Observation 8d5b0919-12b9-4e92-9c07-e49c1634f233 · outbound

This paper cites Clip-fo3d: Free open-vocabulary 3d object detection.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Clip-fo3d: Free open-vocabulary 3d object detection

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.351669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.304389Z digest=sha256:9775b52856358e226efd86cd7fb3e4c0feb6e8fc9e1878ce41836a6fa7af98ad

Observation 477a519a-5234-4981-8330-4f43719a18d1 · outbound

This paper cites Lawrence Zitnick.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Lawrence Zitnick

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:11.411236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:11.411236Z digest=sha256:28cce15fff9a375a90253dd22735298b4c89d06811c26a8b38bde51c06a375c5

Observation b6c26b46-2073-413f-8ddc-4a8cae9da2cc · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:11.655343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:11.655343Z digest=sha256:ca461936a19e63d93a22714bf93e7ead4ff4c4984baee1a02341c699aeb724f8

Observation fe58fcc7-f628-4440-bfae-ae9e85f96351 · outbound

This paper cites Clip2scene: Scene-level 3d open-world understanding via vision-language founda- tion models.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Clip2scene: Scene-level 3d open-world understanding via vision-language founda- tion models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.152642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.737320Z digest=sha256:a5d3b796a03fb97544366fdfca6308f38040dd08677302163c0a50d022bc268e

Observation 539f97de-c60c-477e-9b88-11918bd0f6cd · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d an- notation.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Open-vocabulary point-cloud object detection without 3d an- notation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:15.041547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.830279Z digest=sha256:54c302f534797b705e59ccc5dcfded20481fa99026e82e2d8594c77e983a5ba1

Observation 08a8471c-a351-47d6-8b84-14d3d47fb821 · outbound

This paper cites Open-vocabulary point-cloud object detection without 3d annotation.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Open-vocabulary point-cloud object detection without 3d annotation

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.929498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.924582Z digest=sha256:28b081716e51bf2eec8f43f885d203118fff8c6ed6500dfd7768a5ccb4a3ddc9

Observation af9fe8da-467c-466b-bd6e-e3a14c99f102 · outbound

This paper cites Unidepth: Universal monocular metric depth estimation.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Unidepth: Universal monocular metric depth estimation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.835280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.005130Z digest=sha256:c76801602b68465da1f3858c709c0284eae9a5bd3146fe2d3a43c35b5cb913aa

Observation 14c6e116-7669-48b2-b122-fb2ab3df6388 · outbound

This paper cites Frustum pointnets for 3d object detection from rgb- d data.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Frustum pointnets for 3d object detection from rgb- d data

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.708710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.090961Z digest=sha256:e171a0016a3797a71c9424ec4d076f1e2be9adf3b3db35ff2b07cbd444aa4600

Observation f3baef2b-99b5-4c70-8f27-c9b2cd5f15d7 · outbound

This paper cites You only look once: Unified, real-time object de- tection.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop You only look once: Unified, real-time object de- tection

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:12.171860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:12.171860Z digest=sha256:20f62b85cb76dfeb98dfd087b97f8d1cdc3027d6e663f5b965e982ec73ea4f50

Observation 24ab1c56-722d-429e-8729-c159eb926b7e · outbound

This paper cites Faster r-cnn: Towards real-time object detection with region proposal networks.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Faster r-cnn: Towards real-time object detection with region proposal networks

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.605512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.269015Z digest=sha256:8c2b99dc2a7a52868b82efc15fc5be653df9a8952686d75244df10c795ef8047

Observation addb06c4-4bc4-4e50-b1c0-bead2b9f0df8 · outbound

This paper cites Scalability in perception for autonomous driving: Waymo open dataset.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Scalability in perception for autonomous driving: Waymo open dataset

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.422435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.350142Z digest=sha256:8bcdeb5c474c9108006bc503c30c112e3ac911e0056b818826925119f9e71e31

Observation a7ee932b-59e2-4b8e-9a25-66ab4caaa0d3 · outbound

This paper cites Fsd: Few-shot object detection in 3d scenes.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Fsd: Few-shot object detection in 3d scenes

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.169948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.441833Z digest=sha256:d98e4fb15d3b62e98217eb34e577cb478d6d43afd6c6a7be61f2f833b28ce8c8

Observation e4e9ebd6-b558-4430-94c1-6636f4c4d062 · outbound

This paper cites 3D for Free: Crossmodal Transfer Learning using HD Maps.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop 3D for Free: Crossmodal Transfer Learning using HD Maps

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:53:12.548365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:53:12.548365Z digest=sha256:0f81add7c04f9620f62e5c435ab651e09fa03a02b81baca55f9ce690a513a01d

Observation 8f86db43-4429-4499-bd31-f4ab68a64794 · outbound

This paper cites Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Posecnn: A convolutional neural network for 6d object pose estimation in cluttered scenes

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:14.002820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.613800Z digest=sha256:012dcf10613a5b48da70a9de181eeb237c027af5facda4ace74a7170b991d1b7

Observation 5466189b-1c80-493c-95d8-1aedf36084a0 · outbound

This paper cites Ulip: Learning unified representation of language, image and point cloud for 3d understanding.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Ulip: Learning unified representation of language, image and point cloud for 3d understanding

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:13.826577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.727849Z digest=sha256:8bdbbd27c505ac2ea8dcc7509b098675bd3858ecde18606250017044e2364b7e

Observation 68912b25-2825-4e79-8201-c26f5eb6fceb · outbound

This paper cites Open-vocabulary object de- tection using captions.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Open-vocabulary object de- tection using captions

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:13.627903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.797385Z digest=sha256:06956815be4fe2943ea196e1942f64ccf082fdba5bbfba3e3e8f0c8db95450d8

Observation 1c65b2cf-3b03-4a79-a162-374e891afec3 · outbound

This paper cites Pointclip: Point cloud understanding by clip.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Pointclip: Point cloud understanding by clip

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:13.430144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.899553Z digest=sha256:743337209dc3c46a4e4a3c6f07ed62b3677892afdbd63adc93d5edcb79a9bcf0

Observation 24ae7b24-1551-450c-9d73-6fd0a2e46731 · outbound

This paper cites Regionclip: Region- based language-image pretraining.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Regionclip: Region- based language-image pretraining

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:53:13.251207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:12.973793Z digest=sha256:d93b9506f72b019af3acd05bdbed1ec6b2706f1b5f093842af13b4dfc7d3cb19

Observation a1f7c993-07b9-46ab-b21d-bd7f69b8a49f · outbound

This paper cites an unresolved cited work.

Just Add Geometry: Gradient-Free Open-Vocabulary 3D Detection Without Human-in-the-Loop Unresolved cited work

Reference 2014

Resolution
unresolved
raw_fallback, observed 2026-08-06T19:53:15.258223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-06T19:53:11.514722Z digest=sha256:586f687531538531ad4b14ecd6f61e58fa468ded1893f20ecef197f91afeeaf9

Pith citing papers

No inbound Pith citation observations are available.