Pith. sign in

Paper Citation Record · LEDGER

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval

As of 22 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2607.08541.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.08541 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-10T05:41:42.555295Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact7
  • verified fuzzy7
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 70556889-5243-4bd0-b9a7-6ea04e5e479f · outbound

This paper cites YOLOv3: An Incremental Improvement.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval YOLOv3: An Incremental Improvement

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.372021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:c46a2e2f5f4009f1162faa5a60f683558cbe71c1292061306e083b974bfae133

Observation c2f38646-a600-48c4-8071-40d22107e002 · outbound

This paper cites Yolov10: Real-time end-to-end object detection.Advances in neural information processing systems, 37:107984–108011, 2024.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Yolov10: Real-time end-to-end object detection.Advances in neural information processing systems, 37:107984–108011, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.344134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:4268c84d0dd8e8e3cb723a5ecc312ec502b4b3b117ca6f8a19fba6c5dab82352

Observation 2583a07e-e757-4c97-af57-add0354006c5 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval SAM 3: Segment Anything with Concepts

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.380835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:50eb7f13cde4e51cf14e5a86f0e75245f3d0747d35165a749c24aa8a94e81ebd

Observation af2f0723-60e8-44d3-854d-0c8de1008935 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.339534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:e32b3e270e614dbc759ca11c8241492756daa2096c1cf10aa19fa5674786e7d8

Observation 6b4d0b50-11d3-4e14-bb79-ecf2686a2c74 · outbound

This paper cites T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval T-Rex2: Towards Generic Object Detection via Text-Visual Prompt Synergy

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T05:46:50.378120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:8b7f5f7f773e8e6cf7d8d23d589b1195e5b0a75ee87c90a485253f4351eb7e42

Observation 8a14c43d-d37d-4003-b40f-780a694c0f02 · outbound

This paper cites DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval DINO-X: A Unified Vision Model for Open-World Object Detection and Understanding

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-07-10T05:46:50.369655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:73e8c8d8fa3726ce0301b651956503aa27b8a95bdab11db59496e73a8bc2c881

Observation c32ae220-8ec5-4796-928a-87cf4650dda0 · outbound

This paper cites Insid3: Training-free in-context segmentation with dinov3.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Insid3: Training-free in-context segmentation with dinov3

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.341566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:fe798161b8b9b61afd2959354e2262693e2d389b3381203f0b16752a161878a5

Observation ea19ca4f-672a-454e-9bb0-572256233ba7 · outbound

This paper cites Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.383223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:8aef2985390216f914a0495c70cebf0000b59b019f1e025e3b9153a65811990c

Observation 3e1a2fad-c6d4-4523-bce3-f748641cebac · outbound

This paper cites Detect anything via next point prediction.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Detect anything via next point prediction

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.346201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:6b704bca11b11a93e0566afe0281e33d393743b3f30203b320c4017d3b347fc4

Observation 3c950d6e-caa4-4295-b9a0-cc4aa3a00e1c · outbound

This paper cites No time to train! training-free reference-based instance segmentation.arXiv preprint arXiv:2507.02798, 2025.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval No time to train! training-free reference-based instance segmentation.arXiv preprint arXiv:2507.02798, 2025

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-07-10T05:46:50.367048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:0820223f9b8df456af458c91d752098d6caff9d17f066d73cc67be469c7733e0

Observation e70a137f-65c9-4817-a58a-946517171727 · outbound

This paper cites DINOv2: Learning Robust Visual Features without Supervision.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval DINOv2: Learning Robust Visual Features without Supervision

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.364243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:e2db3f3cb4bb3c370670a451d1b97affdaf5bfbc53af002ea7a7dcaa8455af3a

Observation 9f901118-bf2b-4062-8848-ce82f7b4bb78 · outbound

This paper cites Sam 2: Segment anything in images and videos.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Sam 2: Segment anything in images and videos

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.335228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:67e2b1640a963aff6730f058378b26fbd084bc055b194eab482ee054061244fc

Observation 1691cce1-3c41-4e75-9338-f357937d72ab · outbound

This paper cites DINOv3.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval DINOv3

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.361737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:b4551a221f8d8938780ee1937a5e7f36e7be525eadbdae0a9b7588bc05a7b92f

Observation a030e303-da83-4a06-808c-f7f78949bff9 · outbound

This paper cites Modern hierarchical, agglomerative clustering algorithms.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Modern hierarchical, agglomerative clustering algorithms

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-10T05:46:50.374699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:e1f2c74303f2c7ac508cbb2f8a7b8ad919a91174265fd4a15dc60d7dc25d5c19

Observation 705a2b59-4cbe-41aa-9e9e-e8d1883643ad · outbound

This paper cites Milvus: A purpose-built vector data management system.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Milvus: A purpose-built vector data management system

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.337427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:7b71d96758a0696399c87eeac94bf8a4ab72296f0f6e65df7fe1486be717a7c5

Observation 34ba29b0-a720-409d-8700-d15524b99d01 · outbound

This paper cites Ua-detrac: A new benchmark and protocol for multi-object detection and tracking.Computer Vision and Image Understanding, 193:102907, 2020.

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval Ua-detrac: A new benchmark and protocol for multi-object detection and tracking.Computer Vision and Image Understanding, 193:102907, 2020

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-10T06:26:52.332459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-10T05:41:42.555295Z digest=sha256:a0a65953a981126fa22c46f966404eb1ceb3a89d4cfb22c03829ed97f39af1e5

Pith citing papers

No inbound Pith citation observations are available.