Pith. sign in

Paper Citation Record · LEDGER

An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2401.02361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.02361 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:15.950528Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T10:27:02.358494Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4b86998f-e0b9-4839-8036-bb0f69105d49 · inbound

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification cites this paper.

SemSegBench & DetecBench: Benchmarking Reliability and Generalization Beyond Classification An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:15.950528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:15.950528Z digest=sha256:1dfc43ebe585545fada03ed160f7e141334d367798990bcdf42fb95c280bae10

Observation 1efe85d8-c42e-4f8a-a403-3af50ae9eedf · inbound

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models cites this paper.

DINO-R1: Incentivizing Reasoning Capability in Vision Foundation Models An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:43:41.692788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:43:41.692788Z digest=sha256:e4dd2f98cb176f09a851b9ea298dbca126b6f2bbed1fb43d8cc215e5aa634bf8

Observation 45e8d754-781f-409c-999e-6debc3a739d7 · inbound

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection cites this paper.

3D-MOOD: Lifting 2D to 3D for Monocular Open-Set Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T10:42:01.065564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:42:01.065564Z digest=sha256:b5f012dadc1e204e005f5b3021f62f00f862567ce69dd50067363ac3f34dbe03

Observation e99a1f33-3852-4387-84a7-c1421d2a9899 · inbound

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model cites this paper.

FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T10:18:44.680887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:18:44.680887Z digest=sha256:ca62c704680efdd26a0c8923381e9f2637d67ff7ed40e4faa09bab47ca1f0756

Observation 5c190997-ec5f-41ba-b72f-d5d5304883e6 · inbound

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding cites this paper.

Detector-Empowered Video Large Language Model for Efficient Spatio-Temporal Grounding An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:58:46.594864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T00:54:53.789523Z digest=sha256:59a03967827e2cb445dd19e47846ae2012512cb836e46422a93ea24428d49ad8

Observation e88c9d94-71db-4dde-bb93-b8fa46b96bbd · inbound

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework cites this paper.

Focus on What Really Matters in Low-Altitude Governance: A Management-Centric Multi-Modal Benchmark with Implicitly Coordinated Vision-Language Reasoning Framework An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-16T11:07:48.345220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T11:03:37.176829Z digest=sha256:f70230f0e43389174dbb85d759b20624f8593733ae617d883da7033a7410244e

Observation 7f335ee2-8af0-41d3-bc57-160cecb213a3 · inbound

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training cites this paper.

PET-DINO: Unifying Visual Cues into Grounding DINO with Prompt-Enriched Training An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-13T22:23:21.416722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T22:19:21.723963Z digest=sha256:7f9715a785eef05badd540a0fc0dd7d6cbb05afde7011c7741213455cbf20bd1

Observation 540a292e-b921-46dc-a206-39bcb5603324 · inbound

Training a Student Expert via Semi-Supervised Foundation Model Distillation cites this paper.

Training a Student Expert via Semi-Supervised Foundation Model Distillation An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:53:00.102253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T16:50:57.376622Z digest=sha256:74f16b432e5acf0bf6f9b0ed854a26d84f58842ee9557d35257cc57e07127045

Observation 209b0ead-cc99-482c-8f98-a1d7eaf71716 · inbound

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval cites this paper.

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:50.534296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:42:23.224746Z digest=sha256:af36fc744bae174705c0f39a82836934ef60714b5f5b5aba758a75ba0c94c3b3

Observation 463a44b0-acb2-4e75-b519-aeb67b593d8c · inbound

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results cites this paper.

The Second Challenge on Cross-Domain Few-Shot Object Detection at NTIRE 2026: Methods and Results An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 127

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:21:00.707452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T16:05:16.070144Z digest=sha256:985596bfc1f3cdc32849c69ad7a020a938144865c0dd79202f7a9681dab73c5d

Observation 2f0e812a-c15e-440b-9b25-eca914c50af2 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T12:10:21.951732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T12:07:21.203513Z digest=sha256:eb85d1fab4a7ae4d84b9e0eae9a2ffca57e673fab9da17bea3ca23b9185638e6

Observation f9b45074-bd74-46d8-9c0b-962acfd60426 · inbound

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts cites this paper.

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T20:05:20.702092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:05:20.702092Z digest=sha256:35ccf14959bf11ed3074213fc191573205cfb0fbd78c15365fac30e8c95ad97e

Observation 1cd0ed13-1760-4394-8ae9-705145babb7e · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:30.276802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T01:24:56.236650Z digest=sha256:acb3fbcec2ba9453ef61ec1b9f5e56b97e385df77a2fc705af375af860cbc689

Observation 9d01a686-8cd3-4429-8355-412294974194 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:00:35.797463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T19:14:12.092478Z digest=sha256:95399933b8c4450eef2132754ddbf7b0a5f1c30e80319c79a2e1317dc5a6d28a

Observation baa1efb4-1ade-4b71-a36d-6a1bf977a9e0 · inbound

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection cites this paper.

VL-SAM-v3: Memory-Guided Visual Priors for Open-World Object Detection An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:48.454982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T01:33:55.317107Z digest=sha256:4a7e31327bd5cd9f765b0f194fec029ec39aa589da8f53a06e91ae5e6aaad6f8

Observation 61f9333f-e049-465d-bad7-866ed712caeb · inbound

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer cites this paper.

DetRefiner: Model-Agnostic Detection Refinement with Feature Fusion Transformer An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:26:18.879905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:25:49.511875Z digest=sha256:6583be5374b0acd35360727e6b7da164988ec03b6c46776a71c515bc35761038

Observation 9ddef33d-510d-4ed1-9cb8-6cc45a0324a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 69

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T12:59:53.215985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T05:33:25.373918Z digest=sha256:1cece17d4c58281abfae2a43b853287d62a5214211274c5369995178b19bdfae

Observation 439bc8e2-3a67-4be8-8230-69ae5c0e53a5 · inbound

Robust Onion: Peeling Open Vocab Object Detectors Under Noise cites this paper.

Robust Onion: Peeling Open Vocab Object Detectors Under Noise An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:15:49.989638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T00:42:36.605869Z digest=sha256:790641abe185c03dfa790acaeefc40caa04ad56e6cdea885688af333fdc501a8

Observation 509f46f5-f5b3-43df-bde3-c13a696c041e · inbound

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images cites this paper.

DroneFINE: Domain-Aware Parameter-Efficient Fine-Tuning of Vision-Language Detectors for Drone Images An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:27:04.695200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T15:23:42.245556Z digest=sha256:8b13498bedfcbaeb65da15c9b6d27d235f7dc65e15497606246af7c392afa37f

Observation 5fc73722-0554-4add-84c8-b5d95d9b9f21 · inbound

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery cites this paper.

UniRef-UAV: A Multimodal Benchmark for Universal Referring in UAV Imagery An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-10T10:27:02.359976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-10T10:22:14.308951Z digest=sha256:9f9dabd39b8ba996e84adcd24b58c30499f908b50885cc9d3d9f8ab8583dfb48

Observation c4342e8e-9501-4f06-ab95-cf5278aac247 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO An Open and Comprehensive Pipeline for Unified Object Grounding and Detection

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:51.853525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:51.853525Z digest=sha256:d819ca3a4bdf336c6ffefa26a792516f19c0e9476a11c256dadcfdad55fa48a7