Pith. sign in

Paper Citation Record · LEDGER

VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2410.13860.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13860 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:12:19.231195Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:59:20.317233Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation edcfabd1-07f9-4cd5-b156-2823a5aa0bc2 · inbound

Zero-Shot 3D Visual Grounding from Vision-Language Models cites this paper.

Zero-Shot 3D Visual Grounding from Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:19.231195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:19.231195Z digest=sha256:51bee7983dc91c0fd3358f83596b533b609af9b959934a3a5d33d2f759bea163

Observation 6b33626f-3e6d-49ac-a012-b6b4d3b5ab12 · inbound

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models cites this paper.

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:08.430220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:08.430220Z digest=sha256:809d6f92aa3e543c332f3a7c357b12bc313768d41b5caf5d936f4ff1c26d1f4b

Observation e338d33b-6e96-42b4-afde-065288427dff · inbound

SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding cites this paper.

SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T14:52:56.561970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:52:56.561970Z digest=sha256:c484b2c4fbc83a61a95cd802c7fc2c80e9f4281e73da457b753d7987129028bf

Observation f4b65a09-a4b9-479e-9008-e80826c52478 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:05:10.065051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T13:04:30.544504Z digest=sha256:892bec0f6053c9ee8484d7d2a65caa9f2e6e9d3c970425d6aaff0338b74766d9

Observation 3354e406-cad3-44e4-a168-aa6be5881022 · inbound

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies cites this paper.

PoseVLA: Universal Pose Pretraining for Generalizable Vision-Language-Action Policies VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-02T21:37:15.172508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:37:15.172508Z digest=sha256:b618b34781fa488e38ea35c6e8e6613a592ffc1d888e4090c23021ef59b5dd60

Observation 85c84e12-c257-4e9a-9a62-ff8967061ac8 · inbound

QuadAgent: A Responsive Agent System for Vision-Language Guided Quadrotor Agile Flight cites this paper.

QuadAgent: A Responsive Agent System for Vision-Language Guided Quadrotor Agile Flight VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:18:13.144096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:16:36.843013Z digest=sha256:43906b7c2130a3b43f1027cd445873f312dad5374ef6778f617b1082cf6a79a6

Observation a173309e-a90b-4a2b-81a9-2be06ae79344 · inbound

Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding cites this paper.

Multiple Consistent 2D-3D Mappings for Robust Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:46:26.479613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T13:48:28.695094Z digest=sha256:a56ef08639d09be5d9e133aff579f247744879316afb89dffbc5947ee17c9586

Observation a42499ad-265c-4a59-a797-7260600de9e4 · inbound

SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching cites this paper.

SceneGraphGrounder: Zero-Shot 3D Visual Grounding via Structured Scene Graph Matching VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:51:18.439609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T08:47:25.712575Z digest=sha256:c48a901f49ac82753c4ac9a5dffc12d0b1136d36286625fa3069ec1a5fd9fa04

Observation 2d930105-f2a8-4667-b03d-01ac17814941 · inbound

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning cites this paper.

Beyond 3D VQAs: Injecting 3D Spatial Priors into Vision-Language Models for Enhanced Geometric Reasoning VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:53:13.672837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T07:47:52.739735Z digest=sha256:8b5ef7f67bb6f647ad97ac56bc9a932681716eee4407c7a09b1ea7a7b00c5bf5

Observation c656486a-9678-4d86-8d69-55b8f97fdd94 · inbound

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning cites this paper.

ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:59:20.320423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T20:49:01.269513Z digest=sha256:bd9154c991e078e3734670445ebec511b2e9aee7f906c01f8f931cfdf04278d2

Observation 75455b46-ea1f-4277-91c1-37e3a05e8699 · inbound

G$^2$TAM: Geometry Grounded Track Anything Model cites this paper.

G$^2$TAM: Geometry Grounded Track Anything Model VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T23:56:52.009530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T23:56:52.009530Z digest=sha256:83553405cd6102a3f541b279ffe4c99f6e8864729be1cdd884153d8500da633e

Observation 3adfee7b-12a5-4b11-9ef3-89a8da7c8d6c · inbound

TDVR: Joint Text Disambiguation and Viewpoint Reasoning for Zero-Shot 3D Visual Grounding cites this paper.

TDVR: Joint Text Disambiguation and Viewpoint Reasoning for Zero-Shot 3D Visual Grounding VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T13:09:12.444281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:09:12.444281Z digest=sha256:1f72084e404bbb4b3cafea95ee324886cfa9955e49b0ce7020526ad9fc3fb445

Observation 174dfef7-f6eb-4c34-9be9-a3e11ae3c7dd · inbound

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching cites this paper.

Talk2Sensors: 3D Visual Grounding in Autonomous Driving via Sensor-Adaptive Physical Cue Matching VLM-Grounder: A VLM Agent for Zero-Shot 3D Visual Grounding

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T21:57:45.060750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:57:45.060750Z digest=sha256:dc6e42d5e5055613d169cb369049d356652c78a3cd451e74783ef4b184399065