Pith. sign in

Paper Citation Record · LEDGER

3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2308.04352.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.04352 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:32:38.776704Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T13:38:27.437632Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 915e2c5a-d132-45a5-88a1-b290a4436b8b · inbound

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding cites this paper.

ViGiL3D: A Linguistically Diverse Dataset for 3D Visual Grounding 3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T22:32:38.776704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:32:38.776704Z digest=sha256:8894f58428f0e04e778f3f7ed871da7960df2548dd6d43fcef4c85334309094d

Observation 5d8763a3-5e9b-45ea-b66e-fe3eee26bf00 · inbound

Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning cites this paper.

Embodied Spatial Intelligence: from Implicit Scene Modeling to Spatial Reasoning 3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment

Reference 180

Resolution
verified exact
local_arxiv, observed 2026-08-05T13:38:27.556157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:38:10.684697Z digest=sha256:e51e6c20125f8a9dd82dd137fc59afbce9fc0cf46c568208c9c6ca39663edd3e

Observation d7e7e5be-d45e-46a3-8b8c-08614967c8ad · inbound

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval cites this paper.

CR-Refiner: An Object-Centric Optimal Transport Reranker for Edit-Conditioned 3D Scene Retrieval 3D-VisTA: Pre-trained Transformer for 3D Vision and Text Alignment

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-01T13:28:41.116997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:28:41.116997Z digest=sha256:620e1d6fc8577dc3d75d32df6890d6e9d29d20eb7f0349295298cc839296be39