Pith. sign in

Paper Citation Record · LEDGER

Training Vision Transformers for Image Retrieval

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2102.05644.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.05644 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:50:11.550843Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T14:04:51.544148Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2abb3203-1345-484a-8aa4-82127b40de8a · inbound

Emerging Properties in Self-Supervised Vision Transformers cites this paper.

Emerging Properties in Self-Supervised Vision Transformers Training Vision Transformers for Image Retrieval

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:04:51.546508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-16T14:04:51.458382Z digest=sha256:44874118d7bc7160cb726f71a7a394d7da6580eab82c2f4f63fd1681353ff20a

Observation 35492412-cdff-44a6-a619-efcafa076a9a · inbound

Vision Transformers Need Registers cites this paper.

Vision Transformers Need Registers Training Vision Transformers for Image Retrieval

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:41:38.139493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-13T09:41:37.937046Z digest=sha256:c15cadf41945c82f8b72651262b792752cecaf236e1d59d9399e8644096671be

Observation 87601ccd-7095-4e79-ad9a-2244a30a7abe · inbound

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data cites this paper.

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data Training Vision Transformers for Image Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:50:11.550843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:50:11.550843Z digest=sha256:2f4619b1c9e6ca8cf5da7d3517a04a75453e160957bf72e3fda30f4c2e75405e

Observation 07d81f0f-2041-4108-b269-6725a2a47e8a · inbound

Triplet Synthesis For Enhancing Composed Image Retrieval via Counterfactual Image Generation cites this paper.

Triplet Synthesis For Enhancing Composed Image Retrieval via Counterfactual Image Generation Training Vision Transformers for Image Retrieval

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T16:57:40.147683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:57:40.147683Z digest=sha256:e8fc126d8fa16324b22a2b4d32d1d166739fdfa0f8906538486f8df300bfb9aa

Observation 1ad6a074-8bef-4e57-8474-68e236ba1e03 · inbound

EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations cites this paper.

EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations Training Vision Transformers for Image Retrieval

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T14:55:29.494550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:55:29.494550Z digest=sha256:e323701a6e0613af43f5cd1c0d1b166ab0d91cc21bbd59b2f3cd73e8b5171535

Observation e2a50346-eafe-4b05-a04e-e3a3da02de7a · inbound

Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers cites this paper.

Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers Training Vision Transformers for Image Retrieval

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:08:25.213225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T23:07:13.098350Z digest=sha256:677a19b7c9e0bcf57aaf85c3c2c41194feaf8a08392380ed1ed7d594c3ed6ac3