Pith. sign in

Paper Citation Record · LEDGER

Training Vision Transformers for Image Retrieval

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2102.05644.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.05644 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:50:11.550843Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T14:04:51.544148Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 2abb3203-1345-484a-8aa4-82127b40de8a · inbound

Emerging Properties in Self-Supervised Vision Transformers cites this paper.

Emerging Properties in Self-Supervised Vision Transformers Training Vision Transformers for Image Retrieval

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:04:51.546508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T14:04:51.458382Z digest=sha256:a384d70d60ebfbabb172af9592992a6a3670635d64157e94606d439a05a9f625

Observation 35492412-cdff-44a6-a619-efcafa076a9a · inbound

Vision Transformers Need Registers cites this paper.

Vision Transformers Need Registers Training Vision Transformers for Image Retrieval

Reference 174

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T09:41:38.139493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-13T09:41:37.937046Z digest=sha256:636907ba2b68110e2d4c85b5d966015f37d60bf4466f9b122523e5831dd5bc25

Observation 87601ccd-7095-4e79-ad9a-2244a30a7abe · inbound

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data cites this paper.

MATCHED: Multimodal Authorship-Attribution To Combat Human Trafficking in Escort-Advertisement Data Training Vision Transformers for Image Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:50:11.550843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T12:50:11.550843Z digest=sha256:2f4619b1c9e6ca8cf5da7d3517a04a75453e160957bf72e3fda30f4c2e75405e

Observation 07d81f0f-2041-4108-b269-6725a2a47e8a · inbound

Triplet Synthesis For Enhancing Composed Image Retrieval via Counterfactual Image Generation cites this paper.

Triplet Synthesis For Enhancing Composed Image Retrieval via Counterfactual Image Generation Training Vision Transformers for Image Retrieval

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T16:57:40.147683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T16:57:40.147683Z digest=sha256:e8fc126d8fa16324b22a2b4d32d1d166739fdfa0f8906538486f8df300bfb9aa

Observation 1ad6a074-8bef-4e57-8474-68e236ba1e03 · inbound

EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations cites this paper.

EndoFinder: Online Lesion Retrieval for Explainable Colorectal Polyp Diagnosis Leveraging Latent Scene Representations Training Vision Transformers for Image Retrieval

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T14:55:29.494550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:55:29.494550Z digest=sha256:e323701a6e0613af43f5cd1c0d1b166ab0d91cc21bbd59b2f3cd73e8b5171535

Observation e2a50346-eafe-4b05-a04e-e3a3da02de7a · inbound

Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers cites this paper.

Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers Training Vision Transformers for Image Retrieval

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:08:25.213225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T23:07:13.098350Z digest=sha256:027fd362a07def0eeddac25f03d27537d28318a163f8e532fddea156643dc399