Pith. sign in

Paper Citation Record · LEDGER

CvT: Introducing Convolutions to Vision Transformers

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2103.15808.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2103.15808 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:18:21.009470Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.785007Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e042909f-f0fb-4025-816c-8f3439786601 · inbound

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer cites this paper.

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer CvT: Introducing Convolutions to Vision Transformers

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:46:35.184858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-20T20:46:35.073600Z digest=sha256:f2a6c91633e1e5a29d2d7b3355c09ae8aa0b40cf9b0727b135d2e818a87075ce

Observation ad6687e0-d433-4f45-b37f-1f302491447b · inbound

Convolutional Vision Transformer for Cosmology Parameter Inference cites this paper.

Convolutional Vision Transformer for Cosmology Parameter Inference CvT: Introducing Convolutions to Vision Transformers

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T15:18:21.009470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:18:21.009470Z digest=sha256:0717f704d988241601e99b2be334cd2745bffc0821d056c85d132d4dfca22e8e

Observation c18b2d64-e096-4035-98aa-d7ef4360257b · inbound

Unified Local and Global Attention Interaction Modeling for Vision Transformers cites this paper.

Unified Local and Global Attention Interaction Modeling for Vision Transformers CvT: Introducing Convolutions to Vision Transformers

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T04:34:02.303884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:34:02.303884Z digest=sha256:9546f17ea3a7185b4dc97b4fbe07ad4f9b688e492808a90c4d02e99961085809

Observation 3e529189-274c-41c4-9b48-e5ace37efa68 · inbound

Residual Transformer Fusion Network for Salt and Pepper Image Denoising cites this paper.

Residual Transformer Fusion Network for Salt and Pepper Image Denoising CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T23:01:17.102588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:01:17.102588Z digest=sha256:2f44cd8a6f024d83cea12684ffab6978f4b96829e251b5de2bb743494f4379f2

Observation 348189fe-11a6-49c0-bba1-d12b81a82286 · inbound

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment cites this paper.

JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment CvT: Introducing Convolutions to Vision Transformers

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T13:14:37.600229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T13:14:37.600229Z digest=sha256:15dcab5ccac920f4134ae826b7143c383d32126ddcad8d823cf27cebb93429eb

Observation 342d9fec-7898-4ea9-95be-6a72925cb26a · inbound

Enhancing compact convolutional transformers with super attention cites this paper.

Enhancing compact convolutional transformers with super attention CvT: Introducing Convolutions to Vision Transformers

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T16:06:14.447164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:06:14.447164Z digest=sha256:7265bdfc84cface6f70c9050b63a82c15c88673b28222749506bb9233a1df185

Observation 07cf10e3-67cb-4e75-beef-b94f2606b45f · inbound

Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach cites this paper.

Single Domain Generalization in Diabetic Retinopathy: A Neuro-Symbolic Learning Approach CvT: Introducing Convolutions to Vision Transformers

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T11:24:06.954661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:24:06.954661Z digest=sha256:b3966527d96f67567a2c018238d0d63d1b77540ee861df1013d1f56f5a4ecdfe

Observation 3c04529a-b094-413f-a3a5-dce46cf245bc · inbound

Dual-attention ResNet outperforms transformers in HER2 prediction on DCE-MRI cites this paper.

Dual-attention ResNet outperforms transformers in HER2 prediction on DCE-MRI CvT: Introducing Convolutions to Vision Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T09:55:52.457277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:55:52.457277Z digest=sha256:efcc45f34ce07f18e6d523fdd27274bced36b94eb3bca99b247221f7a47cfb0c

Observation f0367cbd-7e53-4350-8a75-ec95f82c3a95 · inbound

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents cites this paper.

ClawEnvKit: Automatic Environment Generation for Claw-Like Agents CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.786479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-07-05T11:32:36.356530Z digest=sha256:fb207702f99116467ded82338746a8790477e6e184d8ff7bd5880de0e1d77f5d

Observation 422776b5-69fd-47eb-942a-37aec0eb6af9 · inbound

Advancing Vision Transformer with Enhanced Spatial Priors cites this paper.

Advancing Vision Transformer with Enhanced Spatial Priors CvT: Introducing Convolutions to Vision Transformers

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:23:37.655388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=pdf_text observed=2026-05-10T05:22:21.264807Z digest=sha256:d01b51eef57b1a6bda980424a6515982dcd747260c859a59847f7e4118d5678c