Pith. sign in

Paper Citation Record · LEDGER

DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2106.02034.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2106.02034 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:48:48.087616Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:36:24.028635Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation eadc14ab-b024-49fc-b509-35fcd455a5a5 · inbound

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer cites this paper.

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:46:35.159941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T20:46:35.073600Z digest=sha256:2853c8ac188742ec09b49e482fee95266f3b430a4ead3fe0c9e36c4239584680

Observation fd173311-25a9-4ddf-b0af-d96955baa203 · inbound

Compact Vision Transformer by Reduction of Kernel Complexity cites this paper.

Compact Vision Transformer by Reduction of Kernel Complexity DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T16:48:48.087616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:48:48.087616Z digest=sha256:8b75119ba05aef7e7cb243c2f1c029edec6ae95986c4f6b9f46db8c88d0e8020

Observation 1cbac0c9-ca1c-4f56-adb7-065b17f59cd7 · inbound

CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition cites this paper.

CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:24:13.649881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:24:13.649881Z digest=sha256:dd8af2788e6e1ab579c70551ce803484f96dd1b253caba1e4df25e200ff61d17

Observation 16108892-33f1-4062-9814-51b0cd456f28 · inbound

DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking cites this paper.

DC-DiT: Adaptive Compute and Elastic Inference for Visual Generation via Dynamic Chunking DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T15:10:05.849862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T15:10:00.146694Z digest=sha256:83ab382ce71a8c6a318cb67d5b9fd8313695d18ca3d12be129c6e1f28fe3f64d

Observation e03e55f8-3645-4a8a-b5d1-50e7e61b4a65 · inbound

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs cites this paper.

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.030852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:04:43.270155Z digest=sha256:383d507491e71d5470e481b8b463f889893d8b4ba492d6926da98a2817c809df

Observation fa90b629-bd8d-4f05-8132-fcc0ef69b948 · inbound

Foveation-Guided Dynamic Token Selection for Robust and Efficient Vision Transformers cites this paper.

Foveation-Guided Dynamic Token Selection for Robust and Efficient Vision Transformers DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T02:43:47.779443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T02:43:47.779443Z digest=sha256:2287b8db2571a6bf588aa89f915d15e935944f287b134128feb0a0c657bd452a

Observation 24449e0d-0655-4c2c-b96f-e4c7f936677e · inbound

Searching for Task-Specific Vision Paths: Evolutionary Block Pruning Across Vision-Language Models cites this paper.

Searching for Task-Specific Vision Paths: Evolutionary Block Pruning Across Vision-Language Models DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-01T19:13:27.889997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T19:13:27.889997Z digest=sha256:1160da6bb9f0bc9ec6a735e100dd08a3397521e546eb85fd74671329a8df7a09

Observation e95ad9f4-fe05-433b-9d6a-1e36fc9ffd25 · inbound

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models cites this paper.

CRAFT: Compression via Recursive Adaptive Fusion of Video Tokens for Vision-Language Models DynamicViT: Efficient Vision Transformers with Dynamic Token Sparsification

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T23:46:51.535123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:46:51.535123Z digest=sha256:7139542f741c23397cdd5d87345dc2c9b87012b74a9aa54c2cd1e6578c0359c5