Pith. sign in

Paper Citation Record · LEDGER

LocalViT: Analyzing Locality in Vision Transformers

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2104.05707.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2104.05707 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:27:42.235093Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T18:01:54.629944Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 362472ac-08a6-4a50-adb1-050590edded2 · inbound

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer cites this paper.

MobileViT: Light-weight, General-purpose, and Mobile-friendly Vision Transformer LocalViT: Analyzing Locality in Vision Transformers

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:46:35.141309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T20:46:35.073600Z digest=sha256:8edf7433d8bffa78b71b04ea3f30d8eef0e07d78274a2bb8ce9899ed608e4886

Observation 9c8065b1-37c3-4d02-bd73-0309dfd71eb5 · inbound

Cross Paradigm Representation and Alignment Transformer for Image Deraining cites this paper.

Cross Paradigm Representation and Alignment Transformer for Image Deraining LocalViT: Analyzing Locality in Vision Transformers

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-22T18:01:54.631773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T17:57:05.010694Z digest=sha256:446cd05f5ecefe0fbb3e319ed3e5762373d1209c990dfb9f0145116fc281795e

Observation e99e742e-a5f6-45df-9cbb-e05bff0368dc · inbound

RoadFormer : Local-Global Feature Fusion for Road Surface Classification in Autonomous Driving cites this paper.

RoadFormer : Local-Global Feature Fusion for Road Surface Classification in Autonomous Driving LocalViT: Analyzing Locality in Vision Transformers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.235093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.235093Z digest=sha256:d0ae5696cb7722b32ed3faa29ed52b0accfd079f77c233220d65c94fc516b5c4

Observation 52397976-09f5-4ae8-9038-7b709136d20c · inbound

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer cites this paper.

DuoFormer: Leveraging Hierarchical Representations by Local and Global Attention Vision Transformer LocalViT: Analyzing Locality in Vision Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T00:42:00.924074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:42:00.924074Z digest=sha256:1161e17b4139b133983b28f3caa08dc3a22365d66c7071025ec549f8a04b16fc

Observation b8b2ed47-6315-410e-ba62-97af32577f5b · inbound

DFYP: A Dynamic Fusion Framework with Spectral Channel Attention and Adaptive Operator learning for Crop Yield Prediction cites this paper.

DFYP: A Dynamic Fusion Framework with Spectral Channel Attention and Adaptive Operator learning for Crop Yield Prediction LocalViT: Analyzing Locality in Vision Transformers

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T19:22:27.607433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:22:27.607433Z digest=sha256:09293ae588eae56c2bc0d2ff453c98bb8193977c7e1e51ef0b6e9b21a407ed74

Observation 584165ce-fe80-42d9-944f-a70814732004 · inbound

HydraMamba: Multi-Head State Space Model for Global Point Cloud Learning cites this paper.

HydraMamba: Multi-Head State Space Model for Global Point Cloud Learning LocalViT: Analyzing Locality in Vision Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T14:04:42.182463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:04:42.182463Z digest=sha256:f0f1e1348a9268d6db13bd516262c87f4852ff8183c68df5329be801bf3f5645

Observation cb19b660-e1cd-4b84-ab86-8ae59c661fbe · inbound

From Local Windows to Adaptive Candidates via Individualized Exploratory: Rethinking Attention for Image Super-Resolution cites this paper.

From Local Windows to Adaptive Candidates via Individualized Exploratory: Rethinking Attention for Image Super-Resolution LocalViT: Analyzing Locality in Vision Transformers

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T10:54:55.619439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:54:55.619439Z digest=sha256:e507b56d97acb75cba3a739cfda3246cfb07894aa5d2854de74b0277a0a81b4c

Observation 25b10df3-c0cb-4b17-b663-652c629e760d · inbound

KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition cites this paper.

KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition LocalViT: Analyzing Locality in Vision Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:36:09.284615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T08:31:35.014294Z digest=sha256:70ce3d10bb01a692b387ce4f1cd8c6228f00e6056842cb9306b065a0f6d7eaed