Pith. sign in

Paper Citation Record · LEDGER

Screening Is Effective for Visual Recognition

As of 7 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:2607.13983.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.13983 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T03:08:45.931239Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f50e1e6c-f2f4-4ec1-b6cc-ce6525dcd32e · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Screening Is Effective for Visual Recognition Imagenet: A large-scale hierarchical image database

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.625633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.625633Z digest=sha256:30fa8f7322b8008312696db30e3359ab26bfeb3db444f8526d05d2c645db7af8

Observation fb68a3cf-1f0f-4f60-85be-0367b4915b24 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Screening Is Effective for Visual Recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.712232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.712232Z digest=sha256:efacc279c585fe79abed544a7a20f416f68495d3252d44a1f0fac4bc4e72cc13

Observation 21ffc286-8f2c-4e58-8774-0b085b00e7af · outbound

This paper cites Masked autoencoders are scalable vision learners.

Screening Is Effective for Visual Recognition Masked autoencoders are scalable vision learners

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.792791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.792791Z digest=sha256:592f6ff4b20ec912f442e16f6790217b49d88427e35e803e660bf560d56c69e9

Observation 1e6dd1dc-6674-48f3-ade0-27f829df1c6a · outbound

This paper cites Rotary position embedding for vision transformer.

Screening Is Effective for Visual Recognition Rotary position embedding for vision transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.866848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.866848Z digest=sha256:9051840424030e335e2a4e24f3d158e5f12df805c6a790e779b0f83afd67f7de

Observation 81de5a9a-631b-48f2-8f50-37d41aef5ce5 · outbound

This paper cites Segment any- thing.

Screening Is Effective for Visual Recognition Segment any- thing

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.922564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.922564Z digest=sha256:5ab0e7c0d2291d1dc373388f2125c8aa21def6b928c135cd44151ad99aad026f

Observation cfd16ff1-cfbe-40c7-bb80-8be6f2438ebf · outbound

This paper cites Big transfer (bit): General visual representation learning.

Screening Is Effective for Visual Recognition Big transfer (bit): General visual representation learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:44.988601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:44.988601Z digest=sha256:906d57fa4ac6e458c4a644da8b0994863089931d7e0e4fad8082aae87526cf04

Observation b328bd25-e0dd-4c1e-ac92-f08f5bf9f89f · outbound

This paper cites Learning multiple layers of features from tiny images.

Screening Is Effective for Visual Recognition Learning multiple layers of features from tiny images

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.110278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.110278Z digest=sha256:769be9f06fb0b9c9535a70785f98c2979592ea7e56f51002f190b498fd7ec6a6

Observation 32b0d7c4-8a60-47ff-ab4e-addfeb0bc8a7 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Screening Is Effective for Visual Recognition Swin transformer: Hierarchical vision transformer using shifted windows

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.158294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.158294Z digest=sha256:33758c7a603d407b13757bca5aa0f9055aa55628b047d2010eb4c08b97a2d053

Observation ff52a691-8f71-4350-9486-d3e2419facd9 · outbound

This paper cites Screening Is Enough.

Screening Is Effective for Visual Recognition Screening Is Enough

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.224306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.224306Z digest=sha256:3a3952218f2e1c8a45862939e43023a28d629efbe02dfe91c88fc5f7d7741717

Observation 2429aa89-af95-4e09-94cd-aac9c58bd661 · outbound

This paper cites Learning transferable visual models from natural language supervi- sion.

Screening Is Effective for Visual Recognition Learning transferable visual models from natural language supervi- sion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.305530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.305530Z digest=sha256:b8fd3cabd85130d11c5fdbed63d0fe79a7ec63bd5510de823042902f43b36fa8

Observation ebba460c-3cad-45ec-bfdd-c603b8daea9e · outbound

This paper cites How to train your ViT? Data, Augmentation, and Regularization in Vision Transformers.

Screening Is Effective for Visual Recognition How to train your ViT? Data, Augmentation, and Regularization in Vision Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.376016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.376016Z digest=sha256:d1edc62f0a00279c3b4cb358f6a2e155aac5f58076f10927036f0d5683caa0e7

Observation 6ac32eeb-84e1-4a82-95c5-949623c2f68c · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063,.

Screening Is Effective for Visual Recognition Roformer: Enhanced transformer with rotary position embedding.Neurocomputing, 568:127063,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.442903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.442903Z digest=sha256:e2c0c6e30074e2226dbbac57991a56e5fe466602124d08945fe482bf9eb07f64

Observation 970419ec-61db-4cc7-b537-505e50ee01dc · outbound

This paper cites Mlp-mixer: An all-mlp architecture for vision.Advances in neural information processing systems, 34:24261–24272,.

Screening Is Effective for Visual Recognition Mlp-mixer: An all-mlp architecture for vision.Advances in neural information processing systems, 34:24261–24272,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.532987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.532987Z digest=sha256:db0042b2d7a1f7da1822d6c305f7090ec875858016e1f08cbead3f78f31fb047

Observation 19aed776-c5ce-4eed-ad5b-20c2d8414526 · outbound

This paper cites Training data-efficient image transformers & distillation through at- tention.

Screening Is Effective for Visual Recognition Training data-efficient image transformers & distillation through at- tention

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.612160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.612160Z digest=sha256:cecf1012366fd6c4d8c9265f602de1381b019d4fd55ca81745c770ae96100cea

Observation 73f93eda-8c93-4fb3-a341-14cc6088dbe0 · outbound

This paper cites Resmlp: Feedforward networks for image classification with data-efficient training.IEEE transactions on pattern analysis and machine intelligence, 45(4):5314–5321, 2022.

Screening Is Effective for Visual Recognition Resmlp: Feedforward networks for image classification with data-efficient training.IEEE transactions on pattern analysis and machine intelligence, 45(4):5314–5321, 2022

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.706332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.706332Z digest=sha256:01cabb6d400a65df29d37fd833ca18b1b708893b2251e8fcde418f9a276cbf71

Observation 02c4c6a6-84b7-4ddc-ba47-125c3176be3b · outbound

This paper cites Segformer: Simple and efficient design for semantic segmentation with transform- ers.Advances in neural information processing systems, 34: 12077–12090, 2021.

Screening Is Effective for Visual Recognition Segformer: Simple and efficient design for semantic segmentation with transform- ers.Advances in neural information processing systems, 34: 12077–12090, 2021

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.831713Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.831713Z digest=sha256:9a88f18ac3b3f5a5293d0b92c85b03ec3c234c17ebf67bd4be8a9e205e0be75e

Observation 88c0535b-1857-477b-a2da-5a24e74c17e8 · outbound

This paper cites Metaformer is actually what you need for vision.

Screening Is Effective for Visual Recognition Metaformer is actually what you need for vision

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T03:08:45.931239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:08:45.931239Z digest=sha256:45e4f2409472dce27ecc780a87f0a59abc53fe4122366f95fe502031651ac99d

Pith citing papers

No inbound Pith citation observations are available.