Pith. sign in

Paper Citation Record · LEDGER

What Do Self-Supervised Vision Transformers Learn?

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2305.00729.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.00729 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:29:04.769709Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:29:29.165580Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fabb99a7-f2ce-4e71-945f-ce5a5b93dea0 · inbound

On the Surprising Effectiveness of Attention Transfer for Vision Transformers cites this paper.

On the Surprising Effectiveness of Attention Transfer for Vision Transformers What Do Self-Supervised Vision Transformers Learn?

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T20:27:06.912366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:27:06.912366Z digest=sha256:d19dd0909bab85cbeb10d2af639f6031eab45d972188e0233d6c82ca91581a05

Observation 30d36c02-f4b6-46db-b23f-b3e5ebda92f5 · inbound

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition cites this paper.

Relational Contrastive Learning and Masked Image Modeling for Scene Text Recognition What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T18:51:51.115642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T18:51:51.115642Z digest=sha256:fdd37d14e76210422394678ceca8d09fa394a8e1e3247139d368d772702e9766

Observation 53052e91-1044-4956-bd42-e17f72a7fbbe · inbound

Beyond Scalars: Concept-Based Alignment Analysis in Vision Transformers cites this paper.

Beyond Scalars: Concept-Based Alignment Analysis in Vision Transformers What Do Self-Supervised Vision Transformers Learn?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T19:32:03.000362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:32:03.000362Z digest=sha256:527417d7cd27f2d20f1ee19fafb83589c1bb558cc37395b4593943adb655ea9c

Observation a58a8716-1fab-4a4d-90eb-31829caef806 · inbound

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving cites this paper.

Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-16T12:29:04.769709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:29:04.769709Z digest=sha256:72e819c97d9cce4e9db7dd78c793252c1d627cf4c6e817d6d3fd48563a8cfd7a

Observation c4e1a270-117c-4352-a6d8-f674418d3192 · inbound

FourierFlow: Frequency-aware Flow Matching for Generative Turbulence Modeling cites this paper.

FourierFlow: Frequency-aware Flow Matching for Generative Turbulence Modeling What Do Self-Supervised Vision Transformers Learn?

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:19.301548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:02:19.301548Z digest=sha256:c48dd0156cfaacb8c3d41269bde9c58dd05a9899a1c484e25c311f9b4bd6e843

Observation d4ef4d47-f9b0-4f19-b51a-22e64076b132 · inbound

Self-Guided Masked Autoencoder cites this paper.

Self-Guided Masked Autoencoder What Do Self-Supervised Vision Transformers Learn?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:55.768397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:13:55.768397Z digest=sha256:c0078e364d2ccdb70c0f9ad55ac5ab0ff1d3aab1bee69d257c5aef2db9e242e6

Observation 2b39d459-1780-468e-8f95-86b3e6480a83 · inbound

Dynamic Pattern Alignment Learning for Pretraining Lightweight Human-Centric Vision Models cites this paper.

Dynamic Pattern Alignment Learning for Pretraining Lightweight Human-Centric Vision Models What Do Self-Supervised Vision Transformers Learn?

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T22:22:58.793132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:22:58.793132Z digest=sha256:4612385e6a9d3880f6417b734bdf1eedca3257f5a4a3949f281a11b3bcf89ad6

Observation 2069efb1-1c01-4349-98c5-a891ba92175e · inbound

Agentic AI in Remote Sensing: Foundations, Taxonomy, and Emerging Systems cites this paper.

Agentic AI in Remote Sensing: Foundations, Taxonomy, and Emerging Systems What Do Self-Supervised Vision Transformers Learn?

Reference 99

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:31:10.824852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T18:28:33.277442Z digest=sha256:bc2bb61f9ffc8994ef397128e329657ab1e9cc6d26a87de2b4c60c69c10aa753

Observation cf2828a1-66e3-4d08-a901-df90ddda564d · inbound

Unsupervised Semantic Segmentation Facilitates Model Understanding cites this paper.

Unsupervised Semantic Segmentation Facilitates Model Understanding What Do Self-Supervised Vision Transformers Learn?

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:43:15.488660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T08:37:02.350175Z digest=sha256:6e5880702bdbb5467c26360ac311e104b34201c246788e2894222d2955ed1d24

Observation 77a0e991-0237-48ca-afef-02cb99a45e70 · inbound

Unsupervised Semantic Segmentation Facilitates Model Understanding cites this paper.

Unsupervised Semantic Segmentation Facilitates Model Understanding What Do Self-Supervised Vision Transformers Learn?

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:39:16.406737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-04T00:34:21.224797Z digest=sha256:d2716503c08d76fbd37f16a0646c40191a531003a21de999511c732815feca08

Observation 067f32a6-dd26-463c-bc9a-3dc03c5a2897 · inbound

The Hidden Evolution of Disguised Visual Context inside the VLM cites this paper.

The Hidden Evolution of Disguised Visual Context inside the VLM What Do Self-Supervised Vision Transformers Learn?

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:29:29.169229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T18:08:56.044278Z digest=sha256:ca64bba695588c5f2684ecc19f78d33add62114fc2c26008f5f52b0e8cd0bb8c

Observation 299be697-32bd-403b-bfda-b5d5256831b5 · inbound

Understanding Geometric Representations in Self-Supervised Vision Transformers via Subspace Intervention cites this paper.

Understanding Geometric Representations in Self-Supervised Vision Transformers via Subspace Intervention What Do Self-Supervised Vision Transformers Learn?

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:38:33.120461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-03T15:34:54.954593Z digest=sha256:063dbc4582dfdfac3774ba7335c44c18bddcf9ec870a8f4720efbce3c712fbaf