Pith. sign in

Paper Citation Record · LEDGER

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging

As of 7 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 0 inbound Pith citation observations for arXiv:2507.01788.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.01788 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:48:06.271519Z

measured 31 of 31 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

31 of 31 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b8590585-bd2b-42b6-bcf5-e2d0ab33e881 · outbound

This paper cites On the opportunities and risks of foundation models,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging On the opportunities and risks of foundation models,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.558040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.178125Z digest=sha256:a7132670871653e92736c7fbb2ef81449ae37dfad63d70f1a94b65302936f1a4

Observation e82de7a8-9f9f-4596-97c4-c419e026fa51 · outbound

This paper cites Gpt-4 technical report,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Gpt-4 technical report,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:48:06.183085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:48:06.183085Z digest=sha256:108810996510d8cfe94b9ca020813d505cbdf44e953e1cde419f7930163f862f

Observation 6b2e7be4-1945-47eb-b470-00002d852414 · outbound

This paper cites ChatGPT goes to law school,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging ChatGPT goes to law school,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.544313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.186692Z digest=sha256:a97596ddcc45e907dffa5bf787cd828bb263aae0b2c19d3114d8f12ca018d802

Observation e204a648-03f0-4fcd-bf81-0dd46e627ebe · outbound

This paper cites ProteinBERT: a universal deep- learning model of protein sequence and function,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging ProteinBERT: a universal deep- learning model of protein sequence and function,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.534886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.189644Z digest=sha256:0e759918a59d1d892f095331a007e174ace85a64db3543c079b9609811efb6e4

Observation 99e11830-2207-4a17-acd9-a03b41c76a7b · outbound

This paper cites Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Performance of ChatGPT on USMLE: Potential for AI-assisted medical education using large language models,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.525573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.192663Z digest=sha256:b604efb7fddbcf070f8a9ab6ef71f9927ab2865d0aaa7745be85e3582529120e

Observation e4d9eaea-185d-43db-a97a-32bbe34008cd · outbound

This paper cites Attention is all you need,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Attention is all you need,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.516834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.195871Z digest=sha256:7accfb3db11bbc7fb81166122d58d575cb2f9dcadd1e3b98be410e1f6e847e5c

Observation e68378a1-50a6-4fbe-92bf-8908865448a0 · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:48:06.198963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:48:06.198963Z digest=sha256:2964026d6522ffb8c91be6b0992dd42873db80ca760a5a6f17906f41cfe4d00b

Observation 2c20c33a-3b11-478f-a665-f617ce33976a · outbound

This paper cites Bert: Pre-training of deep bidirectional trans- formers for language understanding,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Bert: Pre-training of deep bidirectional trans- formers for language understanding,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.501586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.201823Z digest=sha256:049d8d4cd60ed3fba6107bbcc7f21b27017b76918b0ae686bbe8930d2acc93b7

Observation f8805545-0dfa-4707-a453-fc1f8b00f5ad · outbound

This paper cites Transformers in medical imaging: A survey,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Transformers in medical imaging: A survey,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.493044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.204797Z digest=sha256:6b2298184f560d4c3e3356c0280b606a980d9c8115980d206af425a5bc90290d

Observation 1ff4674a-63c2-48f1-8475-8c1417a39f86 · outbound

This paper cites A comparative study between vision transformers and cnns in digital pathology,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging A comparative study between vision transformers and cnns in digital pathology,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.484726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.208264Z digest=sha256:6b17da0f1203edde1b750ab21cd8afb6326b3e7054ca06033a204f891549231e

Observation caef6f7d-8819-4628-ac77-388d7f7ebb73 · outbound

This paper cites Mil-vt: Multiple instance learning enhanced vision trans- former for fundus image classification,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Mil-vt: Multiple instance learning enhanced vision trans- former for fundus image classification,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.475527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.211258Z digest=sha256:17c4637a57f6169adc243fc249cf62781cd9393f72edf9f937c050a2a17a70ae

Observation 93c4edc1-de78-4347-b9b7-b0d397b66b2f · outbound

This paper cites Med- vit: A robust vision transformer for generalized medical image classification,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Med- vit: A robust vision transformer for generalized medical image classification,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.467320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.214082Z digest=sha256:d2d918122be1aa20ac34d80f70c96302e993abc8946c4c2e80f614804bd044bd

Observation 9787056f-749f-45ce-a93f-facce120bb23 · outbound

This paper cites A recent survey of vision transformers for medical image segmentation,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging A recent survey of vision transformers for medical image segmentation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.458347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.217095Z digest=sha256:664bc72aacd758a13269ed0090fa36152ab1a774b25f85eaf9d03a841d9b4aad

Observation 38f040a5-fd81-47a8-ad03-2f4d6819d063 · outbound

This paper cites Comparing cnns and vits for medical image classification leveraging transfer learning,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Comparing cnns and vits for medical image classification leveraging transfer learning,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.449939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.220647Z digest=sha256:e51151411de30c4ee7d6145a2cf8cd3c0b58c21aeb428d13c223b4cb76ca2148

Observation 0ed3c8a3-687c-4d89-bcb7-69d5c002da39 · outbound

This paper cites Biomedclip: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Biomedclip: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.441524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.223339Z digest=sha256:46087420a47daf1507c7bd49bc32354556ac67b0357dde273deddb70fd1957e0

Observation d2408800-9548-4be7-92bf-8f18e21a73e8 · outbound

This paper cites Pmc-clip: Contrastive language-image pre-training using biomedical documents,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Pmc-clip: Contrastive language-image pre-training using biomedical documents,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.432980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.226268Z digest=sha256:8c77a926d084a0135277bbe015564a28539e3008ed392ff2f4ca55ec46d24ddb

Observation c8cbbe66-18f8-4b59-9b4d-594fb035f5ed · outbound

This paper cites Explaining and harnessing adversarial examples,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Explaining and harnessing adversarial examples,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.424780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.229080Z digest=sha256:71bf34e644ddb81e7e801a4a19c2118bead5f1f94fc3aa9a789f9fed87b35589

Observation 2d80b1f1-d874-42be-a925-1f5575ca5314 · outbound

This paper cites Intriguing properties of neural networks,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Intriguing properties of neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.416445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.231798Z digest=sha256:4e135019201f0bfe25e6c17793e466477fec974f403b597e978c7e7b303cda74

Observation dd802563-c8bf-4920-bb64-f00154100a5a · outbound

This paper cites Towards deep learn- ing models resistant to adversarial attacks,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Towards deep learn- ing models resistant to adversarial attacks,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.408266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.234602Z digest=sha256:ffc0340b256b07698a32636e4964b3e09e06cf1f0695bccd1d0af20b77c1fddf

Observation d9ae84a6-12c2-45c7-89b3-7357a7e37da3 · outbound

This paper cites Survey on adversarial attack and defense for medical image analysis: Methods and challenges,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Survey on adversarial attack and defense for medical image analysis: Methods and challenges,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.397482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.237518Z digest=sha256:bdd810e070f231ec81fd132259b621c584bdbb1f9f35a89e14c8018f0f71114d

Observation f367ca67-4dec-4617-b6f7-37134f2d2d2e · outbound

This paper cites Generalizability vs. robustness: Ad- versarial examples for medical imaging,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Generalizability vs. robustness: Ad- versarial examples for medical imaging,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.388914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.240417Z digest=sha256:944bed74a15ad2b3308805c89187c90dcb3e303dac3d02f9cc5e0f3f0e44bd1d

Observation 66bbd2e3-55fc-43b5-b4ed-cddfaa9be4c2 · outbound

This paper cites Adversarial attacks on medical machine learning,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Adversarial attacks on medical machine learning,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.380033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.244021Z digest=sha256:5890e74be8b51ef061eac08d614552672f14d34dc713fbc0ca06d91cf708e380

Observation 6edd2d57-7429-456b-88fb-d8ca1c2b07a8 · outbound

This paper cites Understanding adversarial attacks on deep learning based medical image analysis systems,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Understanding adversarial attacks on deep learning based medical image analysis systems,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.370807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.246835Z digest=sha256:0ad76bad532a41ffa9d012dd75df7cc98f03380510c998a61f57da2c97ff00c2

Observation 5c8ddaf1-e292-4932-91f0-c02693fafb65 · outbound

This paper cites Adversarial attacks and adversarial robustness in computational pathology,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Adversarial attacks and adversarial robustness in computational pathology,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.361777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.249761Z digest=sha256:745cfe2e30f361d1d851be49efb0dca2aa8e2bd9cd4342305e0a98b4581b398c

Observation ab1cf7ac-2b95-4088-a250-dc4af2ac58e9 · outbound

This paper cites Un- derstanding robustness of transformers for image classifica- tion,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Un- derstanding robustness of transformers for image classifica- tion,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.352734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.252924Z digest=sha256:ac4b9b6d3c4eda55d594b201dd80e3b28e0508ef646d5ec97b32a1fc93f3d586

Observation 7a2e4772-29b9-4664-85a3-570cc2b67765 · outbound

This paper cites Intriguing equivalence structures of the embedding space of vision transformers,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Intriguing equivalence structures of the embedding space of vision transformers,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.343626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.256436Z digest=sha256:88cfa0b1f2f0493e8fc11815dff77b1b915317e529f167ee522d47963417d5bf

Observation 2f8df52d-5b2a-43ae-aaae-efb206723585 · outbound

This paper cites Medmnist v2 - a large-scale lightweight benchmark for 2d and 3d biomedical image classification,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Medmnist v2 - a large-scale lightweight benchmark for 2d and 3d biomedical image classification,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.333993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.259300Z digest=sha256:a949e257309740828c69367b7a2275712de2e7aab365ca3ddaaf635d00b75777

Observation 66c8090a-1d8e-4606-b512-31911850b2d3 · outbound

This paper cites Image quality metrics: Psnr vs. ssim,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Image quality metrics: Psnr vs. ssim,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.324883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.262068Z digest=sha256:5da36e147cb5ddb9cf324b22bbd308380265ad48d3a7c3b126b7d7f6d51597f4

Observation d02b0a57-775d-468a-af7f-310524196fcd · outbound

This paper cites Feature forwarding for efficient single image dehazing,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Feature forwarding for efficient single image dehazing,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.316067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.264868Z digest=sha256:a3fbdda0c60cda891f2ccba0d99dc3a9a8f9e1ec18150868815f6b2206034448

Observation f25fdd7e-4008-4677-9c08-c5c139ecf1e2 · outbound

This paper cites Understanding zero-shot adversarial robustness for large-scale models,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Understanding zero-shot adversarial robustness for large-scale models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.307349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.268622Z digest=sha256:1eed91e76fff83fa2a18d7eae5d1f689d8fe9ad9eb49ef9ac1e3610e195c6666

Observation 4dc02add-cb5f-4c48-9b8b-58c28e757e26 · outbound

This paper cites Malicious path manipu- lations via exploitation of representation vulnerabilities of vision-language navigation systems,.

Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging Malicious path manipu- lations via exploitation of representation vulnerabilities of vision-language navigation systems,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:48:06.298100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T20:48:06.271519Z digest=sha256:606ad24211deb6207fd1cbddf03544905af476cd18f1dc451da2953b83014df8

Pith citing papers

No inbound Pith citation observations are available.