Pith. sign in

Paper Citation Record · LEDGER

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI

As of 9 August 2026, this Paper Citation Record lists 100 of 251 outbound references and 0 inbound Pith citation observations for arXiv:2608.05258.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05258 v1

Coverage vector

measured 100 of 251 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:59:11.105981Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 251 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved96
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d11799d7-cbac-4c50-a9c6-c44c89d41f55 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.556089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.556089Z digest=sha256:bd34aabb27a2a58101732478fbf1d7cff8e62edef3c57fc7f0a5f70947ef3e04

Observation 6f4bef24-111c-4059-9ebc-51ec3ae2795f · outbound

This paper cites Representation learning and na- ture encoded fusion for heterogeneous sensor networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Representation learning and na- ture encoded fusion for heterogeneous sensor networks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.561346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.561346Z digest=sha256:15f445eec05a066c0825c576f33fef4905d7301ecc7e813f308f8ac70f5448a8

Observation 0e2d9a63-413d-4eb4-9378-4d3f3b328dee · outbound

This paper cites Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.565742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.565742Z digest=sha256:a67bee370c0ef1f9ac1f872c5981b83d75239b58350093efc63367b72015b84b

Observation 165e5521-1a09-4255-b96d-1b3629368d1d · outbound

This paper cites Enhanced robustness by symmetry enforcement,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhanced robustness by symmetry enforcement,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.570437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.570437Z digest=sha256:d26864ad2d45ee1dd113f80013c84159f37428f6950c63fc766e4c3d893a107d

Observation 0a432fff-c026-428e-b74b-259d3a45497a · outbound

This paper cites Partial interference alignment for heterogeneous cellular networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Partial interference alignment for heterogeneous cellular networks,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.575167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.575167Z digest=sha256:a038d4078f1d4a034e8dcd3667ff153c87e8c7dec5e297e85249120ec51d8b93

Observation d61ebdc0-a3e0-4c47-9539-b20da33af6c3 · outbound

This paper cites Optimization for user centric massive mimo cell free networks via large system analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Optimization for user centric massive mimo cell free networks via large system analysis,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.580043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.580043Z digest=sha256:9747e6fdcf066495cb8435cee1a262f818ce620b7d5d85a8a50b82ea0d697575

Observation 2ee2b689-e79e-4419-9d9c-87ebb4c5be66 · outbound

This paper cites Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.589472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.589472Z digest=sha256:df14ca17e9731a2eecfa8e53d6795f37d35a9070b64efef2545fc46e7ffe23c3

Observation 9fa2d7cb-02eb-49cc-9232-b08d59763459 · outbound

This paper cites Deep reinforcement learning based computation offloading for mobility-aware edge computing,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Deep reinforcement learning based computation offloading for mobility-aware edge computing,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.593676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.593676Z digest=sha256:cd190a1936afa06dc0b34a361cd437dcb058eb89034a63d0455e7c09d7924bf0

Observation e4dffcba-d670-4690-a55a-66053719d781 · outbound

This paper cites Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.598216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.598216Z digest=sha256:c47b98ddf6daebf33c7700b0b5bceda8706ff2b0e2e9ade3f7da7769243d375e

Observation 436732dc-d1bc-4b90-83ff-e0eb42420b21 · outbound

This paper cites Low complexity optimization for user centric cel- lular networks via large dimensional analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Low complexity optimization for user centric cel- lular networks via large dimensional analysis,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.602562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.602562Z digest=sha256:c4cdb3b4bca34299090efecf540e53d26ac224543bb2abf553029725679f2330

Observation 080326ea-9f1c-40ed-a3a6-f66ce1a212ab · outbound

This paper cites Improving robustness of deep neural networks via large-difference transformation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving robustness of deep neural networks via large-difference transformation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.607014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.607014Z digest=sha256:0c96b20ce3b99697e9763be06afacf9f577097615b0d093c6772c48a99116ec2

Observation c0a28969-e207-40c6-b238-b9141d7a21a9 · outbound

This paper cites Looking beyond content: Modeling and detection of fake news from a social context perspective.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Looking beyond content: Modeling and detection of fake news from a social context perspective

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.611370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.611370Z digest=sha256:d11c0d0d51691493e25168ac4c652cdef4ec7415fe031563a2eb13f0ff73e87d

Observation b70425f5-9a7f-40fa-a5f3-fc1c4fc27c6d · outbound

This paper cites Large system analysis for densification of cellular networks with massive mimo,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large system analysis for densification of cellular networks with massive mimo,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.615627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.615627Z digest=sha256:8f63cd0bb7a15202ea7b86bfda57b6a32904f897d6e4f02d1068c8ebcd529709

Observation a73bec4b-fc33-4bd0-ae8e-aa38cba8ae2e · outbound

This paper cites Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.620148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.620148Z digest=sha256:b2b54f2e43f7bc94ea745dfb4fb84a5978d3b53ef70743850c2570ce2c64ba2c

Observation 4187ea7e-2c05-4ca1-aa10-17fbc28ccd8d · outbound

This paper cites Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.624519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.624519Z digest=sha256:71c7f303545dd9d6cedf96dd9d61dd0096273d8b885a8fbb6224c616c414f337

Observation 8a2fd787-2e48-4a4b-9cd7-e89decb85427 · outbound

This paper cites Information theory and represen- tation learning inspired multimodal data fusion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Information theory and represen- tation learning inspired multimodal data fusion,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.629428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.629428Z digest=sha256:0f156c8e8c3660dc54291943f79328a483364a22944e966c51bf7054215e6cab

Observation 560e8a2e-44dc-4470-a3d9-6adbcdc5c858 · outbound

This paper cites Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.286699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.638625Z digest=sha256:d812feb4338d75e95ffdda6c422e5f4e05d1d8f058d2622ea0a7079e1cc2e2df

Observation 5bf5c5cb-b07f-465f-b50e-d916ce2e0b23 · outbound

This paper cites Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.647891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.647891Z digest=sha256:adbbac7583e0e37b8c1b617c2f9cee90db1ebd4d36dc18378648d8348a359f87

Observation fa094813-9e56-4e25-adf3-58b74e9fbf87 · outbound

This paper cites Expert-guided ex- plainable few-shot learning for medical image diagnosis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning for medical image diagnosis,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.652271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.652271Z digest=sha256:3d2cb533473e32802f68084d9c7a73ea0843ed40811b72276cbe67725bc9cc02

Observation ebb2d243-fe70-4e43-9158-7b3e6fd6f25b · outbound

This paper cites GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.656692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.656692Z digest=sha256:68ee5ad2a38c1bcc8d18d3cffab3f38c82be8fc5c253750fb2ac588baa5eacc2

Observation 0a9a379b-e784-4fb8-8ece-6d18600a778f · outbound

This paper cites Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.661437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.661437Z digest=sha256:ab5e72b41ce10a36d0a835e72c59f4434fb22c1d28a47ca6c9cd99cacfd65547

Observation 2c2acf1c-776b-434b-ae53-208c6063aa2d · outbound

This paper cites Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.665772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.665772Z digest=sha256:bb0245a95e1e45331fbc4cfabd3d2bd8614ae93fb54f3f3f53e97c6b5494f592

Observation 17e6e87b-f390-4c29-88ef-d4a356b0cd8e · outbound

This paper cites CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.670721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.670721Z digest=sha256:e399b082e94bf317a75279a6730dc4418e9cf7c3dedce4d0b6ec34811d6dfa12

Observation 308bf2a7-06f6-41e5-b35c-9e8d88885f42 · outbound

This paper cites Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.679820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.679820Z digest=sha256:aa896cb9c4662665ce470a9c0f3d588e02804ae9863000fad1f7f42c2b23ff18

Observation 5a839a49-aa0d-45fb-a851-9a253daada16 · outbound

This paper cites Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.688548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.688548Z digest=sha256:db5b9b93abd767224ccf2c24a1c9cac46681faac57bf6cba42be8561ab8978ee

Observation 68879a51-1b9a-4caf-99e1-89967ab390f8 · outbound

This paper cites Learning to select like humans: Explainable active learning for medical imaging,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning to select like humans: Explainable active learning for medical imaging,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.692920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.692920Z digest=sha256:abdd4260fb949aee615fc7448b08b4719eac5774a3ea6d71b25c489c50150b20

Observation 5e0d7a4c-b915-49b8-803f-988056107507 · outbound

This paper cites Explainable Novel Category Discovery in Semantic Concept Space.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Explainable Novel Category Discovery in Semantic Concept Space

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.233533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.697321Z digest=sha256:97a8c4a015c8007ca546bb2702378dae8ceb07b2b849f0fce95ee9e1cf9945c7

Observation 1887b201-8937-4113-8006-45742510055d · outbound

This paper cites Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.188293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.702144Z digest=sha256:619e5bcc3c5f930d9e5aa795def8f68ca676fa2370d16c18f155d5665a17e992

Observation 785d0c0a-8fd3-48d4-8273-6b16a02a5c28 · outbound

This paper cites Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.706882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.706882Z digest=sha256:2b0f5954bf42aaf7ee5f6fd6405bf75941be9f95519c00bfb08c4b62a9afd19c

Observation 0eb09aea-aa89-4ab3-912a-63997d1f0ec2 · outbound

This paper cites Large dimensional analysis of cooperative multicell precoding with local individual csi,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large dimensional analysis of cooperative multicell precoding with local individual csi,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.715737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.715737Z digest=sha256:55e456b77c2babd70180ebf3756b1713150d586ab9305488fb949ed407e13a7f

Observation 76993563-aada-40e3-8a68-19ca3b42c1ad · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.720161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.720161Z digest=sha256:ecd5863f94f6c281aa672b21493e388ad7d1838836f6deea01671b6efc8cda29

Observation c152891f-6aea-4391-9770-f81c10223d9d · outbound

This paper cites Pytorch library for cam methods,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pytorch library for cam methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.724493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.724493Z digest=sha256:d1a25b85e58fa9516a5b6d042ee53919070c481636a85c65bcccc379742c67ee

Observation abf010bf-e3a3-4d15-b72c-98b4ee4e40eb · outbound

This paper cites ImageNet Large Scale Visual Recognition Challenge,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI ImageNet Large Scale Visual Recognition Challenge,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.729099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.729099Z digest=sha256:9de4d7f86386fa6dbaeda9044a5c010c97bc5d0d6215cc3dbe5724d5575015f3

Observation dee03cc4-5b09-42d9-a478-cde19bd6ed33 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.733515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.733515Z digest=sha256:f96c6f5f0194312a99b409ddcd149e3f341cacd3e74b2dd46fc566954a7d5ff3

Observation b840d73a-bd35-41ba-b90e-c89e2322738a · outbound

This paper cites BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.746918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.746918Z digest=sha256:4ef3dfc9d5764e040f965c8b8e386b04bc7491fba909148b7b7ce7feeab6ac6b

Observation a6c68623-1328-45ae-ab66-0afe03014e9f · outbound

This paper cites Training data-efficient image trans- formers &; distillation through attention,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Training data-efficient image trans- formers &; distillation through attention,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.751469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.751469Z digest=sha256:3ffc59147fd7b53c109f855d7ddf191ac7e5a778ac015dca29910e353580734c

Observation 696a8e51-6cb3-42bd-88a1-ce695928b7bc · outbound

This paper cites Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.755770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.755770Z digest=sha256:329ad7e2df06feeac0a33ec41994b11404ecce0e6f0c47f9d662114eaef610b9

Observation 6414c085-63b4-4f74-8954-2810626e624f · outbound

This paper cites Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.773034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.773034Z digest=sha256:cbe21733c1bc2c5fb06fb6bdeadaa36e62da3e1b627a5f2b3722fdcf6b0e68e9

Observation 82ef59c6-4706-47f0-a4d7-7dba82f80efd · outbound

This paper cites Full-gradient representation 18 for neural network visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Full-gradient representation 18 for neural network visualization,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.777847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.777847Z digest=sha256:7cfed82a7ba455d18411724434e72533995caa7ee5fd901b292a4ef7fc03ea66

Observation 5fed5e13-e4f3-403a-b455-b126fd1519d5 · outbound

This paper cites Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.782410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.782410Z digest=sha256:dfd73a464186ff4dbbf9ac25b5da3b77c36c1e96ff9df0374da279e37ebba143

Observation cd3ccd1c-3dec-4a3a-9b36-2839533aabb0 · outbound

This paper cites Learning Deep Features for Discriminative Lo- calization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning Deep Features for Discriminative Lo- calization ,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.786756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.786756Z digest=sha256:dd87908ff7900eb4e9406aef93d58ada6b0640c80e076852c1d3826ded5d4572

Observation e5ef305a-6724-4d60-857d-ce9c68ff429a · outbound

This paper cites Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.808587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.808587Z digest=sha256:aa58306d93eee6b4c4bdcfd2539e548272b0bb37cfef98f378bc594567ffaf2a

Observation 16c8297c-2f6b-4232-8c4e-fbe9ed8eba55 · outbound

This paper cites Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.826360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.826360Z digest=sha256:3bcf510a7632feffb4ac1b988eed76a54421f371bf828f97bce96da0bb5cb011

Observation 56931160-3718-46e4-a740-1a38aa352710 · outbound

This paper cites Enhancing prompt generation with adaptive refinement for camouflaged object detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing prompt generation with adaptive refinement for camouflaged object detection,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.843724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.843724Z digest=sha256:97b6f1b146cb343c0ff7fc14f799ec63a28623b9d5674948ae602c6862b7b59d

Observation be5f9b4b-445c-4075-b715-e6e5f63cc018 · outbound

This paper cites Quantifying attention flow in transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Quantifying attention flow in transformers,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.848312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.848312Z digest=sha256:7934014a5ca70e9d896ce8697297b364250a70cca9f28257bc162c4261cddaee

Observation 66fc7b3d-3284-4d49-90a0-16d4df5bba58 · outbound

This paper cites Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.854375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.854375Z digest=sha256:4cecaa7082650039c7c0fed842ff211c40c1e871859bb8dca016ca6ce91bfb96

Observation e7d02dc3-01c1-42c1-8479-e6d76c39c4dd · outbound

This paper cites the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.859102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.859102Z digest=sha256:6dafc7c2fd3c7bab8d3d71b1504f47a77c1a7618b63338ad848069e41fcaf661

Observation 11f0496e-6463-4aa2-8fbc-b70039ffd91b · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.863825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.863825Z digest=sha256:f04a93a6fe68a7e9dd050572a9e6ddb5d0ae5140652c26b1a7ba2b3e1e8a6b33

Observation b3a3ab5f-8b3b-4e34-9e52-235dae4b36c1 · outbound

This paper cites Rather than using an attention map as the attribution representation, the method operates on token-level feature activations.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than using an attention map as the attribution representation, the method operates on token-level feature activations

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.868882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.868882Z digest=sha256:9db27d68ec1ab400040936b23084e5b9dc1cee77d3acebf0b101240e530d27bf

Observation df55f346-798b-46c6-84f3-1a175cd1ff09 · outbound

This paper cites Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.873765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.873765Z digest=sha256:e153141167200f7f784c380933e2797115e9c3c24fbdf3162e31572f8691b09b

Observation 03d04be6-ae68-4e9e-96cd-af6d03164695 · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.879223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.879223Z digest=sha256:8ade9f9e7ab691e94bde2212f25d278f8079c4eeb372e48da8463a9ac576b1a2

Observation a7d66fd2-ced5-4a14-9c46-570c15028ffe · outbound

This paper cites Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.884275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.884275Z digest=sha256:2a6cfee4e9e2e8db97d2fb9466343d276b05fe3e025ce44c8521301890729275

Observation ea63f919-8c90-4164-b286-bae2091db538 · outbound

This paper cites matching.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI matching

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.889056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.889056Z digest=sha256:3de347f14ecf6ed3f88876b7257ba4a071e35fbf30fd80eeea17509a537f2878

Observation 896927b2-4c58-4e89-bab1-255af39476c8 · outbound

This paper cites A dog on a white bed.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A dog on a white bed

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.893576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.893576Z digest=sha256:0481baa76f792d2f737bd1522bbcf25d27b88f4a5b9eb3c972e70bb0b27a7d14

Observation 89e53fb0-a68d-40e6-9a66-da144f0ab390 · outbound

This paper cites Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.897806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.897806Z digest=sha256:0e4b8d8fe9755863b1a106b0f4aa7613f34e51f324572b5f51a63dee0561917c

Observation 54a8e770-36e0-4c28-96ff-64fa6d237f88 · outbound

This paper cites BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.901874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.901874Z digest=sha256:26117b0b0883cfee2f04b6beaa7c0f267b99ee7d96e87f5b4bc3cd945c686075

Observation d76fc51a-f104-4dbe-b7ad-4d69d5bbb578 · outbound

This paper cites Microsoft coco: Common objects in context,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Microsoft coco: Common objects in context,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.906013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.906013Z digest=sha256:554e568824e13215aeea06b75c6b0f85e458045175c640f60dda17c31636dbc2

Observation bcd781eb-9620-480d-9bd2-343bcd340632 · outbound

This paper cites Transformer in- terpretability beyond attention visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transformer in- terpretability beyond attention visualization,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.910505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.910505Z digest=sha256:15ae8f8a8d533677c8c80908ca7db64f5aa8e33204cde91de28a1e0e3a169f79

Observation 348bd8db-0aa6-4056-a8e9-124e79fe93b7 · outbound

This paper cites Transreid: Transformer-based object re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transreid: Transformer-based object re-identification,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.914869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.914869Z digest=sha256:8c48f098bebf01331908e69a922c47d12e18fe9d179718257f3fc39880356861

Observation e30f169d-cf73-4ba5-a336-278ebc76c72e · outbound

This paper cites Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.918840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.918840Z digest=sha256:ce799296ad56c82d47cd76317d57c0cbe8bae2e669ea45ff00c86b062a2a846d

Observation 572b76d5-12af-4075-b5b9-679acd5c8875 · outbound

This paper cites Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.923344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.923344Z digest=sha256:f3db226ba5b0ba2f243dab12443baf25e712dfdc98d2bb4018510a0d871a5b63

Observation 529da6a4-e53c-4da3-b876-01c18d88157d · outbound

This paper cites Analogous to evolutionary algorithm: Designing a unified sequence model,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Analogous to evolutionary algorithm: Designing a unified sequence model,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.927583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.927583Z digest=sha256:ffde972c131aa8283c4c83923d361a048719a21e6468da342e4d391135d0bfa8

Observation 3332e09e-68ef-453d-974c-e65a0d2ab1b4 · outbound

This paper cites Passive attention in artificial neural networks predicts human visual selectivity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Passive attention in artificial neural networks predicts human visual selectivity,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.932385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.932385Z digest=sha256:77d5e7665baa1820e18080e88e1cb0bd39d1bc3bef16aee636448b86d7174080

Observation f616e9ae-43fe-4b85-9ff3-fcc8cee7ec56 · outbound

This paper cites Vitae: Vision transformer advanced by exploring intrinsic inductive bias,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vitae: Vision transformer advanced by exploring intrinsic inductive bias,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.936914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.936914Z digest=sha256:e27d0703b154a74126e8dec7cf47ede34eb2f9e3db28acf2fc1248d3d2dcf956

Observation 7a6cbe1b-3acb-4ea5-aedd-678e82902f72 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before fuse: Vision and language representation learning with momentum distillation,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.941698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.941698Z digest=sha256:7d2c763c26a5121b8664a2a4e3735dec73557f6d88c3bebad475251b5ca60d46

Observation 7d534b61-812d-4bd6-8096-e1accd4878e1 · outbound

This paper cites VLMAE: Vision-Language Masked Autoencoder.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI VLMAE: Vision-Language Masked Autoencoder

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.128574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.946360Z digest=sha256:2a1a29ecf6600c84db063ca8e9daf675145b22083ee69ddd5fcea5b298f7df38

Observation e90b8372-757a-4350-af5f-4898e7a6a024 · outbound

This paper cites A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.951037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.951037Z digest=sha256:ec4ce8d1bde3eeed4be2d19df167bdce6215a3e7a979f2ec05b7b8b127815543

Observation d0d4f501-0a2b-4007-83e3-1ccef1a65f85 · outbound

This paper cites A challenging benchmark of anime style recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A challenging benchmark of anime style recognition,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.955330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.955330Z digest=sha256:2f2f17a9c90dc2b82b7495a8ba1979dea35d0967122bf9e820c990aa0b92e706

Observation 7ec2ad21-43de-4333-99d6-ba9ecface9a1 · outbound

This paper cites Metaformer is actually what you need for vision,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Metaformer is actually what you need for vision,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.960419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.960419Z digest=sha256:ee8484dc1e2c7f609fb981e50ab4e3e64f3337179796b14dd9bb72659187fb4d

Observation 0869e2ca-474f-46aa-bf2b-667d754c1d32 · outbound

This paper cites Delving deep into the generalization of vision transformers under distribution shifts,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving deep into the generalization of vision transformers under distribution shifts,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.965280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.965280Z digest=sha256:3e57a8443e511aa33fa5ddb7ac2cc69741694ec0386d83235b8d096cdc993fc7

Observation 3992634d-9567-458d-893d-75ca64e133cb · outbound

This paper cites General facial representation learning in a visual- linguistic manner,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI General facial representation learning in a visual- linguistic manner,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.970138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.970138Z digest=sha256:8051d7be6c5cf131d514a6e5132e3409c60d91039603db7a4c90cf093c90c39e

Observation ce334212-4a11-43ec-bcb4-c49ea2364d55 · outbound

This paper cites Multi-modal alignment using representa- tion codebook,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal alignment using representa- tion codebook,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.974492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.974492Z digest=sha256:c430bb18a0fe7b1f243757600ba8fd69f3b621878eccc11f9a571d2de1c33adc

Observation 136ae6d9-9f8c-473a-9a0c-572f10a23b29 · outbound

This paper cites Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.979504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.979504Z digest=sha256:ed30c4e0704fbadcf2b1b6766fc725ed96059c6d2c7841bb2ea2d9e56fe8deab

Observation 16f7963a-715b-4d2f-800a-927664864f8c · outbound

This paper cites Inception transformer,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Inception transformer,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.983594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.983594Z digest=sha256:10b73e0c3a853c43c88891a9dd45585612585f1f6a4629c7d361bf043ae6b69b

Observation e6bb17c9-8e60-480f-b346-9f91e16eba1c · outbound

This paper cites Delving into sequential patches for deep- fake detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving into sequential patches for deep- fake detection,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.988241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.988241Z digest=sha256:3ab56f483b3932fd795a4dbd5b6924d71bd19f12662642d145d5d67c94c477bb

Observation 221817f8-c088-4d15-a179-8aac556bfc75 · outbound

This paper cites Adversarial normalization: I can visualize everything (ice),.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Adversarial normalization: I can visualize everything (ice),

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.992974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.992974Z digest=sha256:90a14f03830c1860e91275c3e2f4f1a6db127ee2490b320d1462f40226343fec

Observation 435411ee-05f4-4aa7-b21d-b796ff46f1e4 · outbound

This paper cites A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.997910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.997910Z digest=sha256:aaf9131eed3607c4cb09083e21ad10d29ccc6758bfbfec32aee23a023f018eea

Observation 578198fe-54c4-47d8-8a41-bb65e2d10af6 · outbound

This paper cites Selfme: Self-supervised motion learning for micro- expression recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Selfme: Self-supervised motion learning for micro- expression recognition,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.002177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.002177Z digest=sha256:1a6a2009f4e2a1f52ba1fe397a2979a31e19a44b5e1efb7346e4da6e08f3e2a0

Observation 2eb616b1-5ff8-4cde-9aa5-a6fcc5828b3b · outbound

This paper cites Marlin: Masked autoencoder for facial video representation learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Marlin: Masked autoencoder for facial video representation learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.006873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.006873Z digest=sha256:c49b604092d67a258f622912189c7b5362b9b8991f1941cccb8d2da4b3c1c820

Observation 00360c11-ac10-4e8c-a3e0-a892efc24898 · outbound

This paper cites Blackvip: Black-box visual prompting for robust transfer learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Blackvip: Black-box visual prompting for robust transfer learning,

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.011351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.011351Z digest=sha256:4c19c364f9b2a9657e86fba145affcc6f9a363cf0cbc28726a5cbbf76cdd6f42

Observation b51f83f7-e2d6-43a8-b78d-3ca76bfc2530 · outbound

This paper cites To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.015441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.015441Z digest=sha256:9866999b76a26a424619715fe28e771c9e976324e9e39a91d1bd2695122081ae

Observation 50068a82-55d8-46f2-befa-fd67dd965d89 · outbound

This paper cites Boost vision trans- former with gpu-friendly sparsity and quantization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boost vision trans- former with gpu-friendly sparsity and quantization,

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.019933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.019933Z digest=sha256:66ce5462f295ba6a88861872c45589992eb8466c28548399bffb50371ec2a664

Observation 829603f1-bf1b-4491-b10e-32f38c1bceec · outbound

This paper cites Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.024488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.024488Z digest=sha256:d491ae506c2ec080ea71bea2eb553d4177f8776ff8f12629e6c1aaec05d3010a

Observation f313bd3d-1844-4a9a-bd85-3d2672d52b36 · outbound

This paper cites Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.029301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.029301Z digest=sha256:66c4daf254253a420e7c9a1e82b395493d152033273ddb77bdbd6759e8467c2f

Observation 04a60b94-805b-4d54-98f9-74f3c0d9738b · outbound

This paper cites Vision trans- formers with mixed-resolution tokenization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision trans- formers with mixed-resolution tokenization,

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.034092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.034092Z digest=sha256:050e246fae8f28d4f038e3b552f02a45b33fc56fed3163ce2657d1e469fec369

Observation bf297aea-97c5-4447-9571-90ed5b5de311 · outbound

This paper cites D3former: Debiased dual distilled transformer for incremental learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI D3former: Debiased dual distilled transformer for incremental learning,

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.039180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.039180Z digest=sha256:143b3c6a9717b3f113d14d6603bbdeb0498623432060bfc1ca35c47e491aa473

Observation 6c8475f5-969f-4ffa-9351-3c69d53c8376 · outbound

This paper cites Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.043469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.043469Z digest=sha256:19db54f5712183d7521b384ccc413506130c4730206def4359decf7d8e68dda1

Observation 55df5f88-0300-4922-b572-9723c58761fc · outbound

This paper cites Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.047846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.047846Z digest=sha256:57d81963ca0bda6beb98d09ef3e33be381f6413b96102e68613a7dc3c195cb0b

Observation ceba726d-42b0-4a9d-8219-f66a38dffa55 · outbound

This paper cites Masked autoencoding does not help natural language supervision at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Masked autoencoding does not help natural language supervision at scale,

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.052256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.052256Z digest=sha256:6491b9c69b21f8973609c19ec6c77c12a3baefc81e9bddedad462da1d3a27a18

Observation c2bc21db-f7a6-4764-8f71-0d34ef5184d3 · outbound

This paper cites Vilem: Visual- language error modeling for image-text retrieval,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vilem: Visual- language error modeling for image-text retrieval,

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.057254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.057254Z digest=sha256:c1ecb88f327b80476d85e751d918efe07a9e70ff5ffa53ee3117e5ffe074da85

Observation 2da5988c-6dde-40e8-9825-d48802689472 · outbound

This paper cites Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.061757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.061757Z digest=sha256:a43c86d7c1018f36c0f44c524f04802a663cc50e5e0e1cc3967b943b4285906f

Observation a62b514b-987f-4434-bd4f-ac0f10068217 · outbound

This paper cites Zero-shot referring image segmentation with global-local context features,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Zero-shot referring image segmentation with global-local context features,

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.068004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.068004Z digest=sha256:ac61a79b2eeb393c9a10f31d9b3d727864f65a46c596130e44963fc6849b7c08

Observation c32fa6b7-ffff-4f1e-8ccc-54d1715fb53a · outbound

This paper cites Improving visual grounding by encouraging consistent gradient-based explanations,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving visual grounding by encouraging consistent gradient-based explanations,

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.072944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.072944Z digest=sha256:5b74182d1bbe85b949e808dafaf2f299b3cdcea6b9f348d9e02b86a1f57df2ca

Observation 7b42af4c-c1e8-401f-b212-beadef5ce0a3 · outbound

This paper cites Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.077543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.077543Z digest=sha256:e77fa4dfff1e610eddc77184beadc11ea33f91d8374912be43be3b327e380c18

Observation e73f4ac9-9245-4cf8-8ac6-abe073ded2e0 · outbound

This paper cites Multi-modal representation learn- ing with text-driven soft masks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal representation learn- ing with text-driven soft masks,

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.082156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.082156Z digest=sha256:eccc0064fb304fc530490b8a14fb084a5c2730c9a931cefc96b50b3ca77d4a20

Observation 095db75f-298b-4a0f-98ad-8b46602f274c · outbound

This paper cites From images to textual prompts: Zero-shot visual question answering with frozen large language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI From images to textual prompts: Zero-shot visual question answering with frozen large language models,

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.086478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.086478Z digest=sha256:92f4940fd0162316075bfccfa1c5a4c486e52e37a3e7c4496e83fda7ce95021e

Observation 5c5c3970-f154-4a40-aaf0-cbb4ac470c61 · outbound

This paper cites Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.091386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.091386Z digest=sha256:525b33043ef4c323855edbde5523e9d691c4547e4b71bfae0b389e2eab39062c

Observation ee4af320-2763-4d38-ae43-a9043e532d59 · outbound

This paper cites Semantic information in contrastive learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semantic information in contrastive learning,

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.095850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.095850Z digest=sha256:ca0a2e3fdaf16e1ce0ad5459ec7af8e744657353b5ea01286a07ecd01517381f

Observation fe2cf79a-731b-4777-9221-70e8dabcd219 · outbound

This paper cites Smmix: Self-motivated image mixing for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Smmix: Self-motivated image mixing for vision transformers,

Reference 120

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.100616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.100616Z digest=sha256:5f19abe9629ce34f1a90956c0f8de806bbc6926349fdf925c8296edc5fc9c450

Observation 0257e80a-6b6a-4220-a954-846725895a96 · outbound

This paper cites Cose: A consistency- sensitivity metric for saliency on image classification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Cose: A consistency- sensitivity metric for saliency on image classification,

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.105981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.105981Z digest=sha256:95c94bb4317fa513900bfba183fa3f2826f68d6a4a0b75ab5186349ad5d23cd2

Pith citing papers

No inbound Pith citation observations are available.