Pith. sign in

Paper Citation Record · LEDGER

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI

As of 9 August 2026, this Paper Citation Record lists 100 of 251 outbound references and 0 inbound Pith citation observations for arXiv:2608.05258.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05258 v1

Coverage vector

measured 100 of 251 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:59:11.105981Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 251 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved96
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d11799d7-cbac-4c50-a9c6-c44c89d41f55 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.556089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.556089Z digest=sha256:f650ac84bb2471752c059b004c4707261567088c8a7ac67922b139f8d5ccceb8

Observation 6f4bef24-111c-4059-9ebc-51ec3ae2795f · outbound

This paper cites Representation learning and na- ture encoded fusion for heterogeneous sensor networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Representation learning and na- ture encoded fusion for heterogeneous sensor networks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.561346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.561346Z digest=sha256:2106134b2f82a561a9375b2435c6cc624a4988d313a320e32c8a5573b0990022

Observation 0e2d9a63-413d-4eb4-9378-4d3f3b328dee · outbound

This paper cites Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.565742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.565742Z digest=sha256:f69f5d7168658ddd39350455820e21e38007c8205d91c6ff50019e3f4c68d545

Observation 165e5521-1a09-4255-b96d-1b3629368d1d · outbound

This paper cites Enhanced robustness by symmetry enforcement,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhanced robustness by symmetry enforcement,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.570437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.570437Z digest=sha256:97de0b62feffa8ac58deec3122fd88def955f3ec8ecf0c3a1fd197b4fdae432d

Observation 0a432fff-c026-428e-b74b-259d3a45497a · outbound

This paper cites Partial interference alignment for heterogeneous cellular networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Partial interference alignment for heterogeneous cellular networks,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.575167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.575167Z digest=sha256:6333f3a1bd7547a654d98bf1bd4618b5003b7bdaf4ffa3c29daa80de923e38e1

Observation d61ebdc0-a3e0-4c47-9539-b20da33af6c3 · outbound

This paper cites Optimization for user centric massive mimo cell free networks via large system analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Optimization for user centric massive mimo cell free networks via large system analysis,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.580043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.580043Z digest=sha256:705c7a39bad7e8ae3a4f4fe723e04f7d3018b9072acc0d6400faba1bf95be9cb

Observation 2ee2b689-e79e-4419-9d9c-87ebb4c5be66 · outbound

This paper cites Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.589472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.589472Z digest=sha256:427d6b6052ff7cfdd3da1ab67d1c13fb0e09bfae66b1a28a3884b468b9bfa7ff

Observation 9fa2d7cb-02eb-49cc-9232-b08d59763459 · outbound

This paper cites Deep reinforcement learning based computation offloading for mobility-aware edge computing,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Deep reinforcement learning based computation offloading for mobility-aware edge computing,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.593676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.593676Z digest=sha256:6bc5739844c53000d60fbb6259b8f80017272983d82bb94572aba0e891a62a55

Observation e4dffcba-d670-4690-a55a-66053719d781 · outbound

This paper cites Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.598216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.598216Z digest=sha256:c56a40929e60dac1dd4a32e6de49ab654c756a8193fa5523b66abb4e047917cd

Observation 436732dc-d1bc-4b90-83ff-e0eb42420b21 · outbound

This paper cites Low complexity optimization for user centric cel- lular networks via large dimensional analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Low complexity optimization for user centric cel- lular networks via large dimensional analysis,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.602562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.602562Z digest=sha256:9790ce1f1faf879f2e743faf6b3b1fdc8da776176aa7fd6e4767242c64879925

Observation 080326ea-9f1c-40ed-a3a6-f66ce1a212ab · outbound

This paper cites Improving robustness of deep neural networks via large-difference transformation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving robustness of deep neural networks via large-difference transformation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.607014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.607014Z digest=sha256:b1cecd979d927bc2ba926c556ffc8d639cf1b29d3282377b2acf30b59d48234f

Observation c0a28969-e207-40c6-b238-b9141d7a21a9 · outbound

This paper cites Looking beyond content: Modeling and detection of fake news from a social context perspective.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Looking beyond content: Modeling and detection of fake news from a social context perspective

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.611370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.611370Z digest=sha256:3cf1474719486f948c226fb0e8defd8ce05e2ed151e7aa9bbfda47020789e82f

Observation b70425f5-9a7f-40fa-a5f3-fc1c4fc27c6d · outbound

This paper cites Large system analysis for densification of cellular networks with massive mimo,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large system analysis for densification of cellular networks with massive mimo,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.615627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.615627Z digest=sha256:e78803ac4cc764fe05c4ef5c80ac49f886c0a5161e390e47b09eb65f70261853

Observation a73bec4b-fc33-4bd0-ae8e-aa38cba8ae2e · outbound

This paper cites Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.620148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.620148Z digest=sha256:8ab0a5e9347e485dddebbe3cc8d7a479c0344e9cf86d4b992f0c382836c191a1

Observation 4187ea7e-2c05-4ca1-aa10-17fbc28ccd8d · outbound

This paper cites Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.624519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.624519Z digest=sha256:fdba4946d78dd3d487afd8b2760a2cdc1e9badb2417322bfed7679a83cd9a911

Observation 8a2fd787-2e48-4a4b-9cd7-e89decb85427 · outbound

This paper cites Information theory and represen- tation learning inspired multimodal data fusion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Information theory and represen- tation learning inspired multimodal data fusion,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.629428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.629428Z digest=sha256:3569b114a40cf5c7e314d36f5c0cb9f6029c9746280a4a05f0e98dbe94278ad8

Observation 560e8a2e-44dc-4470-a3d9-6adbcdc5c858 · outbound

This paper cites Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.286699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:59:10.638625Z digest=sha256:68f22b8af16ac2e396af2aad756785d5df98dd53e2400285aa5ae7537754d8b6

Observation 5bf5c5cb-b07f-465f-b50e-d916ce2e0b23 · outbound

This paper cites Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.647891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.647891Z digest=sha256:819a6a1a490595c93a38dd5a44ee99d76368fb9cd536ac087069f15e28ce429e

Observation fa094813-9e56-4e25-adf3-58b74e9fbf87 · outbound

This paper cites Expert-guided ex- plainable few-shot learning for medical image diagnosis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning for medical image diagnosis,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.652271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.652271Z digest=sha256:a1513e9d568ea5dd92d71db851053512b3cdaadeaf6acd885b0806cc670ae764

Observation ebb2d243-fe70-4e43-9158-7b3e6fd6f25b · outbound

This paper cites GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.656692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.656692Z digest=sha256:93dc27a0e59c7e4211e1da6f429e3cbee31850f0057c32c068463267866f1359

Observation 0a9a379b-e784-4fb8-8ece-6d18600a778f · outbound

This paper cites Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.661437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.661437Z digest=sha256:81ad2319060acb973d60272e8a7215da009ceaf6372e29fd725727fe935f7d50

Observation 2c2acf1c-776b-434b-ae53-208c6063aa2d · outbound

This paper cites Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.665772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.665772Z digest=sha256:a44a02ba975c7542f393e6f376d265b8b7889b72661f9ab88a03b95a7410c349

Observation 17e6e87b-f390-4c29-88ef-d4a356b0cd8e · outbound

This paper cites CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.670721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.670721Z digest=sha256:fd7c7178fafbbce6e7c685098cf6b3d4699e5bf0a1f8f8ab81e822c90a158a09

Observation 308bf2a7-06f6-41e5-b35c-9e8d88885f42 · outbound

This paper cites Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.679820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.679820Z digest=sha256:eb043ee308f9111af4c7fcd7992c4cde0850e636436f97bfe43012b88e1f516d

Observation 5a839a49-aa0d-45fb-a851-9a253daada16 · outbound

This paper cites Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.688548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.688548Z digest=sha256:bc1c5db51e00816974507ffe94ba3d2ec92313a0535568fc810b0eb90f152b2d

Observation 68879a51-1b9a-4caf-99e1-89967ab390f8 · outbound

This paper cites Learning to select like humans: Explainable active learning for medical imaging,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning to select like humans: Explainable active learning for medical imaging,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.692920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.692920Z digest=sha256:0ea2c229be4d3f5b8ca8e0fb5b7e5ebb10e5769d3a87611597411e0d6a473918

Observation 5e0d7a4c-b915-49b8-803f-988056107507 · outbound

This paper cites Explainable Novel Category Discovery in Semantic Concept Space.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Explainable Novel Category Discovery in Semantic Concept Space

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.233533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:59:10.697321Z digest=sha256:61cd114ba5a32e0282b1bc78356afa5c0c0a0df413a61933bc3152508729ca47

Observation 1887b201-8937-4113-8006-45742510055d · outbound

This paper cites Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.188293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:59:10.702144Z digest=sha256:3ee54f0af628f709f20e4c8894d16030e2f6362b2e25ecc96dc9ebc7b4332a6e

Observation 785d0c0a-8fd3-48d4-8273-6b16a02a5c28 · outbound

This paper cites Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.706882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.706882Z digest=sha256:447c923961f7d82141f299ad56c262c2cdab9780b85c9a31481ac26ed23671dc

Observation 0eb09aea-aa89-4ab3-912a-63997d1f0ec2 · outbound

This paper cites Large dimensional analysis of cooperative multicell precoding with local individual csi,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large dimensional analysis of cooperative multicell precoding with local individual csi,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.715737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.715737Z digest=sha256:38bbb7e45a64655a3f8b18ca07554aa3a6183b9b3f37ff6589ccde220767d2f9

Observation 76993563-aada-40e3-8a68-19ca3b42c1ad · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.720161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.720161Z digest=sha256:3a11e69af78cb26f70a6a0b3c087db8089ceb044bb46842c4922552fd3aded0a

Observation c152891f-6aea-4391-9770-f81c10223d9d · outbound

This paper cites Pytorch library for cam methods,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pytorch library for cam methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.724493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.724493Z digest=sha256:20d129fa985db3e2f896878532be5e25fed2c93ee86a02fd762f8f286669f961

Observation abf010bf-e3a3-4d15-b72c-98b4ee4e40eb · outbound

This paper cites ImageNet Large Scale Visual Recognition Challenge,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI ImageNet Large Scale Visual Recognition Challenge,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.729099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.729099Z digest=sha256:bc801fc7f39495347c1f7253dca6cece73d7876310c6e59001ce819d8321b8ef

Observation dee03cc4-5b09-42d9-a478-cde19bd6ed33 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.733515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.733515Z digest=sha256:9ea10287193145bb7434364b4c36d16376e521d5146fc81e20032944e63c7877

Observation b840d73a-bd35-41ba-b90e-c89e2322738a · outbound

This paper cites BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.746918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.746918Z digest=sha256:210e51d534a5102c1342a93cd6e5ce1fd07c7fb41d49c700afb47a342f0a413c

Observation a6c68623-1328-45ae-ab66-0afe03014e9f · outbound

This paper cites Training data-efficient image trans- formers &; distillation through attention,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Training data-efficient image trans- formers &; distillation through attention,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.751469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.751469Z digest=sha256:810dd64a94d47fcd1aa3702ea96f5ea7ecf1cee7a628098a3060cab39cbef1f0

Observation 696a8e51-6cb3-42bd-88a1-ce695928b7bc · outbound

This paper cites Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.755770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.755770Z digest=sha256:dd8be66a0a0d1243b67a67a5b81e431b13e774dcf4a90bf6dec1111753171d3f

Observation 6414c085-63b4-4f74-8954-2810626e624f · outbound

This paper cites Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.773034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.773034Z digest=sha256:73efdd38032870e66acf2b5bf0b45292416cc81d265aed2a2e988dcaee312af0

Observation 82ef59c6-4706-47f0-a4d7-7dba82f80efd · outbound

This paper cites Full-gradient representation 18 for neural network visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Full-gradient representation 18 for neural network visualization,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.777847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.777847Z digest=sha256:25e40cbca45f6c2dab4bd507a20dc8c810b9c1da7adb96cf292d7ffa4bcf0d77

Observation 5fed5e13-e4f3-403a-b455-b126fd1519d5 · outbound

This paper cites Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.782410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.782410Z digest=sha256:124e5cded063f55676989d0b786826300ffab135f0f2a59fb009685cadbd67a0

Observation cd3ccd1c-3dec-4a3a-9b36-2839533aabb0 · outbound

This paper cites Learning Deep Features for Discriminative Lo- calization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning Deep Features for Discriminative Lo- calization ,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.786756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.786756Z digest=sha256:9d5328f5e6d5157f645b6e7880df883a43ec562c621dd36c99e617eea4731439

Observation e5ef305a-6724-4d60-857d-ce9c68ff429a · outbound

This paper cites Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.808587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.808587Z digest=sha256:c1bd3072640592bfb5d06aa16d069eb256d64c7a4fee6d48c923dc28be926d59

Observation 16c8297c-2f6b-4232-8c4e-fbe9ed8eba55 · outbound

This paper cites Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.826360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.826360Z digest=sha256:8014963382325a8150ec7e06212c138f70ce041e9874a3b87c191f950083a56a

Observation 56931160-3718-46e4-a740-1a38aa352710 · outbound

This paper cites Enhancing prompt generation with adaptive refinement for camouflaged object detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing prompt generation with adaptive refinement for camouflaged object detection,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.843724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.843724Z digest=sha256:0d9af92f64d2e66ccafbeb463536f3ab6e3af137aab56539e71023b660a8faef

Observation be5f9b4b-445c-4075-b715-e6e5f63cc018 · outbound

This paper cites Quantifying attention flow in transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Quantifying attention flow in transformers,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.848312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.848312Z digest=sha256:2c6e8290b8c4717bf92d3e74f3585c92499c2fe5ca7859b77fe1ed91b6fb44fc

Observation 66fc7b3d-3284-4d49-90a0-16d4df5bba58 · outbound

This paper cites Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.854375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.854375Z digest=sha256:beabb23cdfc70a0b0b583fce64c3341bf763051cc40bee3fd557ec8ceee53b3e

Observation e7d02dc3-01c1-42c1-8479-e6d76c39c4dd · outbound

This paper cites the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.859102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.859102Z digest=sha256:638e7b3fe59853bf828ec7a64d88270978640f47dfe24a1e62b757a7a7766624

Observation 11f0496e-6463-4aa2-8fbc-b70039ffd91b · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.863825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.863825Z digest=sha256:430c5cfb1a56fdf6c50dae18430d25b972f498990ff21b8a9b20bcb40d54d788

Observation b3a3ab5f-8b3b-4e34-9e52-235dae4b36c1 · outbound

This paper cites Rather than using an attention map as the attribution representation, the method operates on token-level feature activations.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than using an attention map as the attribution representation, the method operates on token-level feature activations

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.868882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.868882Z digest=sha256:c3a2df9fce5ebefe78a7a608a5857884277dacf663e764d5538b3f23de50b803

Observation df55f346-798b-46c6-84f3-1a175cd1ff09 · outbound

This paper cites Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.873765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.873765Z digest=sha256:98a90d0d41037b85dbefeb79bf5afc237a5200284ca30aa6aec1dc37fe9b3110

Observation 03d04be6-ae68-4e9e-96cd-af6d03164695 · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.879223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.879223Z digest=sha256:b8493c43b620d0385ada9d8551ab1cae4a31fcfd912338273092594fdc69e530

Observation a7d66fd2-ced5-4a14-9c46-570c15028ffe · outbound

This paper cites Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.884275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.884275Z digest=sha256:8d9c7c00052bb328a1f41ed14a0078467d45931e3983e45cc2aadb91d7714569

Observation ea63f919-8c90-4164-b286-bae2091db538 · outbound

This paper cites matching.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI matching

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.889056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.889056Z digest=sha256:2b0e881e97a34e2660d3d65f547c8e32a2c06a418de4cebddbcd5c4e7c8c1b9e

Observation 896927b2-4c58-4e89-bab1-255af39476c8 · outbound

This paper cites A dog on a white bed.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A dog on a white bed

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.893576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.893576Z digest=sha256:aa003a214b5f286809b331492a8d7e9bf4872f4517e1e61afea3ff081f0f6928

Observation 89e53fb0-a68d-40e6-9a66-da144f0ab390 · outbound

This paper cites Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.897806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.897806Z digest=sha256:3062049d54f7e372de32ce1b9f6e18492eba21e6f09c95f48ba3572c0dd72974

Observation 54a8e770-36e0-4c28-96ff-64fa6d237f88 · outbound

This paper cites BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.901874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.901874Z digest=sha256:a9e1cb9f07db5ade432ea40c0d124ad79815672ddfaf1d1944883b9a47cc5fe7

Observation d76fc51a-f104-4dbe-b7ad-4d69d5bbb578 · outbound

This paper cites Microsoft coco: Common objects in context,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Microsoft coco: Common objects in context,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.906013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.906013Z digest=sha256:c5acdc47bb735d5ed243bb04f97ec47177faaafbfbd18273d929b3a39895b65d

Observation bcd781eb-9620-480d-9bd2-343bcd340632 · outbound

This paper cites Transformer in- terpretability beyond attention visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transformer in- terpretability beyond attention visualization,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.910505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.910505Z digest=sha256:0e5d1a691b8c7b412b9e6e2e3e1b7cbff950a4c82453ee62c534868c2a82c1a1

Observation 348bd8db-0aa6-4056-a8e9-124e79fe93b7 · outbound

This paper cites Transreid: Transformer-based object re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transreid: Transformer-based object re-identification,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.914869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.914869Z digest=sha256:0c610b7d809d165c023b9189a5e73ecc0e810c0cc38876717968f673cd3a986b

Observation e30f169d-cf73-4ba5-a336-278ebc76c72e · outbound

This paper cites Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.918840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.918840Z digest=sha256:c42e6337262adcf6ae635b7b7d17279f8a7443b1dcd33bea276eff9eb253e2c1

Observation 572b76d5-12af-4075-b5b9-679acd5c8875 · outbound

This paper cites Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.923344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.923344Z digest=sha256:168c9ac21af767b2cfa421beae596fca4cefe9eacc82cb153916f80efe51267a

Observation 529da6a4-e53c-4da3-b876-01c18d88157d · outbound

This paper cites Analogous to evolutionary algorithm: Designing a unified sequence model,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Analogous to evolutionary algorithm: Designing a unified sequence model,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.927583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.927583Z digest=sha256:55cfffd557dd9654519717868db0d2be0bb3dae0f9ad9025376d948f28e6e412

Observation 3332e09e-68ef-453d-974c-e65a0d2ab1b4 · outbound

This paper cites Passive attention in artificial neural networks predicts human visual selectivity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Passive attention in artificial neural networks predicts human visual selectivity,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.932385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.932385Z digest=sha256:703a5aea91127edb531fed134e110893b9744d53406b3167a901dcd6d3e36ae2

Observation f616e9ae-43fe-4b85-9ff3-fcc8cee7ec56 · outbound

This paper cites Vitae: Vision transformer advanced by exploring intrinsic inductive bias,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vitae: Vision transformer advanced by exploring intrinsic inductive bias,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.936914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.936914Z digest=sha256:fdbb805bdc820f64b265d0ffc31a6bfc9e49e0d6dd8fc546de3e4ecfdcd78c6f

Observation 7a6cbe1b-3acb-4ea5-aedd-678e82902f72 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before fuse: Vision and language representation learning with momentum distillation,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.941698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.941698Z digest=sha256:129272244a3e1b666bbc45caea7fcd508056d39e5b7bae95822e13138d9df389

Observation 7d534b61-812d-4bd6-8096-e1accd4878e1 · outbound

This paper cites VLMAE: Vision-Language Masked Autoencoder.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI VLMAE: Vision-Language Masked Autoencoder

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.128574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T16:59:10.946360Z digest=sha256:4c80c15379b1348c999e3bd9cd67b44724433db2bd311f259d8a62c1bcc1b84f

Observation e90b8372-757a-4350-af5f-4898e7a6a024 · outbound

This paper cites A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.951037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.951037Z digest=sha256:db1a99e20c3fd049b5cd79d860311a7bda391c9ba8cc2fbf8b05fa388adc0511

Observation d0d4f501-0a2b-4007-83e3-1ccef1a65f85 · outbound

This paper cites A challenging benchmark of anime style recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A challenging benchmark of anime style recognition,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.955330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.955330Z digest=sha256:9c53996c3ab5a07c1b0fa2c83d10674f7a266594c1e4b89446540aee8f44b3a2

Observation 7ec2ad21-43de-4333-99d6-ba9ecface9a1 · outbound

This paper cites Metaformer is actually what you need for vision,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Metaformer is actually what you need for vision,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.960419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.960419Z digest=sha256:46b002ab946a027ea7f966dcf38f5bb7ee95940be681790c315819303c99d0bc

Observation 0869e2ca-474f-46aa-bf2b-667d754c1d32 · outbound

This paper cites Delving deep into the generalization of vision transformers under distribution shifts,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving deep into the generalization of vision transformers under distribution shifts,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.965280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.965280Z digest=sha256:5f4ae9eea045105d4cb4b300256ad9e91a6a0df4880254405b09fc244badffe9

Observation 3992634d-9567-458d-893d-75ca64e133cb · outbound

This paper cites General facial representation learning in a visual- linguistic manner,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI General facial representation learning in a visual- linguistic manner,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.970138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.970138Z digest=sha256:91bff1a387a4cf52456ae7bdfcb61fc0bfc3e194d6bc927cec5eaf7b2f05b09c

Observation ce334212-4a11-43ec-bcb4-c49ea2364d55 · outbound

This paper cites Multi-modal alignment using representa- tion codebook,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal alignment using representa- tion codebook,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.974492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.974492Z digest=sha256:b1914074c63e7c2b6e2da930546f7250df30a3aa15530e35dad32f5eb4402e63

Observation 136ae6d9-9f8c-473a-9a0c-572f10a23b29 · outbound

This paper cites Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.979504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.979504Z digest=sha256:7929931620fe1aac6b24cb1b11ca094347102d625fa65e90c8d0540ee1bf476d

Observation 16f7963a-715b-4d2f-800a-927664864f8c · outbound

This paper cites Inception transformer,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Inception transformer,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.983594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.983594Z digest=sha256:3910ae72e1d508c2615101a973644e2c65f26a265c88b0c088186da5231fd779

Observation e6bb17c9-8e60-480f-b346-9f91e16eba1c · outbound

This paper cites Delving into sequential patches for deep- fake detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving into sequential patches for deep- fake detection,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.988241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.988241Z digest=sha256:87e85f3f69ba619d1561501f37856c9a82b3e917ef15d53c9b2dc872330d1e90

Observation 221817f8-c088-4d15-a179-8aac556bfc75 · outbound

This paper cites Adversarial normalization: I can visualize everything (ice),.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Adversarial normalization: I can visualize everything (ice),

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.992974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.992974Z digest=sha256:c9d676345483e37a49986d9f1ebbe93087288654fe6d94514aeb5e214208836f

Observation 435411ee-05f4-4aa7-b21d-b796ff46f1e4 · outbound

This paper cites A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.997910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.997910Z digest=sha256:0c4caad8e9d32b25121fef14a368291cb906f14b32c3bd35acfa4e5c87e8bd6d

Observation 578198fe-54c4-47d8-8a41-bb65e2d10af6 · outbound

This paper cites Selfme: Self-supervised motion learning for micro- expression recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Selfme: Self-supervised motion learning for micro- expression recognition,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.002177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.002177Z digest=sha256:a999fe672a83d0910eab2bd17e6bb5bb9a345c0f41b113673b4c938d5b11b4d5

Observation 2eb616b1-5ff8-4cde-9aa5-a6fcc5828b3b · outbound

This paper cites Marlin: Masked autoencoder for facial video representation learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Marlin: Masked autoencoder for facial video representation learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.006873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.006873Z digest=sha256:dfca911144fd6a1a5cdf89ee55083ddd7e50f61872e46dee0de71ea6c102b58e

Observation 00360c11-ac10-4e8c-a3e0-a892efc24898 · outbound

This paper cites Blackvip: Black-box visual prompting for robust transfer learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Blackvip: Black-box visual prompting for robust transfer learning,

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.011351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.011351Z digest=sha256:582a2e5f2a4b67a0b7dde89e117f4c9282a0709fe3220ae8dd745444159352fa

Observation b51f83f7-e2d6-43a8-b78d-3ca76bfc2530 · outbound

This paper cites To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.015441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.015441Z digest=sha256:bc1f18d9e7d0d307d1e11da83ffd6cf6b7955d705d065578038d5054f04d4832

Observation 50068a82-55d8-46f2-befa-fd67dd965d89 · outbound

This paper cites Boost vision trans- former with gpu-friendly sparsity and quantization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boost vision trans- former with gpu-friendly sparsity and quantization,

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.019933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.019933Z digest=sha256:56363a1f423f9b2ce57800a49290ac54dfe62732e97c42da2ef7c267e8a124cf

Observation 829603f1-bf1b-4491-b10e-32f38c1bceec · outbound

This paper cites Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.024488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.024488Z digest=sha256:4b24d2508a20a69f81749fa2172d52240fb087975ed6afa4cd8603d0855c3750

Observation f313bd3d-1844-4a9a-bd85-3d2672d52b36 · outbound

This paper cites Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.029301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.029301Z digest=sha256:1867c50803597e7cbd8f6b01ffdd5da6e66a3deb9a962d7ad53f52b33f432674

Observation 04a60b94-805b-4d54-98f9-74f3c0d9738b · outbound

This paper cites Vision trans- formers with mixed-resolution tokenization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision trans- formers with mixed-resolution tokenization,

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.034092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.034092Z digest=sha256:1cb450a7ff34c4583f5189cbfe2fa7cf0f2c50577f9f4cd3344e7c2145f5b0de

Observation bf297aea-97c5-4447-9571-90ed5b5de311 · outbound

This paper cites D3former: Debiased dual distilled transformer for incremental learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI D3former: Debiased dual distilled transformer for incremental learning,

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.039180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.039180Z digest=sha256:6a5c25078d80ec41fcc845dd23e2db94a206e3312c8875c60c6f4c77f4092d95

Observation 6c8475f5-969f-4ffa-9351-3c69d53c8376 · outbound

This paper cites Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.043469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.043469Z digest=sha256:347aa96b67a85aefbb66533e360a1cb64439f2c473352f50ef0c6408ce5ebf01

Observation 55df5f88-0300-4922-b572-9723c58761fc · outbound

This paper cites Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.047846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.047846Z digest=sha256:d2c363f310c36fb56380287ecc92c804d9c2ae7e9ecd962fc0caf0d73efd607d

Observation ceba726d-42b0-4a9d-8219-f66a38dffa55 · outbound

This paper cites Masked autoencoding does not help natural language supervision at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Masked autoencoding does not help natural language supervision at scale,

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.052256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.052256Z digest=sha256:4e0d1e7852cbe5fafc36621d2b001511da93ce18287a4f279643cfcb644fa234

Observation c2bc21db-f7a6-4764-8f71-0d34ef5184d3 · outbound

This paper cites Vilem: Visual- language error modeling for image-text retrieval,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vilem: Visual- language error modeling for image-text retrieval,

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.057254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.057254Z digest=sha256:473fb76bafeb6ad07ecda6828f4c315b77d781e5a4f44e7c09af9bbbf41f5a6b

Observation 2da5988c-6dde-40e8-9825-d48802689472 · outbound

This paper cites Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.061757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.061757Z digest=sha256:2e9090a1406d4fbc25f62c47235dca0d5cb70eaff5d259b3bd4a84c2f2c93bcd

Observation a62b514b-987f-4434-bd4f-ac0f10068217 · outbound

This paper cites Zero-shot referring image segmentation with global-local context features,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Zero-shot referring image segmentation with global-local context features,

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.068004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.068004Z digest=sha256:b1bf9fa3491875a5cf9171fc26f945ad9101dd6d0009f9ee78fab78065f75931

Observation c32fa6b7-ffff-4f1e-8ccc-54d1715fb53a · outbound

This paper cites Improving visual grounding by encouraging consistent gradient-based explanations,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving visual grounding by encouraging consistent gradient-based explanations,

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.072944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.072944Z digest=sha256:97d8165cc3cf06cf7adc4ad43dd50e616efd404caeb00b9647c9f67739773741

Observation 7b42af4c-c1e8-401f-b212-beadef5ce0a3 · outbound

This paper cites Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.077543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.077543Z digest=sha256:708be6dcc29dcdf3e43b1a861a15433f433446aae688787cb5e0f1c65281553e

Observation e73f4ac9-9245-4cf8-8ac6-abe073ded2e0 · outbound

This paper cites Multi-modal representation learn- ing with text-driven soft masks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal representation learn- ing with text-driven soft masks,

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.082156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.082156Z digest=sha256:2eabfabfe2d9127b56955c114d8992caa2271edb414fc4689d76a7dcd13f8a25

Observation 095db75f-298b-4a0f-98ad-8b46602f274c · outbound

This paper cites From images to textual prompts: Zero-shot visual question answering with frozen large language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI From images to textual prompts: Zero-shot visual question answering with frozen large language models,

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.086478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.086478Z digest=sha256:c5487dc3b628245522527e46fb9a8c1600b752c2b12c9f9a70a973c113233213

Observation 5c5c3970-f154-4a40-aaf0-cbb4ac470c61 · outbound

This paper cites Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.091386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.091386Z digest=sha256:215cd5214cebae06010f439dd7b3e4210ccd0694e963893cd794a90bd7953abf

Observation ee4af320-2763-4d38-ae43-a9043e532d59 · outbound

This paper cites Semantic information in contrastive learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semantic information in contrastive learning,

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.095850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.095850Z digest=sha256:063d954f120124536e557a25671ae78e68bad771afdeaa3b31b59ed14c3d60cd

Observation fe2cf79a-731b-4777-9221-70e8dabcd219 · outbound

This paper cites Smmix: Self-motivated image mixing for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Smmix: Self-motivated image mixing for vision transformers,

Reference 120

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.100616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.100616Z digest=sha256:5dfaaaff196f1b111c697b946c15dd2c89ae28c76dbd69805b5edd7d12f49318

Observation 0257e80a-6b6a-4220-a954-846725895a96 · outbound

This paper cites Cose: A consistency- sensitivity metric for saliency on image classification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Cose: A consistency- sensitivity metric for saliency on image classification,

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.105981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.105981Z digest=sha256:573b5b3d810b218dda2cbe705445a5cb4ad918014cc567d42647c78135e057a8

Pith citing papers

No inbound Pith citation observations are available.