Pith. sign in

Paper Citation Record · LEDGER

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI

As of 9 August 2026, this Paper Citation Record lists 100 of 251 outbound references and 0 inbound Pith citation observations for arXiv:2608.05258.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05258 v1

Coverage vector

measured 100 of 251 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T16:59:11.105981Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 251 outbound references displayed

  • verified exact4
  • verified fuzzy0
  • unresolved96
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d11799d7-cbac-4c50-a9c6-c44c89d41f55 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.556089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.556089Z digest=sha256:5db9283ba4f0b4b4c5b2273b7d8495c2f26cae859da483a417eba076ee66acf4

Observation 6f4bef24-111c-4059-9ebc-51ec3ae2795f · outbound

This paper cites Representation learning and na- ture encoded fusion for heterogeneous sensor networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Representation learning and na- ture encoded fusion for heterogeneous sensor networks,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.561346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.561346Z digest=sha256:22cdaa444664665b357f548fd0c60821be70b2a3f3da5e6ccc70862fb1f26559

Observation 0e2d9a63-413d-4eb4-9378-4d3f3b328dee · outbound

This paper cites Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Congestion aware dynamic user association in heterogeneous cellular network: A stochastic decision approach,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.565742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.565742Z digest=sha256:9ac63ca198c94bcb88a9d6b0eb0ad0ce4f2b60ae452ad8fd2c61772e069e206a

Observation 165e5521-1a09-4255-b96d-1b3629368d1d · outbound

This paper cites Enhanced robustness by symmetry enforcement,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhanced robustness by symmetry enforcement,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.570437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.570437Z digest=sha256:9e04a9a483224b802e88b95147c623523234c572499542fd4a4d8755d770d40d

Observation 0a432fff-c026-428e-b74b-259d3a45497a · outbound

This paper cites Partial interference alignment for heterogeneous cellular networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Partial interference alignment for heterogeneous cellular networks,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.575167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.575167Z digest=sha256:61cb95bf2b7df6fe8e7a8b9cafbb6d6e77406f5dba4378f48f225d3cff17a6dc

Observation d61ebdc0-a3e0-4c47-9539-b20da33af6c3 · outbound

This paper cites Optimization for user centric massive mimo cell free networks via large system analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Optimization for user centric massive mimo cell free networks via large system analysis,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.580043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.580043Z digest=sha256:dd0fa30a8928d131a9900aaa32746e8e4bcba6e17a6ffc158a0fc7106dd6e097

Observation 2ee2b689-e79e-4419-9d9c-87ebb4c5be66 · outbound

This paper cites Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Exploration vs exploitation for distributed channel access in cognitive radio networks: A multi-user case study,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.589472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.589472Z digest=sha256:1ec4e89be5b3e96e96559fbc9ee74e9953fb272390075ea5418e90947afbc57e

Observation 9fa2d7cb-02eb-49cc-9232-b08d59763459 · outbound

This paper cites Deep reinforcement learning based computation offloading for mobility-aware edge computing,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Deep reinforcement learning based computation offloading for mobility-aware edge computing,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.593676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.593676Z digest=sha256:eb17ec111c9a0d35eb86db9669bc0f548592c7e8e357d689eccbd42b601bc2bf

Observation e4dffcba-d670-4690-a55a-66053719d781 · outbound

This paper cites Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Performance analysis of co- operative multicell precoding with global csi and local individual csi in the large dimensional regime,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.598216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.598216Z digest=sha256:3b74dc4c61ba09f99d6b3f17c40aca521d38e64431471229decc86c5ee342df9

Observation 436732dc-d1bc-4b90-83ff-e0eb42420b21 · outbound

This paper cites Low complexity optimization for user centric cel- lular networks via large dimensional analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Low complexity optimization for user centric cel- lular networks via large dimensional analysis,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.602562Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.602562Z digest=sha256:162d2e6936927ea875cfecb6e4397d874d1ec17bdce8b53003bb53f30d11ee8b

Observation 080326ea-9f1c-40ed-a3a6-f66ce1a212ab · outbound

This paper cites Improving robustness of deep neural networks via large-difference transformation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving robustness of deep neural networks via large-difference transformation,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.607014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.607014Z digest=sha256:8b3746b458a99d0ca57e3fc279bfe68e074eee481df351b050069e37802ca0e7

Observation c0a28969-e207-40c6-b238-b9141d7a21a9 · outbound

This paper cites Looking beyond content: Modeling and detection of fake news from a social context perspective.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Looking beyond content: Modeling and detection of fake news from a social context perspective

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.611370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.611370Z digest=sha256:db2941c893427e99b61e3b92c20681c36baddf834a6f315bb0a75db2aa3be65d

Observation b70425f5-9a7f-40fa-a5f3-fc1c4fc27c6d · outbound

This paper cites Large system analysis for densification of cellular networks with massive mimo,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large system analysis for densification of cellular networks with massive mimo,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.615627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.615627Z digest=sha256:0e9cec21d69202db67b6b7ba11a4d27446d127a7497b24ab687ace9c07ca20a5

Observation a73bec4b-fc33-4bd0-ae8e-aa38cba8ae2e · outbound

This paper cites Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Collabora- tive spectrum sharing based on information pooling for cognitive radio networks with channel heterogeneity,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.620148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.620148Z digest=sha256:ac056fb4c18eb1f888969c15c2a3dcf6039b5ab43f3d6d554840f652a97002d8

Observation 4187ea7e-2c05-4ca1-aa10-17fbc28ccd8d · outbound

This paper cites Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Dense Cross-Connected Ensemble Convolutional Neural Networks for Enhanced Model Robustness

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.624519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.624519Z digest=sha256:5c7612665006ec5876c22a796ecef9b022d2d0ab0854855952c7f1aafb8e8972

Observation 8a2fd787-2e48-4a4b-9cd7-e89decb85427 · outbound

This paper cites Information theory and represen- tation learning inspired multimodal data fusion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Information theory and represen- tation learning inspired multimodal data fusion,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.629428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.629428Z digest=sha256:df5ddfc9307bd95182aef538d87cb5624985120c335117d859b581a81818084c

Observation 560e8a2e-44dc-4470-a3d9-6adbcdc5c858 · outbound

This paper cites Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing Adversarial Robustness of Deep Neural Networks Through Supervised Contrastive Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.286699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.638625Z digest=sha256:4724cff0c697737bfae74e2a36b86d5902631a16ef4f2c1c97f171db902fefe6

Observation 5bf5c5cb-b07f-465f-b50e-d916ce2e0b23 · outbound

This paper cites Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-scale unrectified push-pull with channel attention for enhanced corruption robustness,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.647891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.647891Z digest=sha256:c45bc5fc486d39f2ba56382d2d5e0887b67fbebbdf72f675fee3f5b2122fcb3b

Observation fa094813-9e56-4e25-adf3-58b74e9fbf87 · outbound

This paper cites Expert-guided ex- plainable few-shot learning for medical image diagnosis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning for medical image diagnosis,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.652271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.652271Z digest=sha256:b9c3814c897439dc42b07ad5fbb7003cdb53420757453103033d0b9080d92f11

Observation ebb2d243-fe70-4e43-9158-7b3e6fd6f25b · outbound

This paper cites GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI GetNetUPAM: Ecologically Informed Nested Cross-Validation and Noise-Robust Attention for Marine Bioacoustic Monitoring

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.656692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.656692Z digest=sha256:d64462b0a79351d2a04fc2495090af1b1fe402765054b4b4582c61224d4802f1

Observation 0a9a379b-e784-4fb8-8ece-6d18600a778f · outbound

This paper cites Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shape-aware thoracic edge map chest x- ray representation for pulmonary abnormality screening,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.661437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.661437Z digest=sha256:db2176cfe11be1190d31d2b689ac5997a108c70770e73cd06be4facb4410bf5f

Observation 2c2acf1c-776b-434b-ae53-208c6063aa2d · outbound

This paper cites Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Expert-guided ex- plainable few-shot learning with active sample selection for medical image analysis,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.665772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.665772Z digest=sha256:afda5edf6c392cee9c3024656114a6ab3a575e1e834d60b40d35558b7accfb11

Observation 17e6e87b-f390-4c29-88ef-d4a356b0cd8e · outbound

This paper cites CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI CoSwin: Convolution Enhanced Hierarchical Shifted Window Attention For Small-Scale Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.670721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.670721Z digest=sha256:d3f645eddd08a1a64ef6a366b8342370a5294b87caf858ec21960d940e21ef9f

Observation 308bf2a7-06f6-41e5-b35c-9e8d88885f42 · outbound

This paper cites Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Channel-selected stratified nested cross- validation for clinically relevant eeg-based parkinson’s disease detection,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.679820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.679820Z digest=sha256:7ca2720676430d5e1727e4f122020482008689dd1f3258068d6cbf23f721d577

Observation 5a839a49-aa0d-45fb-a851-9a253daada16 · outbound

This paper cites Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Promoting shape bias in cnns: Frequency-based and contrastive regularization for corruption robustness,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.688548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.688548Z digest=sha256:896a1ceb54efecd1edc68d9bcda3046d77df9d3c05681573183fc2facf2c5cf6

Observation 68879a51-1b9a-4caf-99e1-89967ab390f8 · outbound

This paper cites Learning to select like humans: Explainable active learning for medical imaging,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning to select like humans: Explainable active learning for medical imaging,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.692920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.692920Z digest=sha256:ba91171ff376d918494ba73bdc6be2af7320fcba70e0923b8f4602ac3a6e0a0e

Observation 5e0d7a4c-b915-49b8-803f-988056107507 · outbound

This paper cites Explainable Novel Category Discovery in Semantic Concept Space.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Explainable Novel Category Discovery in Semantic Concept Space

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.233533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.697321Z digest=sha256:13ba9916a6feb809fc8d7ed7f65e456779b9e081b809b7b4d1773776e4daa9ba

Observation 1887b201-8937-4113-8006-45742510055d · outbound

This paper cites Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.188293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.702144Z digest=sha256:b4247e938f3306f9993ba2fba37bae54286718cd9dba34c6b029dd37425f0b38

Observation 785d0c0a-8fd3-48d4-8273-6b16a02a5c28 · outbound

This paper cites Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Frequency-aware contrastive learning for robust shape- biased convolutional neural networks,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.706882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.706882Z digest=sha256:5e9b2951cb887f79b5fb9cb081451723cce7cc8842b03236bd1b6a2a95a03523

Observation 0eb09aea-aa89-4ab3-912a-63997d1f0ec2 · outbound

This paper cites Large dimensional analysis of cooperative multicell precoding with local individual csi,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Large dimensional analysis of cooperative multicell precoding with local individual csi,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.715737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.715737Z digest=sha256:27eb505a1906f30700975b60c9062bebab3bac11bda323efc1922f03c0126299

Observation 76993563-aada-40e3-8a68-19ca3b42c1ad · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI An image is worth 16x16 words: Transformers for image recognition at scale,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.720161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.720161Z digest=sha256:2f32d539fc62092fa4c4db0713789995bad9e73612935bfd1a048e71da2419d7

Observation c152891f-6aea-4391-9770-f81c10223d9d · outbound

This paper cites Pytorch library for cam methods,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pytorch library for cam methods,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.724493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.724493Z digest=sha256:9b2fe5fd04401864725f2f866c77f77ca6f733de6e97da4909267fc97b63ca8c

Observation abf010bf-e3a3-4d15-b72c-98b4ee4e40eb · outbound

This paper cites ImageNet Large Scale Visual Recognition Challenge,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI ImageNet Large Scale Visual Recognition Challenge,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.729099Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.729099Z digest=sha256:9d685018f822ef1efd190976947250e571b5875fc354862cd55604071cb88bc8

Observation dee03cc4-5b09-42d9-a478-cde19bd6ed33 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.733515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.733515Z digest=sha256:2308f2706e0a06e1a9547f984a190e8a0d4a4f70fb581c2141238d35e5401632

Observation b840d73a-bd35-41ba-b90e-c89e2322738a · outbound

This paper cites BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.746918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.746918Z digest=sha256:60b82402932f237244ae30ebf431948c7639bc63982e768d2e5570ab12280bc8

Observation a6c68623-1328-45ae-ab66-0afe03014e9f · outbound

This paper cites Training data-efficient image trans- formers &; distillation through attention,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Training data-efficient image trans- formers &; distillation through attention,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.751469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.751469Z digest=sha256:409e77276dded0d0c3cdf19ebfb9a079fb2cac684b3aada9590ab639dfcf4eae

Observation 696a8e51-6cb3-42bd-88a1-ce695928b7bc · outbound

This paper cites Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-CAM++: Generalized Gradient- Based Visual Explanations for Deep Convolutional Net- works ,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.755770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.755770Z digest=sha256:3f3f1b761ab01f83d9e07358d54f81feb2a7c3d16e30f0392a7eace2846e6f1a

Observation 6414c085-63b4-4f74-8954-2810626e624f · outbound

This paper cites Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ablation-CAM: Visual Explanations for Deep Convolutional Network via Gradient-free Localization ,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.773034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.773034Z digest=sha256:449b08f7fda8805ef3804dfd3d9a9ece8a32a41f5dc30fa3e3fd31cf9e3304d4

Observation 82ef59c6-4706-47f0-a4d7-7dba82f80efd · outbound

This paper cites Full-gradient representation 18 for neural network visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Full-gradient representation 18 for neural network visualization,

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.777847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.777847Z digest=sha256:cbc70daa5b393a5a90877024f13c068cc0fc87bfd368773c2c9ee88e52671b84

Observation 5fed5e13-e4f3-403a-b455-b126fd1519d5 · outbound

This paper cites Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Axiom-based grad-cam: Towards accurate visualization and explanation of cnns,

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.782410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.782410Z digest=sha256:8bf132468e397973b5220f50f8e69e3f091260b78b93d63f0d220ab1aa416e9c

Observation cd3ccd1c-3dec-4a3a-9b36-2839533aabb0 · outbound

This paper cites Learning Deep Features for Discriminative Lo- calization ,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Learning Deep Features for Discriminative Lo- calization ,

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.786756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.786756Z digest=sha256:5d99178f30e977941e6c3608d1a2ff55a3df7b1c64574f0cd5b45489ab24c16b

Observation e5ef305a-6724-4d60-857d-ce9c68ff429a · outbound

This paper cites Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boosting the transferability of adversarial attack on vision transformer with adaptive token tuning,

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.808587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.808587Z digest=sha256:e0b71c9d59ace9857429fd36c08f9b4eced99e602a48a46ef43f6994e57dd4ff

Observation 16c8297c-2f6b-4232-8c4e-fbe9ed8eba55 · outbound

This paper cites Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Emergent open-vocabulary semantic segmentation from off-the- shelf vision-language models,

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.826360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.826360Z digest=sha256:b722879f6bd93bb1d6aa579ac7028ed21d470f928ba6dad66671584c8f944bb6

Observation 56931160-3718-46e4-a740-1a38aa352710 · outbound

This paper cites Enhancing prompt generation with adaptive refinement for camouflaged object detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Enhancing prompt generation with adaptive refinement for camouflaged object detection,

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.843724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.843724Z digest=sha256:17bdf7443fca50ab2866161e1992fdb630e3cc0745b4a5a460d3da1e4363a2af

Observation be5f9b4b-445c-4075-b715-e6e5f63cc018 · outbound

This paper cites Quantifying attention flow in transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Quantifying attention flow in transformers,

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.848312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.848312Z digest=sha256:404ed4976f537ddddef3e2ff44627339d6d7de1312fd287a86ae7580a6ac3305

Observation 66fc7b3d-3284-4d49-90a0-16d4df5bba58 · outbound

This paper cites Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unlike attention-map adaptations, the method operates on feature activations from the MLP in the final transformer blocks and uses the true-label class score as the gradient target

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.854375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.854375Z digest=sha256:46529621aa41b972527446b1e943f3aade5581899b028a2a34b2410f774d6023

Observation e7d02dc3-01c1-42c1-8479-e6d76c39c4dd · outbound

This paper cites the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI the best way we found to apply GradCAM was to treat the last attention layer’s [CLS] token as the designated feature map,

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.859102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.859102Z digest=sha256:793ebddb4f92012de37d8aef58eca4a6fe16b440f2c8ee6300c49c4ee3cb8490

Observation 11f0496e-6463-4aa2-8fbc-b70039ffd91b · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.863825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.863825Z digest=sha256:118258a93c992715e8303d98a56adf9e21ec918a3c3846cd10bc5794181af995

Observation b3a3ab5f-8b3b-4e34-9e52-235dae4b36c1 · outbound

This paper cites Rather than using an attention map as the attribution representation, the method operates on token-level feature activations.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than using an attention map as the attribution representation, the method operates on token-level feature activations

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.868882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.868882Z digest=sha256:0be989c38d2c92ad194ae3ccfd206bcbe09ec2bd07520c7b7d88eac27c69f4d5

Observation df55f346-798b-46c6-84f3-1a175cd1ff09 · outbound

This paper cites Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Rather than operating on attention maps, it fuses gradients and intermediate ViT features from a selected transformer layer

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.873765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.873765Z digest=sha256:17043439512bfdb46f3d87267cd80a9fdcbcdb59781cb49e42d8369d5ad33887

Observation 03d04be6-ae68-4e9e-96cd-af6d03164695 · outbound

This paper cites an unresolved cited work.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Unresolved cited work

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.879223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.879223Z digest=sha256:b6cffdeceddf3ac96eacdca5e15fddafa7f4c8e16d666310f90d5ac9a9ea44b0

Observation a7d66fd2-ced5-4a14-9c46-570c15028ffe · outbound

This paper cites Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before Fuse: Vision and Language Representation Learning with Momentum Distillation [11]is best charac- terized as an attention-map-based attribution method

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.884275Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.884275Z digest=sha256:47ffa23a443ca78075f2b0bfcc3782f656eb94d6b5839e8220c6fc678d577fef

Observation ea63f919-8c90-4164-b286-bae2091db538 · outbound

This paper cites matching.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI matching

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.889056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.889056Z digest=sha256:716b3874878db1d9b59e0b3e15188f9d321e873ddfe6320779062fa6a2189ffe

Observation 896927b2-4c58-4e89-bab1-255af39476c8 · outbound

This paper cites A dog on a white bed.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A dog on a white bed

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.893576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.893576Z digest=sha256:751bb6ad9957e079b96ccb14e06356ce3cbd352303bbb43d316c28287f3011d5

Observation 89e53fb0-a68d-40e6-9a66-da144f0ab390 · outbound

This paper cites Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Grad-cam: Visual explana- tions from deep networks via gradient-based localiza- tion,

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.897806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.897806Z digest=sha256:3c6a4e0ffa684bd4ec92a652275d6537d137439e81bc200df94292b987e4a4bd

Observation 54a8e770-36e0-4c28-96ff-64fa6d237f88 · outbound

This paper cites BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI BLIP: Bootstrap- ping language-image pre-training for unified vision- language understanding and generation,

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.901874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.901874Z digest=sha256:d03a9c2c23332a76032f84467f96f28783e66e7505277c5ea5a5aae92dfe19fc

Observation d76fc51a-f104-4dbe-b7ad-4d69d5bbb578 · outbound

This paper cites Microsoft coco: Common objects in context,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Microsoft coco: Common objects in context,

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.906013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.906013Z digest=sha256:f91b7492e30dec839e8ebfe8e2ad1f1079c11535a5455b39f9e299b1a4b758f7

Observation bcd781eb-9620-480d-9bd2-343bcd340632 · outbound

This paper cites Transformer in- terpretability beyond attention visualization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transformer in- terpretability beyond attention visualization,

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.910505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.910505Z digest=sha256:e474811583fa8586ad3f0f1b3f4a3181f91361d314b3234d91f2e3ff49829890

Observation 348bd8db-0aa6-4056-a8e9-124e79fe93b7 · outbound

This paper cites Transreid: Transformer-based object re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Transreid: Transformer-based object re-identification,

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.914869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.914869Z digest=sha256:ccd0ea1b6558c1121ee37bf3603141353b23a530699baef8d4dbf75eb9525c34

Observation e30f169d-cf73-4ba5-a336-278ebc76c72e · outbound

This paper cites Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Generic attentiemer- gent on-model explainability for interpreting bi-modal and encoder-decoder transformers,

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.918840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.918840Z digest=sha256:5fe6f4706965c3c862cbd327958608cd1abb47d67bfb348e7a0d9a9c69133671

Observation 572b76d5-12af-4075-b5b9-679acd5c8875 · outbound

This paper cites Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Ia-redˆ2: Interpretability-aware redundancy reduction for vision transformers,

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.923344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.923344Z digest=sha256:675a778f95efc21b686f94a5e1fe9fe3cfcd192fbb1117fdd44ec9c53a866b9c

Observation 529da6a4-e53c-4da3-b876-01c18d88157d · outbound

This paper cites Analogous to evolutionary algorithm: Designing a unified sequence model,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Analogous to evolutionary algorithm: Designing a unified sequence model,

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.927583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.927583Z digest=sha256:c3d76cd736b3aad651434815cd1775892a5e847102a034a645a422996cce6a4d

Observation 3332e09e-68ef-453d-974c-e65a0d2ab1b4 · outbound

This paper cites Passive attention in artificial neural networks predicts human visual selectivity,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Passive attention in artificial neural networks predicts human visual selectivity,

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.932385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.932385Z digest=sha256:1d6ea539020d44e4d90971db0d5c19d5f38fa3459f24acf8e6b6f1dbc11377d5

Observation f616e9ae-43fe-4b85-9ff3-fcc8cee7ec56 · outbound

This paper cites Vitae: Vision transformer advanced by exploring intrinsic inductive bias,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vitae: Vision transformer advanced by exploring intrinsic inductive bias,

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.936914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.936914Z digest=sha256:5f24803d72fde5087bedb9a0c41c9a686b3c298d16822043b374ff78c63ed594

Observation 7a6cbe1b-3acb-4ea5-aedd-678e82902f72 · outbound

This paper cites Align before fuse: Vision and language representation learning with momentum distillation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Align before fuse: Vision and language representation learning with momentum distillation,

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.941698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.941698Z digest=sha256:5ab4accbf6281cc41b7681c6c246e9b102cfa05323199dd4c9b5b77cd845ca5a

Observation 7d534b61-812d-4bd6-8096-e1accd4878e1 · outbound

This paper cites VLMAE: Vision-Language Masked Autoencoder.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI VLMAE: Vision-Language Masked Autoencoder

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-08T16:59:12.128574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T16:59:10.946360Z digest=sha256:08d31e7601bc6b8b046e3b62a3a51e84494c87635cd8aa6a5c3e3f47a0f25142

Observation e90b8372-757a-4350-af5f-4898e7a6a024 · outbound

This paper cites A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A compre- hensive study of image classification model sensitivity to foregrounds, backgrounds, and visual attributes,

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.951037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.951037Z digest=sha256:57b4ef0e55adcbef767d0e528e8291b97a3bcca2dfe4d37f543e7379376e6d09

Observation d0d4f501-0a2b-4007-83e3-1ccef1a65f85 · outbound

This paper cites A challenging benchmark of anime style recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A challenging benchmark of anime style recognition,

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.955330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.955330Z digest=sha256:97595800b3d591c84658b97e8bcfc6db101225977f2e9c572f2902c9497ff524

Observation 7ec2ad21-43de-4333-99d6-ba9ecface9a1 · outbound

This paper cites Metaformer is actually what you need for vision,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Metaformer is actually what you need for vision,

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.960419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.960419Z digest=sha256:ea8771a289b4d46788466738a3d86a6e0cf1c27b4beefd66fdaaa1ed64f2397c

Observation 0869e2ca-474f-46aa-bf2b-667d754c1d32 · outbound

This paper cites Delving deep into the generalization of vision transformers under distribution shifts,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving deep into the generalization of vision transformers under distribution shifts,

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.965280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.965280Z digest=sha256:f37c86142aaa1d240f5ae8135ab261ea045ad948230db593325a9864ca1d5f43

Observation 3992634d-9567-458d-893d-75ca64e133cb · outbound

This paper cites General facial representation learning in a visual- linguistic manner,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI General facial representation learning in a visual- linguistic manner,

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.970138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.970138Z digest=sha256:a5f88fa7e04a35337a31a5527989febbf70ca17b69f2dc5170df82ea2aafe9ed

Observation ce334212-4a11-43ec-bcb4-c49ea2364d55 · outbound

This paper cites Multi-modal alignment using representa- tion codebook,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal alignment using representa- tion codebook,

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.974492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.974492Z digest=sha256:7221e954892ae983007474094ad0bb3a0921762fbd8a049312db97a7d9906ac3

Observation 136ae6d9-9f8c-473a-9a0c-572f10a23b29 · outbound

This paper cites Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Plug-and-play VQA: Zero-shot VQA by conjoining large pretrained models with zero training,

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.979504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.979504Z digest=sha256:81dfb714c1400d2c653ff56b619d6b4d512e109fd1248a67a9e765b8796d1bc0

Observation 16f7963a-715b-4d2f-800a-927664864f8c · outbound

This paper cites Inception transformer,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Inception transformer,

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.983594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.983594Z digest=sha256:4fa3bf319662e1c3863eab41aa93685ef72b66e06b8c405d6256e844ba69ca69

Observation e6bb17c9-8e60-480f-b346-9f91e16eba1c · outbound

This paper cites Delving into sequential patches for deep- fake detection,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Delving into sequential patches for deep- fake detection,

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.988241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.988241Z digest=sha256:37606e708aafcadc15131f94aa83a99c3f37f6574a3dd6139b5545d6e4341a17

Observation 221817f8-c088-4d15-a179-8aac556bfc75 · outbound

This paper cites Adversarial normalization: I can visualize everything (ice),.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Adversarial normalization: I can visualize everything (ice),

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.992974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.992974Z digest=sha256:2e247963d88b44e1c5320f35c3c9257edc63ca8c443ac0f0d650b3159f71a1ea

Observation 435411ee-05f4-4aa7-b21d-b796ff46f1e4 · outbound

This paper cites A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI A new benchmark: On the utility of synthetic data with blender for bare supervised learning and downstream domain adaptation,

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:10.997910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:10.997910Z digest=sha256:ba64f7048bb228559f34976b3324db21b489d9368c986fc0c69556195891240c

Observation 578198fe-54c4-47d8-8a41-bb65e2d10af6 · outbound

This paper cites Selfme: Self-supervised motion learning for micro- expression recognition,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Selfme: Self-supervised motion learning for micro- expression recognition,

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.002177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.002177Z digest=sha256:38c7e4515eb67372690ca390abb2888eb0fff81cf176d952b0d5fc5e1856b24a

Observation 2eb616b1-5ff8-4cde-9aa5-a6fcc5828b3b · outbound

This paper cites Marlin: Masked autoencoder for facial video representation learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Marlin: Masked autoencoder for facial video representation learning,

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.006873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.006873Z digest=sha256:744462d9eb3e3cc03cd06a12ed538958bb0d215017402dbb9b2702226d3f5a9a

Observation 00360c11-ac10-4e8c-a3e0-a892efc24898 · outbound

This paper cites Blackvip: Black-box visual prompting for robust transfer learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Blackvip: Black-box visual prompting for robust transfer learning,

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.011351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.011351Z digest=sha256:6ed70a80cfada05d9575acd3a39301cf6811717f180c3d79adeedb6a69bd6bb5

Observation b51f83f7-e2d6-43a8-b78d-3ca76bfc2530 · outbound

This paper cites To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI To- kenhpe: Learning orientation tokens for efficient head pose estimation via transformers,

Reference 102

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.015441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.015441Z digest=sha256:61920b1c1f00371bd114138d03a53c725c9d24d8a3746a05a2fcc0d3d55ef806

Observation 50068a82-55d8-46f2-befa-fd67dd965d89 · outbound

This paper cites Boost vision trans- former with gpu-friendly sparsity and quantization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Boost vision trans- former with gpu-friendly sparsity and quantization,

Reference 103

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.019933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.019933Z digest=sha256:ae5d2a4fa6ff1703283c30862f547f8f646905845ac757ffd0511f366de61d85

Observation 829603f1-bf1b-4491-b10e-32f38c1bceec · outbound

This paper cites Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision diffmask: Faithful interpretation of vision transformers with differentiable patch masking,

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.024488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.024488Z digest=sha256:1bc095a1fc3efa6fa2e822ab311013bc7346f20663556e34a35f84574e72eb46

Observation f313bd3d-1844-4a9a-bd85-3d2672d52b36 · outbound

This paper cites Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Pha: Patch-wise high-frequency augmentation for transformer-based person re-identification,

Reference 105

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.029301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.029301Z digest=sha256:7ea9a171ac216d7fc297e1704756ca9872b7990a540131e47b4e2e20b4f9bd6e

Observation 04a60b94-805b-4d54-98f9-74f3c0d9738b · outbound

This paper cites Vision trans- formers with mixed-resolution tokenization,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vision trans- formers with mixed-resolution tokenization,

Reference 106

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.034092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.034092Z digest=sha256:8f89d772331b0ca24cd2ca27c5c1c4525705289bbe907a8d1bdb78627347a741

Observation bf297aea-97c5-4447-9571-90ed5b5de311 · outbound

This paper cites D3former: Debiased dual distilled transformer for incremental learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI D3former: Debiased dual distilled transformer for incremental learning,

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.039180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.039180Z digest=sha256:c8eb9386a6cb7da6e3c32ee485ba34107ca6b8dc8be8743f004471ad357dbd7a

Observation 6c8475f5-969f-4ffa-9351-3c69d53c8376 · outbound

This paper cites Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Shared inter- est...sometimes: Understanding the alignment between human perception, vision architectures, and saliency map techniques,

Reference 108

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.043469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.043469Z digest=sha256:85e34517135fd08044428911f58e379d1c0a484a18ba119581e75c3afaa14cf3

Observation 55df5f88-0300-4922-b572-9723c58761fc · outbound

This paper cites Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semicvt: Semi- supervised convolutional vision transformer for seman- tic segmentation,

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.047846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.047846Z digest=sha256:2d75457a7a554126448051d3f69f86474b55b9aa475c8f98688cde4e2235f338

Observation ceba726d-42b0-4a9d-8219-f66a38dffa55 · outbound

This paper cites Masked autoencoding does not help natural language supervision at scale,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Masked autoencoding does not help natural language supervision at scale,

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.052256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.052256Z digest=sha256:e362defb79eec7fce89ade5f76654e8859a355fb2a2df9460ca20cdbbfa63b12

Observation c2bc21db-f7a6-4764-8f71-0d34ef5184d3 · outbound

This paper cites Vilem: Visual- language error modeling for image-text retrieval,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Vilem: Visual- language error modeling for image-text retrieval,

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.057254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.057254Z digest=sha256:52e19b2a3fd6ac7da97b472037fe91da396514e8a9d23fa74d2703947eadcccb

Observation 2da5988c-6dde-40e8-9825-d48802689472 · outbound

This paper cites Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Fashionsap: Symbols and attributes prompt for fine-grained fashion vision-language pre-training,

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.061757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.061757Z digest=sha256:177bbf7fc986ff1a4f0a728d219a169761abba4612ec6d1e276c56d1106e445d

Observation a62b514b-987f-4434-bd4f-ac0f10068217 · outbound

This paper cites Zero-shot referring image segmentation with global-local context features,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Zero-shot referring image segmentation with global-local context features,

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.068004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.068004Z digest=sha256:9a404d49193c4bd26256ad66a016ce1a41c5b11d1d8e0398bd3e3f0be6d3e6a3

Observation c32fa6b7-ffff-4f1e-8ccc-54d1715fb53a · outbound

This paper cites Improving visual grounding by encouraging consistent gradient-based explanations,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Improving visual grounding by encouraging consistent gradient-based explanations,

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.072944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.072944Z digest=sha256:06a8b459423f6c242c89b61288476916fdc6d8941b1c553a8e6fec2adcacbabd

Observation 7b42af4c-c1e8-401f-b212-beadef5ce0a3 · outbound

This paper cites Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Clip is also an efficient segmenter: A text-driven approach for weakly supervised semantic segmentation,

Reference 115

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.077543Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.077543Z digest=sha256:ad0b365477436a5465dfa650ab13c34c75e9f1a2f6b104771ee38a5bf3c63a34

Observation e73f4ac9-9245-4cf8-8ac6-abe073ded2e0 · outbound

This paper cites Multi-modal representation learn- ing with text-driven soft masks,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Multi-modal representation learn- ing with text-driven soft masks,

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.082156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.082156Z digest=sha256:e1cea6d95e620b58663a78a74628a595738f40c8efee0ba4bc9fefd3f90740db

Observation 095db75f-298b-4a0f-98ad-8b46602f274c · outbound

This paper cites From images to textual prompts: Zero-shot visual question answering with frozen large language models,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI From images to textual prompts: Zero-shot visual question answering with frozen large language models,

Reference 117

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.086478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.086478Z digest=sha256:4403a970ff6b8df500b8fb6b56886d548e862da376c639eea7c7303175749743

Observation 5c5c3970-f154-4a40-aaf0-cbb4ac470c61 · outbound

This paper cites Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Sparse multi- modal vision transformer for weakly supervised seman- tic segmentation,

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.091386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.091386Z digest=sha256:5bb28ce38a5d29b7fee97f344f664f3397fa636bd0d54e7f1d4c444da93d9ecb

Observation ee4af320-2763-4d38-ae43-a9043e532d59 · outbound

This paper cites Semantic information in contrastive learning,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Semantic information in contrastive learning,

Reference 119

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.095850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.095850Z digest=sha256:5e8862575ca5ab3459c07ec5c1c4e010eb336048760142016e75899aedadb9b7

Observation fe2cf79a-731b-4777-9221-70e8dabcd219 · outbound

This paper cites Smmix: Self-motivated image mixing for vision transformers,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Smmix: Self-motivated image mixing for vision transformers,

Reference 120

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.100616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.100616Z digest=sha256:9121fcf108c4c6956162da4c10e27fef41c46122b88484304da5e0cf1104cef1

Observation 0257e80a-6b6a-4220-a954-846725895a96 · outbound

This paper cites Cose: A consistency- sensitivity metric for saliency on image classification,.

Grad-CAM for Vision Transformers: A Systematic Taxonomy and Audit of Methodological Ambiguity in Explainable AI Cose: A consistency- sensitivity metric for saliency on image classification,

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-08T16:59:11.105981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T16:59:11.105981Z digest=sha256:957f5a9b442fbc4d8b3cba9911c5b8d01105fbd434ab9e9597bc0202e639620f

Pith citing papers

No inbound Pith citation observations are available.