Pith. sign in

REVIEW 24 cited by

Quantifying Attention Flow in Transformers

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2005.00928 v2 pith:K5W5IRX2 submitted 2020-05-02 cs.LG cs.AIcs.CL

classification cs.LGcs.AIcs.CL
keywords attentionflowinformationinputtokensmethodsweightsquantifying
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In the Transformer model, "self-attention" combines information from attended embeddings into the representation of the focal embedding in the next layer. Thus, across layers of the Transformer, information originating from different tokens gets increasingly mixed. This makes attention weights unreliable as explanations probes. In this paper, we consider the problem of quantifying this flow of information through self-attention. We propose two methods for approximating the attention to input tokens given attention weights, attention rollout and attention flow, as post hoc methods when we use attention weights as the relative relevance of the input tokens. We show that these methods give complementary views on the flow of information, and compared to raw attention, both yield higher correlations with importance scores of input tokens obtained using an ablation method and input gradients.

Discussion (0). Sign in to comment.

Forward citations

Cited by 24 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Transient Reserves, Sink Dampers, and the Failure of Eigenvalue Reasoning in the Attention Propagator

    cond-mat.dis-nn 2026-07 conditional novelty 7.0 of 10

    Resolvent analysis of trained causal attention shows sinks act as transient dampers, routing heads carry excess Kreiss reserve, and eigenvalue depth predictions fail by 7–11 orders of magnitude.

  2. AGNFormer I: Reconstruction of AGN spectra using a probabilistic transformer model

    astro-ph.GA 2026-07 conditional novelty 6.0 of 10

    An uncertainty-aware transformer reconstructs masked AGN broad lines and spectral halves with 4-16% flux errors and beats eleven purpose-built Lyα-reconstruction algorithms on a blind benchmark.

  3. Feature-level Interaction Explanations in Multimodal Transformers

    cs.LG 2026-03 conditional novelty 6.0 of 10

    FL-I2MoE separates unique, synergistic, and redundant cross-modal evidence at the token/patch level and uses SII and redundancy-gap scores to rank pairs whose removal degrades performance more than random masking.

  4. Full-Frequency Temporal Patching and Structured Masking for Enhanced Audio Classification

    cs.SD 2025-08 conditional novelty 6.0 of 10

    Replacing square spectrogram patches with full-frequency temporal patches plus patch-aligned masking improves audio classification accuracy and reduces compute for Transformer and Mamba models.

  5. DSS-Prompt: Dynamic-Static Synergistic Prompting for Few-Shot Class-Incremental Learning

    cs.CV 2025-08 conditional novelty 6.0 of 10

    DSS-Prompt combines static prompts with instance-aware dynamic prompts generated from BLIP multi-modal features to achieve state-of-the-art few-shot class-incremental learning on four benchmarks without incremental training.

  6. Towards White-Box Deep Wireless Sensing

    cs.LG 2025-07 conditional novelty 6.0 of 10

    RF-CRATE derives a fully complex-valued white-box transformer for RF sensing from the sparse rate reduction principle and shows it matches black-box baselines across five datasets.

  7. Self-Guided Masked Autoencoder

    cs.CV 2025-07 conditional novelty 6.0 of 10

    A Masked Autoencoder that masks the object cluster found by its own early patch-clustering signal learns better representations than random masking, with no external labels or models.

  8. CytoSAE: Interpretable Cell Embeddings for Hematology

    cs.CV 2025-07 conditional novelty 6.0 of 10

    CytoSAE learns sparse, expert-validated morphological concepts from blood-cell images that generalize across datasets and can classify AML subtypes at patient level with F1 0.83.

  9. VIP: Visual Information Protection through Adversarial Attacks on Vision-Language Models

    eess.IV 2025-07 conditional novelty 6.0 of 10

    A perturbation computed from early attention and value matrices can make LLaVA, Instruct-BLIP, and BLIP2-T5 fail to detect objects inside a specified image region while keeping the rest of the image usable.

  10. Foveation-Guided Dynamic Token Selection for Robust and Efficient Vision Transformers

    cs.CV 2026-07 conditional novelty 5.0 of 10

    FDT adds foveation and binary fixation modules to DeiT so multi-scale tokens are selected dynamically in one pass, improving ImageNet100 accuracy, MACs, and robustness without adversarial training.

  11. From Features to Actions: Explainability in Traditional and Agentic AI Systems

    cs.AI 2026-02 conditional novelty 5.0 of 10

    Attribution explanations that work for static classifiers do not diagnose failures in multi-step AI agents; trace-grounded rubric evaluation does, with state-tracking inconsistency 2.7x more common in failed agent runs.

  12. Robust Representation Learning in Masked Autoencoders

    cs.LG 2026-02 conditional novelty 5.0 of 10

    Masked Autoencoders build class-separable representations across depth and keep their embeddings directionally stable under blur and occlusion, which tracks their robust classification.

  13. Revisiting 2D Foundation Models for Scalable 3D Medical Image Classification

    cs.CV 2025-12 conditional novelty 5.0 of 10

    A frozen 2D vision foundation model with lightweight LoRA adapters and attention-based slice fusion achieves state-of-the-art 3D medical image classification across 12 tasks with about 1M trainable parameters per task.

  14. An Autoencoder and Vision Transformer-based Interpretability Analysis of the Differences in Automated Staging of Second and Third Molars

    cs.CV 2025-09 conditional novelty 5.0 of 10

    An autoencoder-plus-ViT pipeline improves dental staging accuracy and uses latent space, reconstructions, and attention maps to attribute the weaker third-molar performance to high intra-class data variability.

  15. Attention Maps in 3D Shape Classification for Dental Stage Estimation with Class Node Graph Attention Networks

    cs.CV 2025-09 conditional novelty 5.0 of 10

    CGAT, a graph attention network with a CLS node, achieves 0.76 weighted F1 on Demirjian stage classification of 3D third-molar meshes and generates attention maps that highlight roots and furcation regions.

  16. Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts

    cs.AI 2025-08 conditional novelty 5.0 of 10

    A method and case study for visualizing cross-attention maps in Wan video diffusion transformers, showing token-region alignment over time and their use as artistic material.

  17. Decoding the Multimodal Maze: A Systematic Review on the Adoption of Explainability in Multimodal Attention-based Models

    cs.LG 2025-08 unverdicted novelty 5.0 of 10

    A systematic review of 55 papers finds explainability for multimodal attention-based models is dominated by attention-weight visualizations, while evaluation remains mostly qualitative and non-standardized.

  18. User Experience Estimation in Human-Robot Interaction Via Multi-Instance Learning of Multimodal Social Signals

    cs.RO 2025-07 reject novelty 5.0 of 10

    A multimodal Transformer with multi-instance learning estimates user experience (UX) questionnaire ratings from facial expressions and voice during human-robot interaction, reporting accuracy above third-party human raters.

  19. Fair-FLIP: Fair Deepfake Detection with Fairness-Oriented Final Layer Input Prioritising

    cs.LG 2025-07 conditional novelty 5.0 of 10

    Fair-FLIP improves fairness parity in deepfake detection by reweighting final-layer features based on between-ethnicity variance, with negligible accuracy loss.

  20. Safer Skin Lesion Classification with Global Class Activation Probability Map Evaluation and SafeML

    cs.CV 2025-08 conditional novelty 4.0 of 10

    A pixel-level argmax over per-class Grad-CAM maps, combined with a selective predictor, is proposed to detect unreliable skin lesion classifications.

  21. Generalizable Federated Learning using Client Adaptive Focal Modulation

    cs.CV 2025-08 reject novelty 4.0 of 10

    The abstract describes AdaptFED, a claimed federated learning method, but the full text is an unrelated graph theory paper, so the claimed results are absent from the submission.

  22. PiPViT: Patch-based Visual Interpretable Prototypes for Retinal Image Analysis

    cs.CV 2025-06 conditional novelty 4.0 of 10

    PiPViT combines vision transformers and prototype learning to classify retinal OCT scans while showing the spatial extent of the biomarker that drove the decision.

  23. Learning from Limited and Imperfect Data

    cs.LG 2025-07 unverdicted novelty 3.0 of 10

    A doctoral thesis compiling nine peer-reviewed papers on long-tailed image generation, long-tailed recognition, semi-supervised learning, and domain adaptation.

  24. The Hitchhiker's Guide to Agentic AI: From Foundations to Systems

    cs.AI 2026-06 unverdicted novelty 2.0 of 10

    A survey-style reference book mapping the full agentic-AI stack from transformer internals to production deployment, with no new research result.

Pith tools