REVIEW 1 cited by
AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification
read the original abstract
Multiple Instance Learning (MIL) effectively analyzes whole slide images but faces overfitting due to attention over-concentration. While existing solutions rely on complex architectural modifications or additional processing steps, we introduce Attention Entropy Maximization (AEM), a simple yet effective regularization technique. Our investigation reveals the positive correlation between attention entropy and model performance. Building on this insight, we integrate AEM regularization into the MIL framework to penalize excessive attention concentration. To address sensitivity to the AEM weight parameter, we implement Cosine Weight Annealing, reducing parameter dependency. Extensive evaluations demonstrate AEM's superior performance across diverse feature extractors, MIL frameworks, attention mechanisms, and augmentation techniques. Here is our anonymous code: https://github.com/dazhangyu123/AEM.
Forward citations
Cited by 1 Pith paper
-
Cluster-Level Sparse Multi-Instance Learning for Whole-Slide Images
csMIL adds K-means cluster sparsity to attention-based MIL, reporting CAMELYON16 AUC 0.951 and TCGA-NSCLC AUC 0.933, but with test-set-tuned hyperparameters and a borrowed Lasso bound.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.