REVIEW 2 cited by
Augmenting Convolutional networks with attention-based aggregation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We show how to augment any convolutional network with an attention-based global map to achieve non-local reasoning. We replace the final average pooling by an attention-based aggregation layer akin to a single transformer block, that weights how the patches are involved in the classification decision. We plug this learned aggregation layer with a simplistic patch-based convolutional network parametrized by 2 parameters (width and depth). In contrast with a pyramidal design, this architecture family maintains the input patch resolution across all the layers. It yields surprisingly competitive trade-offs between accuracy and complexity, in particular in terms of memory consumption, as shown by our experiments on various computer vision tasks: object classification, image segmentation and detection.
Forward citations
Cited by 2 Pith papers
-
FECT: Classification of Breast Cancer Pathological Images Based on Fusion Features
FECT fuses cell, tissue, and edge features, achieving 65.8% weighted F1 on the seven-class BRACS breast cancer classification benchmark.
-
Scaling Structure Aware Virtual Screening to Billions of Molecules with SPRINT
SPRINT co-embeds drugs and proteins with a structure-aware language model and attention pooling, achieving leading virtual screening enrichment and billion-scale retrieval speed.
Discussion (0). Continue with ORCID to comment.